Transcription of Introduction to Text Mining - VSCSE
{{id}} {{{paragraph}}}
Abbott Analytics, Inc. 2001-2013 Introduction to Text MiningVirtual Data Intensive Summer SchoolJuly 10, 2013 Dean AbbottAbbott Analytics, : : : @deanabb1 Wednesday, July 10, 13 Abbott Analytics, Inc. 2001-2013 Why Text? How much data? zettabytes ( trillion GB) Most of the World s Data is Unstructured 2009 HP survey: 70% Gartner: 80% Jerry Hill (Teradata), Anant Jhingran (IBM): 85% Structured (stored) data often misses elements critical to predictive modeling Un-transcribed fields, notes, comments Ex: examiner/adjuster notes, surveys with free-text fields, medical charts2 Wednesday, July 10, 13 Abbott Analytics, Inc.
mining classification methods, based on models trained on labeled examples. 4. Web Mining: Data and Text Mining on the Internet with a specific focus on the scale and interconnectedness of the web. 9 Wednesday, July 10, 13
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}
Text Mining: A Thematic Exploration of Don Quixote, Text Mining, And analysis, Analysis, Mining, Text, Text Mining and Analysis, Text Mining for Health Care and Medicine, Text analysis, Analysis of Voice of Customer: Text Mining, Text mining and topic models, Text mining for central banks, Text Mining Process, Techniques and Tools