Practical Text Mining and Statistical Analysis for Non-structured Text Data Applications

Practical Text Mining and Statistical Analysis for Non-structured Text Data Applications

3.35 (17 ratings by Goodreads)
By (author)  , By (author)  , By (author)  , By (author)  , By (author)  , By (author) 

List price: US$79.99

Currently unavailable

We can notify you when this item is back in stock

Add to wishlist

AbeBooks may have this title (opens in new window).

Try AbeBooks


The world contains an unimaginably vast amount of digital information which is getting ever vaster ever more rapidly. This makes it possible to do many things that previously could not be done: spot business trends, prevent diseases, combat crime and so on. Managed well, the textual data can be used to unlock new sources of economic value, provide fresh insights into science and hold governments to account. As the Internet expands and our natural capacity to process the unstructured text that it contains diminishes, the value of text mining for information retrieval and search will increase dramatically.

This comprehensive professional reference brings together all the information, tools and methods a professional will need to efficiently use text mining applications and statistical analysis.

The Handbook of Practical Text Mining and Statistical Analysis for Non-structured Text Data Applications presents a comprehensive how- to reference that shows the user how to conduct text mining and statistically analyze results. In addition to providing an in-depth examination of core text mining and link detection tools, methods and operations, the book examines advanced preprocessing techniques, knowledge representation considerations, and visualization approaches. Finally, the book explores current real-world, mission-critical applications of text mining and link detection using real world example tutorials in such varied fields as corporate, finance, business intelligence, genomics research, and counterterrorism activities.

-Extensive case studies, most in a tutorial format, allow the reader to 'click through' the example using a software program, thus learning to conduct text mining analyses in the most rapid manner of learning possible

-Numerous examples, tutorials, power points and datasets available via companion website on

-Glossary of text mining terms provided in the appendix

show more

Product details

  • Electronic book text | 1000 pages
  • 191 x 235mm
  • Academic Press Inc
  • United Kingdom
  • English
  • Approx. 200 illustrations
  • 9780123870117

Review quote

"They've done it again. From the same industry leaders who brought you the "bible" of data mining comes the definitive, go-to text mining resource. This book empowers you to dig in and seize value, with over two dozen hands-on tutorials that drive an incredible range of applications such as predicting marketing success and detecting customer sentiment, criminal lies, writing authorship, and patient schizophrenia. These step-by-step tutorials immediately place you in the practitioner's driver's seat, executing on text analytics. Beyond this, 17 more chapters cover the latest methods and the leading tools, making this the most comprehensive resource, and earning it a well-deserved place on your desk aside the authors' data mining handbook." - Eric Siegel, Ph.D., Founder, Predictive Analytics World, Text Analytics World and Prediction Impact, Inc.

"Of the number of statistics books that are published each year, only a few stand out as really being important, meaning that they positively influence how future research is done in the subject area of the text. I believe that Practical Text Mining is just such a book.” - Joseph M. Hilbe, JD, PhD, Arizona State University and Jet Propulsion Laboratory

"When you want real help extracting insight from the mountains of text that you're facing, this is the book to turn to for immediate practical advice.” - Karl Rexer, PhD, President, Rexer Analytics, Boston, MA

"The underlying premise is that almost all data in databases takes the form of unstructured text, or summaries of unstructured text, and that historians, marketers, crime investigators, and others need to know how to search that text for meaningful patterns - a very different process than reading. Contributors in a range of fields share their insights and experience with the process. After setting out the principles, they present tutorials and case studies, then move on to advanced topics." - Reference and Research Book News, Inc. "The authors of Practical Text Mining and Statistical Analysis for Nonstructured Text Data Applications have managed to produce three books in one. First, in 17 chapters they give a friendly yet comprehensive introduction to the huge field of text mining, a field comprising techniques from several different disciplines and a variety of different tasks. Miner and his colleagues have produced a readable overview of the area that is sure to help the practitioner navigate this large and unruly ocean of techniques. Second, the authors provide a comprehensive list and review of both the commercial and free software available to perform most text data mining tasks. Finally, and most importantly, the authors have also provided an amazing collection of tutorials and case studies. The tutorials illustrate various text mining scenarios and paths actually taken by researchers, while the case studies go into even more depth, showing both the methodology used and the business decisions taken based on the analysis. These practical step-by-step guides are impressive not only in the breadth of their applications but in the depth and detail that each case study delivers. The studies are authored by several guest authors in addition to the book authors and are built on real problems with real solutions. These case studies and tutorials alone make the book worth having. I have never seen such a collection of real business problems published in any field, much less in such a new field as text mining. These, together with the explanations in the chapters, should provide the practitioner wishing to get a broad view of the text mining field an invaluable resource for both learning and practice. - Richard De Veaux Professor of Statistics; Dept. of Mathematics and Statistics; Williams College; Williamstown MA 01267 "In writing Practical Text Mining and Statistical Analysis for Nonstructured Text Data Applications, the six authors (Miner, Delen, Elder, Fast, Hill, and Nisbet) accepted the daunting task of creating a cohesive operational framework from the disparate aspects and activities of text mining, an emerging field that they appropriately describe as the "Wild West" of data mining. Tapping into their unique expertise and applying a wide cross-application lens, they have succeeded in their mission. Rather than listing the facets of text mining simply as independent academic topics of discussion, the book leans much more to the practical, presenting a conceptual road map to assist users in correlating articulated text mining techniques to categories of actual commonly observed business needs. To finish out the job, summaries for some of the most prevalent commercial text mining solutions are included, along with examples. In this way, the authors have uniquely presented a text mining resource with value to readers across that breadth of business applications."  - Gerard Britton, J.D. V.P., GRC Analytics, Opera Solutions LLC "Text Mining is one of those phrases people throw around as though it describes something singular. As the authors of Practical Text Mining and Statistical Analysis for Non-structured Text Data Applications show us, nothing could be further from the truth. There is a rich, diverse ecosystem of text mining approaches and technologies available. Readers of this book will discover a myriad of ways to use these text mining approaches to understand and improve their business. Because the authors are a practical bunch the book is full of examples and tutorials that use every approach, multiple commercial and open source tools, and that show the power and trade-offs each involves. The case studies are worked through in detail by the authors so you can see exactly how things would be done and learn how to apply it to your own problems. If you are interested in text mining, and you should be, this book will give you a perspective that is broad, deep and approachable." - James Taylor CEO Decision Management Solutions
show more

Table of contents

Part I Basic Text Mining Principles 1. The History of Text Mining 2. The Seven Practice Areas of Text Analytics 3. Conceptual Foundations of Text Mining and Preprocessing Steps 4. Applications and Use Cases for Text Mining 5. Text Mining Methodology 6. Three Common Text Mining Software Tools

Part II Introduction to the Tutorial and Case Study Section of This Book AA. CASE STUDY: Using the Social Share of Voice to Predict Events That Are about to Happen BB. Mining Twitter for Airline Consumer Sentiment A. Using STATISTICA Text Miner to Monitor and Predict Success of Marketing Campaigns Based on Social Media Data B. Text Mining Improves Model Performance in Predicting Airplane Flight Accident Outcome C. Insurance Industry: Text Analytics Adds "Lift” to Predictive Models with STATISTICA Text and Data Miner D. Analysis of Survey Data for Establishing the "Best Medical Survey Instrument” Using Text Mining E. Analysis of Survey Data for Establishing "Best Medical Survey Instrument” Using Text Mining: Central Asian (Russian Language) Study Tutorial 2: Potential for Constructing Instruments That Have Increased Validity F. Using eBay Text for Predicting ATLAS Instrumental Learning G. Text Mining for Patterns in Children's Sleep Disorders Using STATISTICA Text Miner H. Extracting Knowledge from Published Literature Using RapidMiner I. Text Mining Speech Samples: Can the Speech of Individuals Diagnosed with Schizophrenia Differentiate Them from Unaffected Controls? J. Text Mining Using STM, CART, and TreeNet from Salford Systems: Analysis of 16,000 iPod Auctions on eBay K. Predicting Micro Lending Loan Defaults Using SAS Text Miner L. Opera Lyrics: Text Analytics Compared by the Composer and the Century of CompositiondWagner versus Puccini M. CASE STUDY: Sentiment-Based Text Analytics to Better Predict Customer Satisfaction and Net Promoter Score Using IBM SPSS Modeler N. CASE STUDY: Detecting Deception in Text with Freely Available Text and Data Mining Tools O. Predicting Box Office Success of Motion Pictures with Text Mining P. A Hands-On Tutorial of Text Mining in PASW: Clustering and Sentiment Analysis Using Tweets from Twitter Q. A Hands-On Tutorial on Text Mining in SAS: Analysis of Customer Comments for Clustering and Predictive Modeling R. Scoring Retention and Success of Incoming College Freshmen Using Text Analytics S. Searching for Relationships in Product Recall Data from the Consumer Product Safety Commission with STATISTICA Text Miner T. Potential Problems That Can Arise in Text Mining: Example Using NALL Aviation Data U. Exploring the Unabomber Manifesto Using Text Miner V. Text Mining PubMed: Extracting Publications on Genes and Genetic Markers Associated with Migraine Headaches from PubMed Abstracts W. CASE STUDY: The Problem with the Use of Medical Abbreviations by Physicians and Health Care Providers X. Classifying Documents with Respect to "Earnings” and Then Making a Predictive Model for the Target Variable Using Decision Trees, MARSplines, Naïve Bayes Classifier, and K-Nearest Neighbors with STATISTICA Text Miner Y. CASE STUDY: Predicting Exposure of Social Messages: The Bin Laden Live Tweeter Z. The InFLUence Model: Web Crawling, Text Mining, and Predictive Analysis with 2010e2011 Influenza GuidelinesdCDC, IDSA, WHO, and FMC

Part III Advanced Topics 7. Text Classification and Categorization 8. Prediction in Text Mining: The Data Mining Algorithms of Predictive Analytics 9. Entity Extraction 10. Feature Selection and Dimensionality Reduction 11. Singular Value Decomposition in Text Mining 12. Web Analytics and Web Mining 13. Clustering Words and Documents 14. Leveraging Text Mining in Property and Casualty Insurance 15. Focused Web Crawling 16. The Future of Text and Web Analytics 17. Summary



show more

Rating details

17 ratings
3.35 out of 5 stars
5 6% (1)
4 47% (8)
3 35% (6)
2 0% (0)
1 12% (2)
Book ratings by Goodreads
Goodreads is the world's largest site for readers with over 50 million reviews. We're featuring millions of their reader ratings on our book pages to help you find your new favourite book. Close X