Text as Data (eBook)

Computational Methods of Understanding Written Expression Using SAS
eBook Download: EPUB
2021 | 1. Auflage
240 Seiten
John Wiley & Sons (Verlag)
978-1-119-48715-9 (ISBN)

Lese- und Medienproben

Text as Data - Barry DeVille, Gurpreet Singh Bawa
Systemvoraussetzungen
46,99 inkl. MwSt
  • Download sofort lieferbar
  • Zahlungsarten anzeigen
Text As Data: Combining qualitative and quantitative algorithms within the SAS system for accurate, effective and understandable text analytics

The need for powerful, accurate and increasingly automatic text analysis software in modern information technology has dramatically increased. Fields as diverse as financial management, fraud and cybercrime prevention, Pharmaceutical R&D, social media marketing, customer care, and health services are implementing more comprehensive text-inclusive, analytics strategies. Text as Data: Computational Methods of Understanding Written Expression Using SAS presents an overview of text analytics and the critical role SAS software plays in combining linguistic and quantitative algorithms in the evolution of this dynamic field.

Drawing on over two decades of experience in text analytics, authors Barry deVille and Gurpreet Singh Bawa examine the evolution of text mining and cloud-based solutions, and the development of SAS Visual Text Analytics. By integrating quantitative data and textual analysis with advanced computer learning principles, the authors demonstrate the combined advantages of SAS compared to standard approaches, and show how approaching text as qualitative data within a quantitative analytics framework produces more detailed, accurate, and explanatory results.

* Understand the role of linguistics, machine learning, and multiple data sources in the text analytics workflow

* Understand how a range of quantitative algorithms and data representations reflect contextual effects to shape meaning and understanding

* Access online data and code repositories, videos, tutorials, and case studies

* Learn how SAS extends quantitative algorithms to produce expanded text analytics capabilities

* Redefine text in terms of data for more accurate analysis

This book offers a thorough introduction to the framework and dynamics of text analytics--and the underlying principles at work--and provides an in-depth examination of the interplay between qualitative-linguistic and quantitative, data-driven aspects of data analysis. The treatment begins with a discussion on expression parsing and detection and provides insight into the core principles and practices of text parsing, theme, and topic detection. It includes advanced topics such as contextual effects in numeric and textual data manipulation, fine-tuning text meaning and disambiguation. As the first resource to leverage the power of SAS for text analytics, Text as Data is an essential resource for SAS users and data scientists in any industry or academic application.

BARRY DEVILLE is a Data Scientist and Solutions Architect with 18 years of experience working at SAS. He led the development of the KnowledgeSEEKER decision tree package and has given workshops and tutorials on decision trees for Statistics Canada, the American Marketing Association, the IEEE, and the Direct Marketing Association. GURPREET SINGH BAWA is the Data Science Senior Manager at Accenture PLC in India. He delivers advanced analytics solutions for global clients in a variety of corporate sectors.

Preface xi

Acknowledgments xiii

About the Authors xv

Introduction 1

Chapter 1 Text Mining and Text Analytics 3

Chapter 2 Text Analytics Process Overview 15

Chapter 3 Text Data Source Capture 33

Chapter 4 Document Content and Characterization 43

Chapter 5 Textual Abstraction: Latent Structure, Dimension Reduction 73

Chapter 6 Classification and Prediction 103

Chapter 7 Boolean Methods of Classification and Prediction 125

Chapter 8 Speech to Text 139

Appendix A Mood State Identification in Text 157

Appendix B A Design Approach to Characterizing Users Based on Audio Interactions on a Conversational AI Platform 175

Appendix C SAS Patents in Text Analytics 189

Glossary 197

Index 203

Erscheint lt. Verlag 29.9.2021
Reihe/Serie SAS Institute Inc
SAS Institute Inc
Sprache englisch
Themenwelt Mathematik / Informatik Informatik Datenbanken
Mathematik / Informatik Informatik Theorie / Studium
Schlagworte Business data processing • Computer Science • Informatik • SAS • Textanalyse • Wirtschaftsinformatik
ISBN-10 1-119-48715-3 / 1119487153
ISBN-13 978-1-119-48715-9 / 9781119487159
Haben Sie eine Frage zum Produkt?
EPUBEPUB (Adobe DRM)
Größe: 22,5 MB

Kopierschutz: Adobe-DRM
Adobe-DRM ist ein Kopierschutz, der das eBook vor Mißbrauch schützen soll. Dabei wird das eBook bereits beim Download auf Ihre persönliche Adobe-ID autorisiert. Lesen können Sie das eBook dann nur auf den Geräten, welche ebenfalls auf Ihre Adobe-ID registriert sind.
Details zum Adobe-DRM

Dateiformat: EPUB (Electronic Publication)
EPUB ist ein offener Standard für eBooks und eignet sich besonders zur Darstellung von Belle­tristik und Sach­büchern. Der Fließ­text wird dynamisch an die Display- und Schrift­größe ange­passt. Auch für mobile Lese­geräte ist EPUB daher gut geeignet.

Systemvoraussetzungen:
PC/Mac: Mit einem PC oder Mac können Sie dieses eBook lesen. Sie benötigen eine Adobe-ID und die Software Adobe Digital Editions (kostenlos). Von der Benutzung der OverDrive Media Console raten wir Ihnen ab. Erfahrungsgemäß treten hier gehäuft Probleme mit dem Adobe DRM auf.
eReader: Dieses eBook kann mit (fast) allen eBook-Readern gelesen werden. Mit dem amazon-Kindle ist es aber nicht kompatibel.
Smartphone/Tablet: Egal ob Apple oder Android, dieses eBook können Sie lesen. Sie benötigen eine Adobe-ID sowie eine kostenlose App.
Geräteliste und zusätzliche Hinweise

Buying eBooks from abroad
For tax law reasons we can sell eBooks just within Germany and Switzerland. Regrettably we cannot fulfill eBook-orders from other countries.

Mehr entdecken
aus dem Bereich
der Grundkurs für Ausbildung und Praxis

von Ralf Adams

eBook Download (2023)
Carl Hanser Verlag GmbH & Co. KG
29,99
Das umfassende Handbuch

von Wolfram Langer

eBook Download (2023)
Rheinwerk Computing (Verlag)
49,90