Beginning Data Science with R (eBook)
XI, 157 Seiten
Springer International Publishing (Verlag)
978-3-319-12066-9 (ISBN)
The goal of 'Beginning Data Science with R' is to introduce the readers to some of the useful data science techniques and their implementation with the R programming language. The book attempts to strike a balance between the how: specific processes and methodologies, and understanding the why: going over the intuition behind how a particular technique works, so that the reader can apply it to the problem at hand. This book will be useful for readers who are not familiar with statistics and the R programming language.
Dr. Manas A. Pathak received a BTech degree in computer science from Visvesvaraya National Institute of Technology, Nagpur, India, in 2006, and MS and PhD degrees from the Language Technologies Institute at Carnegie Mellon University (CMU) in 2009 and 2012 respectively. His PhD thesis on 'Privacy-Preserving Machine Learning for Speech Processing' was published as a monograph in the Springer best thesis series. His research received significant press coverage, including articles in the Economist and MIT Tech Review. He has many years of experience with data analysis using the R programming language. He is currently working as a staff software engineer at @WalmartLabs.
Dr. Manas A. Pathak received a BTech degree in computer science from Visvesvaraya National Institute of Technology, Nagpur, India, in 2006, and MS and PhD degrees from the Language Technologies Institute at Carnegie Mellon University (CMU) in 2009 and 2012 respectively. His PhD thesis on "Privacy-Preserving Machine Learning for Speech Processing" was published as a monograph in the Springer best thesis series. His research received significant press coverage, including articles in the Economist and MIT Tech Review. He has many years of experience with data analysis using the R programming language. He is currently working as a staff software engineer at @WalmartLabs.
4.4 Interactive Visualizations using Shiny 4.5 Chapter Summary & Further Reading References 5 Exploratory Data Analysis 5.1 Summary Statistics 5.1.1 Dataset Size 5.1.2 Summarizing the Data 5.1.3 Ordering Data by a Variable 5.1.4 Group and Split Data by a Variable 5.1.5 Variable Correlation 5.2 Getting a sense of data distribution 5.2.1 Box plots 5.2.2 Histograms 5.2.3 Measuring Data Symmetry using Skewness and Kurtosis 5.3 Putting it all together: Outlier Detection 5.4 Chapter Summary References 6 Regression 6.1 Introduction 6.1.1 Regression Models 6.2 Parametric Regression Models 6.2.1 Simple Linear Regression 6.2.2 Multivariate Linear Regression 6.2.3 Log-Linear Regression Models 6.3 Non-Parametric Regression Models 6.3.1 Locally Weighted Regression 6.3.2 Kernel Regression 6.3.3 Regression Trees 6.4 Chapter Summary References 7 Classification 7.1 Introduction 7.1.1 Training and Test Datasets 7.2 Parametric Classification Models 7.2.1 Naive Bayes 7.2.2 Logistic Regression 7.2.3 Support Vector Machines 7.3 Non-Parametric Classification Models 7.3.1 Nearest Neighbors 7.3.2 Decision Trees 7.4 Chapter Summary References 8 Text Mining 8.1 Introduction 8.2 Reading Text Input Data 8.3 Common Text Preprocessing Tasks 8.3.1 Stop Word Removal 8.3.2 Stemming 8.4 Term Document Matrix 8.4.1 TF-IDF Weighting Function 8.5 Text Mining Applications 8.5.1 Frequency Analysis 8.5.2 Text Classification 8.6 Chapter Summary
Erscheint lt. Verlag | 8.12.2014 |
---|---|
Zusatzinfo | XI, 157 p. 155 illus., 26 illus. in color. |
Verlagsort | Cham |
Sprache | englisch |
Themenwelt | Mathematik / Informatik ► Informatik |
Technik ► Elektrotechnik / Energietechnik | |
Schlagworte | Creating Tag Clouds • R Code • R Interfacing • R Programming • Social Network Data Analysis • statistical modeling • time series data |
ISBN-10 | 3-319-12066-2 / 3319120662 |
ISBN-13 | 978-3-319-12066-9 / 9783319120669 |
Haben Sie eine Frage zum Produkt? |
Größe: 4,0 MB
DRM: Digitales Wasserzeichen
Dieses eBook enthält ein digitales Wasserzeichen und ist damit für Sie personalisiert. Bei einer missbräuchlichen Weitergabe des eBooks an Dritte ist eine Rückverfolgung an die Quelle möglich.
Dateiformat: PDF (Portable Document Format)
Mit einem festen Seitenlayout eignet sich die PDF besonders für Fachbücher mit Spalten, Tabellen und Abbildungen. Eine PDF kann auf fast allen Geräten angezeigt werden, ist aber für kleine Displays (Smartphone, eReader) nur eingeschränkt geeignet.
Systemvoraussetzungen:
PC/Mac: Mit einem PC oder Mac können Sie dieses eBook lesen. Sie benötigen dafür einen PDF-Viewer - z.B. den Adobe Reader oder Adobe Digital Editions.
eReader: Dieses eBook kann mit (fast) allen eBook-Readern gelesen werden. Mit dem amazon-Kindle ist es aber nicht kompatibel.
Smartphone/Tablet: Egal ob Apple oder Android, dieses eBook können Sie lesen. Sie benötigen dafür einen PDF-Viewer - z.B. die kostenlose Adobe Digital Editions-App.
Buying eBooks from abroad
For tax law reasons we can sell eBooks just within Germany and Switzerland. Regrettably we cannot fulfill eBook-orders from other countries.
aus dem Bereich