Apache Spark 2 for Beginners (eBook)
332 Seiten
Packt Publishing (Verlag)
978-1-78588-669-0 (ISBN)
Develop large-scale distributed data processing applications using Spark 2 in Scala and Python
About This Book
- This book offers an easy introduction to the Spark framework published on the latest version of Apache Spark 2
- Perform efficient data processing, machine learning and graph processing using various Spark components
- A practical guide aimed at beginners to get them up and running with Spark
Who This Book Is For
If you are an application developer, data scientist, or big data solutions architect who is interested in combining the data processing power of Spark from R, and consolidating data processing, stream processing, machine learning, and graph processing into one unified and highly interoperable framework with a uniform API using Scala or Python, this book is for you.
What You Will Learn
- Get to know the fundamentals of Spark 2 and the Spark programming model using Scala and Python
- Know how to use Spark SQL and DataFrames using Scala and Python
- Get an introduction to Spark programming using R
- Perform Spark data processing, charting, and plotting using Python
- Get acquainted with Spark stream processing using Scala and Python
- Be introduced to machine learning using Spark MLlib
- Get started with graph processing using the Spark GraphX
- Bring together all that you've learned and develop a complete Spark application
In Detail
Spark is one of the most widely-used large-scale data processing engines and runs extremely fast. It is a framework that has tools that are equally useful for application developers as well as data scientists.
This book starts with the fundamentals of Spark 2 and covers the core data processing framework and API, installation, and application development setup. Then the Spark programming model is introduced through real-world examples followed by Spark SQL programming with DataFrames. An introduction to SparkR is covered next. Later, we cover the charting and plotting features of Python in conjunction with Spark data processing. After that, we take a look at Spark's stream processing, machine learning, and graph processing libraries. The last chapter combines all the skills you learned from the preceding chapters to develop a real-world Spark application.
By the end of this book, you will have all the knowledge you need to develop efficient large-scale applications using Apache Spark.
Style and approach
Learn about Spark's infrastructure with this practical tutorial. With the help of real-world use cases on the main features of Spark we offer an easy introduction to the framework.
Develop large-scale distributed data processing applications using Spark 2 in Scala and PythonAbout This BookThis book offers an easy introduction to the Spark framework published on the latest version of Apache Spark 2Perform efficient data processing, machine learning and graph processing using various Spark componentsA practical guide aimed at beginners to get them up and running with SparkWho This Book Is ForIf you are an application developer, data scientist, or big data solutions architect who is interested in combining the data processing power of Spark from R, and consolidating data processing, stream processing, machine learning, and graph processing into one unified and highly interoperable framework with a uniform API using Scala or Python, this book is for you.What You Will LearnGet to know the fundamentals of Spark 2 and the Spark programming model using Scala and PythonKnow how to use Spark SQL and DataFrames using Scala and PythonGet an introduction to Spark programming using RPerform Spark data processing, charting, and plotting using PythonGet acquainted with Spark stream processing using Scala and PythonBe introduced to machine learning using Spark MLlibGet started with graph processing using the Spark GraphXBring together all that you've learned and develop a complete Spark applicationIn DetailSpark is one of the most widely-used large-scale data processing engines and runs extremely fast. It is a framework that has tools that are equally useful for application developers as well as data scientists.This book starts with the fundamentals of Spark 2 and covers the core data processing framework and API, installation, and application development setup. Then the Spark programming model is introduced through real-world examples followed by Spark SQL programming with DataFrames. An introduction to SparkR is covered next. Later, we cover the charting and plotting features of Python in conjunction with Spark data processing. After that, we take a look at Spark's stream processing, machine learning, and graph processing libraries. The last chapter combines all the skills you learned from the preceding chapters to develop a real-world Spark application.By the end of this book, you will have all the knowledge you need to develop efficient large-scale applications using Apache Spark.Style and approachLearn about Spark's infrastructure with this practical tutorial. With the help of real-world use cases on the main features of Spark we offer an easy introduction to the framework.
Erscheint lt. Verlag | 6.10.2016 |
---|---|
Sprache | englisch |
Themenwelt | Sachbuch/Ratgeber ► Freizeit / Hobby ► Sammeln / Sammlerkataloge |
Mathematik / Informatik ► Informatik ► Datenbanken | |
Mathematik / Informatik ► Informatik ► Programmiersprachen / -werkzeuge | |
Mathematik / Informatik ► Informatik ► Theorie / Studium | |
ISBN-10 | 1-78588-669-X / 178588669X |
ISBN-13 | 978-1-78588-669-0 / 9781785886690 |
Haben Sie eine Frage zum Produkt? |
Größe: 6,1 MB
Kopierschutz: Adobe-DRM
Adobe-DRM ist ein Kopierschutz, der das eBook vor Mißbrauch schützen soll. Dabei wird das eBook bereits beim Download auf Ihre persönliche Adobe-ID autorisiert. Lesen können Sie das eBook dann nur auf den Geräten, welche ebenfalls auf Ihre Adobe-ID registriert sind.
Details zum Adobe-DRM
Dateiformat: EPUB (Electronic Publication)
EPUB ist ein offener Standard für eBooks und eignet sich besonders zur Darstellung von Belletristik und Sachbüchern. Der Fließtext wird dynamisch an die Display- und Schriftgröße angepasst. Auch für mobile Lesegeräte ist EPUB daher gut geeignet.
Systemvoraussetzungen:
PC/Mac: Mit einem PC oder Mac können Sie dieses eBook lesen. Sie benötigen eine
eReader: Dieses eBook kann mit (fast) allen eBook-Readern gelesen werden. Mit dem amazon-Kindle ist es aber nicht kompatibel.
Smartphone/Tablet: Egal ob Apple oder Android, dieses eBook können Sie lesen. Sie benötigen eine
Geräteliste und zusätzliche Hinweise
Buying eBooks from abroad
For tax law reasons we can sell eBooks just within Germany and Switzerland. Regrettably we cannot fulfill eBook-orders from other countries.
aus dem Bereich