Pro Apache Hadoop - Jason Venner, Sameer Wadkar, Madhu Siddalingaiah

Pro Apache Hadoop

Buch | Softcover
444 Seiten
2014 | 2nd ed.
Apress (Verlag)
978-1-4302-4863-7 (ISBN)
48,14 inkl. MwSt
Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop – the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations. All the old content has been revised too, giving the latest on the ins and outs of MapReduce, cluster design, the Hadoop Distributed File System, and more. This book covers everything you need to build your first Hadoop cluster and begin analyzing and deriving value from your business and scientific data. Learn to solve big-data problems the MapReduce way, by breaking a big problem into chunks and creating small-scale solutions that can be flung across thousands upon thousands of nodes to analyze large data volumes in a short amount of wall-clock time. Learn how to let Hadoop take care of distributing and parallelizing your software—you just focus on the code; Hadoop takes care of the rest.





Covers all that is new in Hadoop 2.0
Written by a professional involved in Hadoop since day one
Takes you quickly to the seasoned pro level on the hottest cloud-computing framework  

Jason Venner has more than 20 years of software engineering, managing, designing, and coding experience. He has been a vice president, director, and consultant. Currently, his interests and expertise are in Java, Hadoop, cloud computing, and more. For more, visit www.prohadoopbook.com.

1. Motivation for Big Data 2. Hadoop Concepts 3. Getting Started with the Hadoop Framework 4. Hadoop Administration 5. Basics of MapReduce Development 6. Advanced MapReduce Development 7. Hadoop Input Output 8. Testing Hadoop Programs 9. Monitoring Hadoop 10. Data Warehousing using Hadoop 11. Data Processing using Pig 12. HCatalog and Hadoop in the Enterprise 13. Log Analysis using Hadoop 14. Building Real-Time Systems using HBase 15. Data Science With Hadoop 16. Hadoop in the Cloud 17. Building a YARN Application 18. Appendix A 19. Appendix B 20. Appendix C

Erscheint lt. Verlag 9.9.2014
Zusatzinfo 70 Illustrations, black and white; XXII, 444 p. 70 illus.
Verlagsort Berlin
Sprache englisch
Maße 178 x 254 mm
Themenwelt Mathematik / Informatik Informatik Betriebssysteme / Server
Informatik Datenbanken Data Warehouse / Data Mining
Informatik Theorie / Studium Künstliche Intelligenz / Robotik
ISBN-10 1-4302-4863-7 / 1430248637
ISBN-13 978-1-4302-4863-7 / 9781430248637
Zustand Neuware
Haben Sie eine Frage zum Produkt?
Mehr entdecken
aus dem Bereich
Auswertung von Daten mit pandas, NumPy und IPython

von Wes McKinney

Buch | Softcover (2023)
O'Reilly (Verlag)
44,90