Pro Apache Hadoop

Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop – the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations....

Full description

Bibliographic Details
Main Authors: Venner, Jason, Wadkar, Sameer (Author), Siddalingaiah, Madhu (Author)
Format: eBook
Language:English
Published: Berkeley, CA Apress 2014, 2014
Edition:2nd ed. 2014
Subjects:
Online Access:
Collection: Springer eBooks 2005- - Collection details see MPG.ReNa
LEADER 02135nmm a2200289 u 4500
001 EB000897031
003 EBX01000000000000000694151
005 00000000000000.0
007 cr|||||||||||||||||||||
008 141008 ||| eng
020 |a 9781430248644 
100 1 |a Venner, Jason 
245 0 0 |a Pro Apache Hadoop  |h Elektronische Ressource  |c by Jason Venner, Sameer Wadkar, Madhu Siddalingaiah 
250 |a 2nd ed. 2014 
260 |a Berkeley, CA  |b Apress  |c 2014, 2014 
300 |a XXII, 444 p. 70 illus  |b online resource 
653 |a Open source software 
653 |a Data mining 
653 |a Data Mining and Knowledge Discovery 
653 |a Open Source 
700 1 |a Wadkar, Sameer  |e [author] 
700 1 |a Siddalingaiah, Madhu  |e [author] 
041 0 7 |a eng  |2 ISO 639-2 
989 |b Springer  |a Springer eBooks 2005- 
856 4 0 |u https://doi.org/10.1007/978-1-4302-4864-4?nosfx=y  |x Verlag  |3 Volltext 
082 0 |a 005.3 
520 |a Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop – the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations. All the old content has been revised too, giving the latest on the ins and outs of MapReduce, cluster design, the Hadoop Distributed File System, and more. This book covers everything you need to build your first Hadoop cluster and begin analyzing and deriving value from your business and scientific data. Learn to solve big-data problems the MapReduce way, by breaking a big problem into chunks and creating small-scale solutions that can be flung across thousands upon thousands of nodes to analyze large data volumes in a short amount of wall-clock time. Learn how to let Hadoop take care of distributing and parallelizing your software—you just focus on the code; Hadoop takes care of the rest. Covers all that is new in Hadoop 2.0 Written by a professional involved in Hadoop since day one Takes you quickly to the seasoned pro level on the hottest cloud-computing framework