Back to Search

Pro Apache Hadoop

AUTHOR Wadkar, Sameer; Venner, Jason; Siddalingaiah, Madhu
PUBLISHER Apress (09/09/2014)
PRODUCT TYPE Paperback (Paperback)

Description
Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop - the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations. All the old content has been revised too, giving the latest on the ins and outs of MapReduce, cluster design, the Hadoop Distributed File System, and more.

This book covers everything you need to build your first Hadoop cluster and begin analyzing and deriving value from your business and scientific data. Learn to solve big-data problems the MapReduce way, by breaking a big problem into chunks and creating small-scale solutions that can be flung across thousands upon thousands of nodes to analyze large data volumes in a short amount of wall-clock time. Learn how to let Hadoop take care of distributing and parallelizing your software--you just focus on the code; Hadoop takes care of the rest.

  • Covers all that is new in Hadoop 2.0
  • Written by a professional involved in Hadoop since day one
  • Takes you quickly to the seasoned pro level on the hottest cloud-computing framework
Show More
Product Format
Product Details
ISBN-13: 9781430248637
ISBN-10: 1430248637
Binding: Paperback or Softback (Trade Paperback (Us))
Content Language: English
Edition Number: 0002
More Product Details
Page Count: 444
Carton Quantity: 9
Product Dimensions: 7.50 x 0.90 x 9.25 inches
Weight: 1.67 pound(s)
Feature Codes: Illustrated
Country of Origin: NL
Subject Information
BISAC Categories
Computers | Data Science - Data Analytics
Computers | Programming - Parallel
Dewey Decimal: 006.312
Descriptions, Reviews, Etc.
publisher marketing
Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop - the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations. All the old content has been revised too, giving the latest on the ins and outs of MapReduce, cluster design, the Hadoop Distributed File System, and more.

This book covers everything you need to build your first Hadoop cluster and begin analyzing and deriving value from your business and scientific data. Learn to solve big-data problems the MapReduce way, by breaking a big problem into chunks and creating small-scale solutions that can be flung across thousands upon thousands of nodes to analyze large data volumes in a short amount of wall-clock time. Learn how to let Hadoop take care of distributing and parallelizing your software--you just focus on the code; Hadoop takes care of the rest.

  • Covers all that is new in Hadoop 2.0
  • Written by a professional involved in Hadoop since day one
  • Takes you quickly to the seasoned pro level on the hottest cloud-computing framework
Show More

Author: Wadkar, Sameer
Sameer Wadkar has over 15 years of experience in Software Development, and over four years of active development experience in Big Data. He has applied Big Data methods in both the Public and Private Sector. He has also applied Big Data and Data Science methods in the Healthcare and Finance Industries. Sameer has a Bachelor in Electrical Engineering from Mumbai University, and a Masters in Applied Mathematics from Johns Hopkins University.
Show More
List Price $44.99
Your Price  $32.39
Paperback