Pro Apache Hadoop(2nd Edition)

Pro Apache Hadoop(2nd Edition)

作者: 
Sameer Wadkar / Madhu Siddalingaiah / Jason Venner
语言: 
ISBN: 
9781430248637
出版日期: 
星期一, 九月 1, 2014

简介

Pro Apache Hadoop, Second Edition brings you up to speed on Hadoop – the framework of big data. Revised to cover Hadoop 2.0, the book covers the very latest developments such as YARN (aka MapReduce 2.0), new HDFS high-availability features, and increased scalability in the form of HDFS Federations. All the old content has been revised too, giving the latest on the ins and outs of MapReduce, cluster design, the Hadoop Distributed File System, and more.
This book covers everything you need to build your first Hadoop cluster and begin analyzing and deriving value from your business and scientific data. Learn to solve big-data problems the MapReduce way, by breaking a big problem into chunks and creating small-scale solutions that can be flung across thousands upon thousands of nodes to analyze large data volumes in a short amount of wall-clock time. Learn how to let Hadoop take care of distributing and parallelizing your software – you just focus on the code; Hadoop takes care of the rest.

目录

About the Authors
About the Technical Reviewer
Acknowledgments
Introduction
Chapter 1: Motivation for Big Data
Chapter 2: Hadoop Concepts
Chapter 3: Getting Started with the Hadoop Framework
Chapter 4: Hadoop Administration
Chapter 5: Basics of MapReduce Development
Chapter 6: Advanced MapReduce Development
Chapter 7: Hadoop Input/Output
Chapter 8: Testing Hadoop Programs
Chapter 9: Monitoring Hadoop
Chapter 10: Data Warehousing Using Hadoop
Chapter 11: Data Processing Using Pig
Chapter 12: HCatalog and Hadoop in the Enterprise
Chapter 13: Log Analysis Using Hadoop
Chapter 14: Building Real-Time Systems Using HBase
Chapter 15: Data Science with Hadoop
Chapter 16: Hadoop in the Cloud
Chapter 17: Building a YARN Application
Appendix A: Installing Hadoop
Appendix B: Using Maven with Eclipse
Appendix C: Apache Ambari
Index