The Super Fun Kids' Graphic Novel Sale

Special Offers see all

Enter to WIN a $100 Credit

Subscribe to
for a chance to win.
Privacy Policy

Visit our stores

    Recently Viewed clear list

    Contributors | September 15, 2015

    Mary Karr: IMG Memoir Tutorials with Mary Karr, Lena Dunham, and Gary Shteyngart

    Editor's note: It's been 20 years since the groundbreaking memoir The Liars' Club sent Mary Karr into the literary spotlight with its phenomenal... Continue »
    1. $17.49 Sale Hardcover add to wish list

      The Art of Memoir

      Mary Karr 9780062223067

Qualifying orders ship free.
New Trade Paper
Ships in 1 to 3 days
Add to Wishlist
Qty Store Section
5 Local Warehouse Internet- Apache

More copies of this ISBN

Hadoop: The Definitive Guide


Hadoop: The Definitive Guide Cover


Out of Print

Synopses & Reviews

Publisher Comments:

Hadoop: The Definitive Guide helps you harness the power of your data. Ideal for processing large datasets, the Apache Hadoop framework is an open source implementation of the MapReduce algorithm on which Google built its empire. This comprehensive resource demonstrates how to use Hadoop to build reliable, scalable, distributed systems: programmers will find details for analyzing large datasets, and administrators will learn how to set up and run Hadoop clusters.

Complete with case studies that illustrate how Hadoop solves specific problems, this book helps you:

  • Use the Hadoop Distributed File System (HDFS) for storing large datasets, and run distributed computations over those datasets using MapReduce
  • Become familiar with Hadoop's data and I/O building blocks for compression, data integrity, serialization, and persistence
  • Discover common pitfalls and advanced features for writing real-world MapReduce programs
  • Design, build, and administer a dedicated Hadoop cluster, or run Hadoop in the cloud
  • Use Pig, a high-level query language for large-scale data processing
  • Take advantage of HBase, Hadoop's database for structured and semi-structured data
  • Learn ZooKeeper, a toolkit of coordination primitives for building distributed systems

If you have lots of data — whether it's gigabytes or petabytes — Hadoop is the perfect solution. Hadoop: The Definitive Guide is the most thorough book available on the subject.

"Now you have the opportunity to learn about Hadoop from a master-not only of the technology, but also of common sense and plain talk."-- Doug Cutting, Hadoop Founder, Yahoo!


Organizations large and small are adopting Apache Hadoop to deal with huge application data sets, and this comprehensive resource provides the key for unlocking the wealth this data holds.

About the Author

Tom White has been an Apache Hadoop committer since February 2007, and is a member of the Apache Software Foundation. He works for Cloudera, a company set up to offer Hadoop support and training. Previously he was as an independent Hadoop consultant, working with companies to set up, use, and extend Hadoop. He has written numerous articles for O'Reilly, and IBM's developerWorks, and has spoken at several conferences, including at ApacheCon 2008 on Hadoop. Tom has a Bachelor's degree in Mathematics from the University of Cambridge and a Master's in Philosophy of Science from the University of Leeds, UK.

Table of Contents

ForewordPrefaceChapter 1: Meet HadoopChapter 2: MapReduceChapter 3: The Hadoop Distributed FilesystemChapter 4: Hadoop I/OChapter 5: Developing a MapReduce ApplicationChapter 6: How MapReduce WorksChapter 7: MapReduce Types and FormatsChapter 8: MapReduce FeaturesChapter 9: Setting Up a Hadoop ClusterChapter 10: Administering HadoopChapter 11: PigChapter 12: HBaseChapter 13: ZooKeeperChapter 14: Case StudiesInstalling Apache HadoopClouderas Distribution for HadoopPreparing the NCDC Weather DataColophon

Product Details

The Definitive Guide
White, Tom
Foreword by:
Cutting, Doug
Cutting, Doug
O'Reilly Media
Programming Languages - JavaScript
Programming Languages - CGI, Javascript, Perl, VBScript
Programming / Parallel
cloud;cloud computing;cluster;database;distributed computing;google;hadoop;java;mapreduce
cloud;cloud com
puting;cluster;database;distributed computing;google;hadoop;java;mapreduce
CourseSmart Subject Description
Edition Description:
Print PDF
Publication Date:
9.19 x 7.00 in

Other books you might like

  1. Pro Hadoop New Trade Paper $39.99

Related Subjects

Computers and Internet » Computers Reference » General
Computers and Internet » Internet » Apache
Computers and Internet » Software Engineering » General

Hadoop: The Definitive Guide New Trade Paper
0 stars - 0 reviews
$44.99 In Stock
Product details 528 pages O'Reilly Media - English 9780596521974 Reviews:
"Synopsis" by ,
Organizations large and small are adopting Apache Hadoop to deal with huge application data sets, and this comprehensive resource provides the key for unlocking the wealth this data holds.
  • back to top


Powell's City of Books is an independent bookstore in Portland, Oregon, that fills a whole city block with more than a million new, used, and out of print books. Shop those shelves — plus literally millions more books, DVDs, and gifts — here at