Summer Reading B2G1 Free

Special Offers see all

Enter to WIN a $100 Credit

Subscribe to
for a chance to win.
Privacy Policy

Visit our stores

    Recently Viewed clear list

    Lists | July 16, 2015

    Annie Liontas: IMG "You Want Me to Smell My Fingers?": Five Unforgettable Greek Idioms

    The word "idiom" originates in the Greek word ídios ("one's own") and means "special feature" or "special phrasing." Idioms are peculiar because,... Continue »
    1. $18.20 Sale Hardcover add to wish list

      Let Me Explain You

      Annie Liontas 9781476789088

Qualifying orders ship free.
New Trade Paper
Ships in 1 to 3 days
Add to Wishlist
Qty Store Section
5 Local Warehouse Internet- Apache

More copies of this ISBN

Hadoop: The Definitive Guide


Hadoop: The Definitive Guide Cover


Out of Print

Synopses & Reviews

Publisher Comments:

Hadoop: The Definitive Guide helps you harness the power of your data. Ideal for processing large datasets, the Apache Hadoop framework is an open source implementation of the MapReduce algorithm on which Google built its empire. This comprehensive resource demonstrates how to use Hadoop to build reliable, scalable, distributed systems: programmers will find details for analyzing large datasets, and administrators will learn how to set up and run Hadoop clusters.

Complete with case studies that illustrate how Hadoop solves specific problems, this book helps you:

  • Use the Hadoop Distributed File System (HDFS) for storing large datasets, and run distributed computations over those datasets using MapReduce
  • Become familiar with Hadoop's data and I/O building blocks for compression, data integrity, serialization, and persistence
  • Discover common pitfalls and advanced features for writing real-world MapReduce programs
  • Design, build, and administer a dedicated Hadoop cluster, or run Hadoop in the cloud
  • Use Pig, a high-level query language for large-scale data processing
  • Take advantage of HBase, Hadoop's database for structured and semi-structured data
  • Learn ZooKeeper, a toolkit of coordination primitives for building distributed systems

If you have lots of data — whether it's gigabytes or petabytes — Hadoop is the perfect solution. Hadoop: The Definitive Guide is the most thorough book available on the subject.

"Now you have the opportunity to learn about Hadoop from a master-not only of the technology, but also of common sense and plain talk."-- Doug Cutting, Hadoop Founder, Yahoo!


Organizations large and small are adopting Apache Hadoop to deal with huge application data sets, and this comprehensive resource provides the key for unlocking the wealth this data holds.

About the Author

Tom White has been an Apache Hadoop committer since February 2007, and is a member of the Apache Software Foundation. He works for Cloudera, a company set up to offer Hadoop support and training. Previously he was as an independent Hadoop consultant, working with companies to set up, use, and extend Hadoop. He has written numerous articles for O'Reilly, and IBM's developerWorks, and has spoken at several conferences, including at ApacheCon 2008 on Hadoop. Tom has a Bachelor's degree in Mathematics from the University of Cambridge and a Master's in Philosophy of Science from the University of Leeds, UK.

Table of Contents

ForewordPrefaceChapter 1: Meet HadoopChapter 2: MapReduceChapter 3: The Hadoop Distributed FilesystemChapter 4: Hadoop I/OChapter 5: Developing a MapReduce ApplicationChapter 6: How MapReduce WorksChapter 7: MapReduce Types and FormatsChapter 8: MapReduce FeaturesChapter 9: Setting Up a Hadoop ClusterChapter 10: Administering HadoopChapter 11: PigChapter 12: HBaseChapter 13: ZooKeeperChapter 14: Case StudiesInstalling Apache HadoopClouderas Distribution for HadoopPreparing the NCDC Weather DataColophon

Product Details

The Definitive Guide
White, Tom
Foreword by:
Cutting, Doug
Cutting, Doug
O'Reilly Media
Programming Languages - JavaScript
Programming Languages - CGI, Javascript, Perl, VBScript
Programming / Parallel
cloud;cloud computing;cluster;database;distributed computing;google;hadoop;java;mapreduce
cloud;cloud com
puting;cluster;database;distributed computing;google;hadoop;java;mapreduce
CourseSmart Subject Description
Edition Description:
Print PDF
Publication Date:
9.19 x 7.00 in

Other books you might like

  1. Pro Hadoop New Trade Paper $39.99

Related Subjects

Computers and Internet » Computers Reference » General
Computers and Internet » Internet » Apache
Computers and Internet » Software Engineering » General

Hadoop: The Definitive Guide New Trade Paper
0 stars - 0 reviews
$44.99 In Stock
Product details 528 pages O'Reilly Media - English 9780596521974 Reviews:
"Synopsis" by ,
Organizations large and small are adopting Apache Hadoop to deal with huge application data sets, and this comprehensive resource provides the key for unlocking the wealth this data holds.
  • back to top


Powell's City of Books is an independent bookstore in Portland, Oregon, that fills a whole city block with more than a million new, used, and out of print books. Shop those shelves — plus literally millions more books, DVDs, and gifts — here at