- Softcover
- New

Seller: Rarewaves.com UK, London, United KingdomRarewaves.com UK
AbeBooks seller since June 11, 2025
Condition: New
US$ 39.11
Quantity: 1 available
Add to basketItem description from seller
If your organization is about to enter the world of big data, you not only need to decide whether Apache Hadoop is the right platform to use, but also which of its many components are best suited to your task. This field guide makes the exercise manageable by breaking down the Hadoop ecosystem into short, digestible sections. You'll quickly understand how Hadoop's projects, subprojects, and related technologies work together. Each chapter introduces a different topic-such as core technologies or data transfer-and explains why certain components may or may not be useful for particular needs. When it comes to data, Hadoop is a whole new ballgame, but with this handy reference, you'll have a good grasp of the playing field.Topics include: Core technologies-Hadoop Distributed File System (HDFS), MapReduce, YARN, and Spark Database and data management-Cassandra, HBase, MongoDB, and Hive Serialization-Avro, JSON, and Parquet Management and monitoring-Puppet, Chef, Zookeeper, and Oozie Analytic helpers-Pig, Mahout, and MLLib Data transfer-Scoop, Flume, distcp, and Storm Security, access control, auditing-Sentry, Kerberos, and Knox Cloud computing and virtualization-Serengeti, Docker, and Whirr.…
Seller Inventory # LU-9781491947937
- Title
- Field Guide to Hadoop
- Author
- Marshall Sitto
- Publisher
- O'Reilly Media, US
- Publication year
- 2015
- Condition
- New
- Binding
- Paperback
- Language
- English
- ISBN 10
- 1491947934
- ISBN 13
- 9781491947937
- Item weight
- 180 grams
If your organization is about to enter the world of big data, you not only need to decide whether Apache Hadoop is the right platform to use, but also which of its many components are best suited to your task. This field guide makes the exercise manageable by breaking down the Hadoop ecosystem into short, digestible sections. You’ll quickly understand how Hadoop’s projects, subprojects, and related technologies work together.
Each chapter introduces a different topic—such as core technologies or data transfer—and explains why certain components may or may not be useful for particular needs. When it comes to data, Hadoop is a whole new ballgame, but with this handy reference, you’ll have a good grasp of the playing field.
Topics include:
- Core technologies—Hadoop Distributed File System (HDFS), MapReduce, YARN, and Spark
- Database and data management—Cassandra, HBase, MongoDB, and Hive
- Serialization—Avro, JSON, and Parquet
- Management and monitoring—Puppet, Chef, Zookeeper, and Oozie
- Analytic helpers—Pig, Mahout, and MLLib
- Data transfer—Scoop, Flume, distcp, and Storm
- Security, access control, auditing—Sentry, Kerberos, and Knox
- Cloud computing and virtualization—Serengeti, Docker, and Whirr
"Synopsis" may belong to another edition of this title.
About the Author
Kevin Sitto is a Field Solutions Engineer with Pivotal Software, providing consulting services to help folks understand and address their big data needs.
He lives in Maryland with his wife and two kids and enjoys making homebrew beer when he's not writing books about big data.
"About the title" may belong to another edition of this title.
Rarewaves.com UK
London, United Kingdom
AbeBooks seller since June 11, 2025
Shipping rates from United Kingdom to U.S.A.
| Item | 60 to 60 business days | 60 to 60 business days |
|---|---|---|
| First item | US$ 86.06 | US$ 132.40 |
Payment methods
Seller's business information
RAREWAVES.COM LIMITED
Elsley Court, 20-22 Great Titchfield Street
London, United Kingdom W1W 8BE
Shipping terms
Please note that we do not offer Priority shipping to any country.
We currently do not ship to the below countries:
Russia
Belarus
Ukraine
Please do not attempt to place orders with any of these countries as a ship to address - they will be cancelled.