Spark in Action, Second Edition: Covers Apache Spark 3 with Examples in Java, Python, and Scala
26 ratings by Goodreads
Language: English
Published by Manning (edition 2nd), 2020
Series: Book 129 of 182 - In Action
- Softcover
- Used

Seller: BooksRun, Philadelphia, PA, U.S.A.BooksRun
5-star seller
AbeBooks seller since February 2, 2016
Softcover
Condition: Used - Fair
US$ 9.40
Free Shipping
Ships within U.S.A.
Quantity: 1 available
Add to basketFree 30-day returns
Item description from seller
The item might be beaten up but readable. May contain markings or highlighting, as well as stains, bent corners, or any other major defect, but the text is not obscured in any way.
Seller Inventory # 1617295523-7-1-13
- Title
- Spark in Action, Second Edition: Covers Apache Spark 3 with Examples in Java, Python, and Scala
- Author
- Perrin, Jean-Georges
- Publisher
- Manning (edition 2nd)
- Publication year
- 2020
- Condition
- Fair
- Binding
- Paperback
- Language
- English
- ISBN 10
- 1617295523
- ISBN 13
- 9781617295522
- Edition
- 2nd.
- Series
- Book 129 of 182: In Action
Summary
The Spark distributed data processing platform provides an easy-to-implement tool for ingesting, streaming, and processing data from any source. In Spark in Action, Second Edition, you’ll learn to take advantage of Spark’s core features and incredible processing speed, with applications including real-time computation, delayed evaluation, and machine learning. Spark skills are a hot commodity in enterprises worldwide, and with Spark’s powerful and flexible Java APIs, you can reap all the benefits without first learning Scala or Hadoop.
Foreword by Rob Thomas.
About the technology
Analyzing enterprise data starts by reading, filtering, and merging files and streams from many sources. The Spark data processing engine handles this varied volume like a champ, delivering speeds 100 times faster than Hadoop systems. Thanks to SQL support, an intuitive interface, and a straightforward multilanguage API, you can use Spark without learning a complex new ecosystem.
About the book
Spark in Action, Second Edition, teaches you to create end-to-end analytics applications. In this entirely new book, you’ll learn from interesting Java-based examples, including a complete data pipeline for processing NASA satellite data. And you’ll discover Java, Python, and Scala code samples hosted on GitHub that you can explore and adapt, plus appendixes that give you a cheat sheet for installing tools and understanding Spark-specific terms.
What's inside
Writing Spark applications in Java
Spark application architecture
Ingestion through files, databases, streaming, and Elasticsearch
Querying distributed datasets with Spark SQL
About the reader
This book does not assume previous experience with Spark, Scala, or Hadoop.
About the author
Jean-Georges Perrin is an experienced data and software architect. He is France’s first IBM Champion and has been honored for 12 consecutive years.
Table of Contents
PART 1 - THE THEORY CRIPPLED BY AWESOME EXAMPLES
1 So, what is Spark, anyway?
2 Architecture and flow
3 The majestic role of the dataframe
4 Fundamentally lazy
5 Building a simple app for deployment
6 Deploying your simple app
PART 2 - INGESTION
7 Ingestion from files
8 Ingestion from databases
9 Advanced ingestion: finding data sources and building
your own
10 Ingestion through structured streaming
PART 3 - TRANSFORMING YOUR DATA
11 Working with SQL
12 Transforming your data
13 Transforming entire documents
14 Extending transformations with user-defined functions
15 Aggregating your data
PART 4 - GOING FURTHER
16 Cache and checkpoint: Enhancing Spark’s performances
17 Exporting data and building full data pipelines
18 Exploring deployment
The Spark distributed data processing platform provides an easy-to-implement tool for ingesting, streaming, and processing data from any source. In Spark in Action, Second Edition, you’ll learn to take advantage of Spark’s core features and incredible processing speed, with applications including real-time computation, delayed evaluation, and machine learning. Spark skills are a hot commodity in enterprises worldwide, and with Spark’s powerful and flexible Java APIs, you can reap all the benefits without first learning Scala or Hadoop.
Foreword by Rob Thomas.
About the technology
Analyzing enterprise data starts by reading, filtering, and merging files and streams from many sources. The Spark data processing engine handles this varied volume like a champ, delivering speeds 100 times faster than Hadoop systems. Thanks to SQL support, an intuitive interface, and a straightforward multilanguage API, you can use Spark without learning a complex new ecosystem.
About the book
Spark in Action, Second Edition, teaches you to create end-to-end analytics applications. In this entirely new book, you’ll learn from interesting Java-based examples, including a complete data pipeline for processing NASA satellite data. And you’ll discover Java, Python, and Scala code samples hosted on GitHub that you can explore and adapt, plus appendixes that give you a cheat sheet for installing tools and understanding Spark-specific terms.
What's inside
Writing Spark applications in Java
Spark application architecture
Ingestion through files, databases, streaming, and Elasticsearch
Querying distributed datasets with Spark SQL
About the reader
This book does not assume previous experience with Spark, Scala, or Hadoop.
About the author
Jean-Georges Perrin is an experienced data and software architect. He is France’s first IBM Champion and has been honored for 12 consecutive years.
Table of Contents
PART 1 - THE THEORY CRIPPLED BY AWESOME EXAMPLES
1 So, what is Spark, anyway?
2 Architecture and flow
3 The majestic role of the dataframe
4 Fundamentally lazy
5 Building a simple app for deployment
6 Deploying your simple app
PART 2 - INGESTION
7 Ingestion from files
8 Ingestion from databases
9 Advanced ingestion: finding data sources and building
your own
10 Ingestion through structured streaming
PART 3 - TRANSFORMING YOUR DATA
11 Working with SQL
12 Transforming your data
13 Transforming entire documents
14 Extending transformations with user-defined functions
15 Aggregating your data
PART 4 - GOING FURTHER
16 Cache and checkpoint: Enhancing Spark’s performances
17 Exporting data and building full data pipelines
18 Exploring deployment
"Synopsis" may belong to another edition of this title.
About the Author
Jean-Georges “jgp” Perrin is a technology leader focusing on building innovative and modern data platforms, author, and president of AIDA User Group. He is passionate about software engineering and all things data, including Data Mesh. He is proud to have been recognized as a Lifetime IBM Champion.
"About the title" may belong to another edition of this title.
BooksRun
Philadelphia, PA, U.S.A.
5-star seller
AbeBooks seller since February 2, 2016
Shipping rates within U.S.A.
| Item | 3 to 8 business days | 3 to 6 business days |
|---|---|---|
| First item | US$ 0.00 | US$ 3.99 |
Payment methods
Store description
BooksRun helps save money on books. Founded in 2014, we are an independent online bookseller with thousands of happy customers and top ratings. With millions of titles in stock, from fiction to textbooks, we have the best book selection and prices 90% below the list price.
We ship all orders the same day or the next business day. Expedited shipping arrives in 2 - 5 business days. Returns are accepted within 30 days of delivery.
We are committed to providing each customer with the highest standard of customer service. Please carefully check the book’s description and condition before ordering. If you have any questions or issues, please contact us first.
Thank you for choosing BooksRun!
Specialty
Сollege textbooks and trade booksSeller's business information
AZ Texts LLC
228 Park Ave S Suite 38827
New York, NY U.S.A. 10003
Terms of sale
30 days hassle-free returns guaranteed!