Building Scalable Data Systems with Apache Spark 4.x
Language: English
Published by Independently Published, 2026
- Softcover
- New

Seller: PBShop.store US, Wood Dale, IL, U.S.A.PBShop.store US
AbeBooks seller since April 7, 2005
Condition: New
US$ 32.41
Quantity: Over 20 available
Add to basketItem description from seller
Seller Inventory # L2-9798195327088
- Title
- Building Scalable Data Systems with Apache Spark 4.x
- Author
- Kevin R Auguste
- Publisher
- Independently Published
- Publication year
- 2026
- Condition
- New
- Binding
- PAP
- Language
- English
- ISBN 13
- 9798195327088
- Item weight
- 426 grams
Are your data pipelines slowing down, breaking under scale, or becoming too complex to maintain?
Modern data systems demand more than scripts that “just work.” They require reliability, performance, and the ability to evolve without constant rewrites. Yet many engineers and analysts struggle with inefficient Spark jobs, unpredictable execution, and rising infrastructure costs.
This book addresses that gap.
Building Scalable Data Systems with Apache Spark 4.x is a practical guide to designing, optimizing, and operating distributed data pipelines using Apache Spark, PySpark, SQL, and lakehouse technologies. It focuses on how Spark actually behaves at scale, so you can build systems that are not only functional, but fast, stable, and production-ready.
You won’t just learn how to write Spark code, you’ll learn how to think like a data systems engineer.
Inside, you will learn how to:
- Design end-to-end pipelines from ingestion to output using PySpark and SQL
- Understand execution internals like DAGs, jobs, stages, and Catalyst optimization
- Optimize performance through partitioning, Adaptive Query Execution (AQE), and efficient joins
- Build reliable streaming systems with Structured Streaming and exactly-once semantics
- Work with modern storage systems like Delta Lake and Apache Iceberg
- Deploy and operate Spark workloads using Kubernetes, monitoring, and resource tuning
Each chapter builds practical intuition, connecting code to execution so you can diagnose bottlenecks, reduce cost, and scale confidently.
If you work as a data engineer, data analyst, backend developer, or data scientist, this book equips you with the skills to move beyond trial-and-error and build systems that perform consistently in real-world environments.
Your data is growing. Your systems should keep up.
Get your copy today and start building data pipelines that scale, perform, and last.
"Synopsis" may belong to another edition of this title.
PBShop.store US
Wood Dale, IL, U.S.A.
AbeBooks seller since April 7, 2005
Shipping rates within U.S.A.
| Item | 7 to 14 business days | 7 to 14 business days |
|---|---|---|
| First item | US$ 0.00 | US$ 0.00 |
Payment methods
Store description
Specialty
Hardbacks, PaperbacksSeller's business information
Pbshop.co.uk Ltd
Unit 22 Horcott Industrial Estate, Horcott Road
Fairford, United Kingdom GL7 4BX
Terms of sale
Returns Policy
We ask all customers to contact us for authorisation should they wish to return their order. Orders returned without authorisation may not be credited.
If you wish to return, please contact us within 14 days of receiving your order to obtain authorisation.
Returns requested beyond this time will not be authorised.
Our team will provide full instructions on how to return your order and once received our returns department will process your refund.
Please note the cost to return any unwanted order to us is borne by the buyer.
Should your order arrive damaged, not as advertised or faulty, we must be advised of this within 14 days of delivery. Please contact us so we can find the best solution for you.
Our Customer Care Team can be contacted via emailing paperback-us@paperbackshop.co.uk, or by calling our UK Office on +441285 712 917. We are available 9:00am till 5:30pm GMT weekdays and 9:00am till 1:00pm GMT on Saturdays.
Shipping terms
Books are shipped from UK warehouse. Delivery thereafter is between 4 and 14 business days dependant upon your location - please do contact us with any queries you may have.