Skip to content
Home/ Spark and Big Data Processing with R and Python: Building Scalable Pipelines with Sparklyr, PySpark,
Spark and Big Data Processing with R and Python: Building Scalable Pipelines with Sparklyr, PySpark,

Spark and Big Data Processing with R and Python: Building Scalable Pipelines with Sparklyr, PySpark,

No customer reviews yet ISBN 9798198822771

Reactive Publishing

Learn big data processing with Spark using R and Python.

This book offers a practical introduction to Apache Spark for large-scale data work. It focuses on building scalable data pipelines using Sparklyr, PySpark, and Databricks.

You will discover how to:

  • Work with large datasets using Apache Spark
  • Use Sparklyr to integrate Spark with R
  • Develop with PySpark in Python
  • Leverage Databricks for cloud-based Spark environments
  • Construct reliable data pipelines for real-world use

The book bridges the R and Python ecosystems, helping data professionals use the right language for different Spark tasks. It covers core concepts including data transformation, distributed processing, and moving projects from development to production.

Written for data analysts, data scientists, and engineers who want to add Spark to their toolkit.

Perfect for: Professionals looking to expand their skills in big data processing with Spark across both R and Python.

About the author

Product details

Pub dateMay 27, 2026
ISBN-109798198822771
ISBN-139798198822771
LanguageEnglish
Last updated 2026-05-28 19:16
$34.95 $39.99 12% off
You save $5.04 · list price $39.99
In stock — ships in 24 hours with free tracking
Delivery by Monday, September 14, 2026
Qty
Sign in to Add to Saved list
Free delivery on orders over $35.
15-day returns. Any reason.
Secure checkout. We never store card details.