• Open Daily: 10am - 10pm
    Alley-side Pickup: 10am - 7pm

    3038 Hennepin Ave Minneapolis, MN
    612-822-4611

Open Daily: 10am - 10pm | Alley-side Pickup: 10am - 7pm
3038 Hennepin Ave Minneapolis, MN
612-822-4611
Spark and Big Data Processing with R and Python: Building Scalable Pipelines with Sparklyr, PySpark, and Databricks

Spark and Big Data Processing with R and Python: Building Scalable Pipelines with Sparklyr, PySpark, and Databricks

Paperback

Programming

Currently unavailable to order

ISBN13: 9798198822771
Publisher: Independently Published
Pages: 424
Weight: 1.12
Height: 1.05 Width: 6.00 Depth: 9.00
Language: English
Reactive Publishing

Learn big data processing with Spark using R and Python.

This book offers a practical introduction to Apache Spark for large-scale data work. It focuses on building scalable data pipelines using Sparklyr, PySpark, and Databricks.

You will discover how to:

  • Work with large datasets using Apache Spark
  • Use Sparklyr to integrate Spark with R
  • Develop with PySpark in Python
  • Leverage Databricks for cloud-based Spark environments
  • Construct reliable data pipelines for real-world use

The book bridges the R and Python ecosystems, helping data professionals use the right language for different Spark tasks. It covers core concepts including data transformation, distributed processing, and moving projects from development to production.

Also from

Schwartz, Alice

Also in

Programming