Skip to content
All videos
releasesDatabricks·July 27, 2023

Nebula: The Journey of Scaling Instacart’s Data Pipelines with Apache Spark™ and Lakehouse

Summary

Instacart migrated its large-scale advertising and data pipelines from a traditional data warehouse to a Databricks lake house powered by Apache Spark and Delta Lake. The presenters demonstrate how this transition unified batch and streaming data processing, reduced storage and compute costs for over 40 petabytes of data, and improved developer productivity through Scala-based modular pipelines and local unit testing.

Summary generated by brickster.ai from the video transcript.