pipelinesIntermediate

Data Engineering

Build production pipelines: ingest with Auto Loader, COPY INTO and Lakeflow Connect, transform with PySpark and SQL, orchestrate with Lakeflow Jobs, and ship it with bundles.

0/52|52 written|~60 h
To doIn progressDonePlannedSwipe the map, tap a concept

The longest path on the site and the one that maps almost one-to-one onto the Data Engineer Associate exam. Work it in order: every stage assumes the one before it.

If you have never opened a Databricks workspace, do Lakehouse Foundations first.

Resources for this path

Courses, books and repos that cover the whole map go here. None have been added yet; per-concept resources appear on each concept page.