← Back to courses
DatabricksIntermediateUpcoming
Databricks Applied
Data Pipelines and Streaming Workloads
8 hours28 lessons Self-study
EE
Emmanuel Edegbo
Lead Data Engineer & Architect
About this course
Build Delta Live Tables pipelines, streaming ingestion, and orchestrated workflows for production data platforms.
Databricks is where enterprise data engineering has consolidated — the Lakehouse pattern, Delta tables, and Unity Catalog are now standard vocabulary in data-platform job specs. Engineers who can build governed, tested pipelines on it, not just run notebooks, are the ones teams hire.
Who this course is for
- ✓Data engineers and analysts moving from SQL/Python fundamentals into production pipeline work
- ✓Professionals whose employers are adopting Databricks and who need real project experience, not clicked-through tutorials
- ✓Career switchers targeting data engineering roles who want one continuous, portfolio-grade project
- ✓SQL and Python track graduates ready to combine both skills on the Lakehouse
Who this course is NOT for
- ✗Complete beginners to data — start with the SQL or Python track first; this track assumes those fundamentals
- ✗Experienced Databricks engineers — Core builds the foundations; join at Applied or Professional
- ✗Anyone after ML/AI model training — this track is data engineering: pipelines, quality, and governance
How you'll learn
- →One continuous production pipeline built across the whole track — every chapter advances the same real project
- →You work in a provisioned environment: your own workspace access, personal Unity Catalog, and Azure DevOps repo
- →Changes ship the way real teams ship them — branch, commit, pull request, review, merge; the repo enforces it
- →Checkpoint states let you verify your build against a known-good reference at the end of each level
- →Course discussion lets you ask the instructor and other learners questions inline with each lesson
By the end of this course, you'll be able to
- ✓Build contract-driven Landing→Bronze→Silver→Gold pipelines on Delta tables
- ✓Apply data-quality validation, quarantine, and audit patterns that survive production
- ✓Work with Unity Catalog governance — catalogs, schemas, and permissions — as a daily tool
- ✓Use Git and pull requests on Azure DevOps the way professional data teams do
- ✓Walk into an interview with a continuous, reviewable pipeline project in your own repo
Datasets used
| SalesPY | Retail dataset loaded as Delta tables for transformation and aggregation work. |
| FinancePY | Compliance dataset for monitoring and aggregation pipelines on the Lakehouse. |
Tools you'll need
- •Databricks Free Edition workspace — compute is provided during the course; upgrade to a paid tier if you want to keep the workspace for post-course practice
- •Any modern web browser — the workspace UI, notebooks and Delta Live Tables all run in-browser; nothing to install locally
What you get when you enrol
- ✓Lifetime access to every lesson, exercise, and update — including future revisions to this course.
- ✓12-month Azure SQL practice access against the same datasets used in the course (read-only). Renews on request for active learners.
- ✓Auto-graded labs in your browser — write SQL, hit Run, get instant feedback against the expected result.
- ✓AI-graded Professional Challenges — open-ended scenarios reviewed against a published rubric, not just a single right answer.
- ✓Course discussion + community — talk to other learners and ask the instructor questions inside the course.
- ✓Basic Certificate on demonstrated capability — awarded when you complete every Hands-On Lab and Module Readiness Check, plus the Professional Challenges. Confirms you can write, run, and defend course-level SQL against real datasets.
- ✓Optional Advanced Certificate on completion of the Databricks Applied multi-project — a separate credential awarded when you complete all three capstone projects, each independently assessed and approved by an instructor. Each project is end-to-end work against a real brief with defined acceptance criteria — proves competence at a level an employer can actually evaluate. The Basic Certificate alone confirms course mastery; the Advanced Certificate confirms you can deliver.
- ✓Optional live training upgrade — instructor-led cohort sessions with capped capacity, sold separately.
What you'll learn
- ✓Build Delta Live Tables (DLT) pipelines with expectations
- ✓Implement streaming ingestion with Auto Loader and Structured Streaming
- ✓Orchestrate multi-task workflows with dependencies
- ✓Apply SCD Type 1 and Type 2 patterns in DLT
- ✓Monitor pipeline health and implement quarantine patterns
- ✓Complete a production data pipeline project
Who this is for
Software testersAdvanced data analystsData engineersData scientists
Prerequisites
- •Databricks Core or equivalent lakehouse foundation
What learners say
How ratings work4.7
38 ratings (time-weighted)
- 5★27
- 4★11
- 3★0
- 2★0
- 1★0
Course discussion
Open to enrolled learnersSign in to read and post in the course discussion.
Sign inContinue on the Databricks Intermediate track
Pair this with the matching format to build skills, evidence and accountability together.
Live Training cohorts
Instructor-led. Pick a date that works.
No Databricks Intermediate live cohort scheduled yet.
See upcoming live sessionsPortfolio projects
Real-data evidence employers can see.
No Databricks Intermediate portfolio project published yet.
Browse all portfolio projects