New Course - Limited Seats

Production Python +
PySpark + Databricks

Go from Zero to Production-Ready Data Engineer in 12 Weeks

12 Weeks
48 Sessions
Mon - Thu, 1 Hour/Day
4 Projects
Enrollment is closed for this cohort

The full syllabus below stays published so you can judge the depth before the next batch opens. If you want to start learning now, the Agentic AI Portfolio course is currently enrolling.

Every Session, Same Format

Consistent structure for maximum learning efficiency

0-3 min

settle In

Quick recap + 1 common mistake from homework

3-13 min

concept

Real-world scenario with 3-5 slides max

13-50 min

live Build

Code the solution together (37 min hands-on)

50-57 min

gotchas

What breaks in real production systems

57-60 min

homework

Task that builds toward your capstone

Your Starting Point Matters

Jump in at the right week based on your experience

W1

Start from

Week 1

No programming experience

W5

Start from

Week 5

Python developer, new to data

W7

Start from

Week 7

Data analyst with SQL + Python

W11

Start from

Week 11

Experienced engineer, needs Databricks

10 Modules, 12 weeks

From Python fundamentals to production-ready Databricks pipelines

Module 1
2 weeks
8 sessions
Python Fundamentals
Build a rock-solid Python foundation with core programming concepts, data structures, and file handling

Week 1: Core Python Essentials

  • Environment Setup: Terminal, Python, VS Code & Virtual Environments
  • Variables, Data Types, Strings & Type Conversion
  • Lists, Loops & Conditional Logic
  • Dictionaries, Sets & Tuples

Week 2: Intermediate Python

  • Functions, Arguments, Scope & Modules
  • File I/O: CSV, JSON & Text Files
  • Comprehensions, Generators & Lambda Functions
  • Error Handling, Custom Exceptions & Logging

4 Portfolio-Ready Projects

Build real production systems that demonstrate your skills

1
Week 3-4
Data Ingestion CLI Tool

Python, Pandas, APIs, Git, error handling

Portfolio-ready deliverable
2
Week 6
Production Python Pipeline

Testing, config, logging, retries, project structure

Portfolio-ready deliverable
3
Week 10
PySpark Analytics Pipeline

Spark transforms, joins, windows, performance, testing

Portfolio-ready deliverable
4
Week 11-12
Capstone: Production Data Platform

Delta, Databricks, medallion, CI/CD, governance

Portfolio-ready deliverable

12-Week Learning Path

Your weekly topic breakdown at a glance

WeekMonTueWedThu
W1
Terminal & SetupVariables & TypesLists & LoopsDicts & Sets
W2
FunctionsFile I/OComprehensionsError Handling
W3
OOP & DataclassesGit BasicsGit WorkflowsPackages
W4
Pandas BasicsPandas AdvancedREST APIsAsync Python
W5
SQL FundamentalsAdvanced SQLData ModelingDimensional Design
W6
Pydantic & TypesConfig & LoggingPytestProject Structure
W7
Cloud StorageDockerResilience PatternsData Quality
W8
Spark IntroDataFrame OpsAggregationsSchemas & Formats
W9
JoinsWindowsComplex TypesSpark SQL
W10
Spark InternalsPerformance TuningCaching LabStreaming
W11
DatabricksDelta LakeMERGE PatternsOptimization
W12
Unity CatalogAuto LoaderMedallionCapstone

Launch Your Data Engineering Career

Graduate with the skills top companies are hiring for

Data Engineer
$120,000 - $180,000
FAANG
Startups
Finance
Healthcare
Analytics Engineer
$110,000 - $160,000
Tech companies
Consulting
E-commerce
Platform Engineer
$130,000 - $190,000
Cloud providers
Data platforms
Enterprise

Ready to Become a Data Engineer?

48 live sessions, 4 portfolio projects, zero to production-ready data engineering in 12 weeks.

Enrollment is closed for this cohort

We keep cohorts in sync rather than letting people join a half-finished syllabus. The next batch has not been dated yet.

Questions? Contact us at hello@thrivewithai.com