Python. SQL. Spark. Power BI.
One person. A whole toolkit.
Learn data engineering through real incidents
One person. A whole toolkit.
and make it flow.
Billions of events, shaped into tables a company can trust.
Your delivery ETA. The CEO's dashboard. The price you pay.
We’re building the place where anyone can become one.
Nobody sees the plumbing. Everything depends on it.
Season 1 · Aankra
Four cases are live. Start with Case 01: sixteen minutes, no setup.
The stack you'll touch in Season 1
Season 1 · Case files
The company scales from a thousand orders a day to millions. Your pipeline grows with it, and so do the ways it breaks. Start with Case 01, then work through the rest in order.
Why the database that runs your app is terrible at answering questions about it — shown byte by byte.
Watch a late order break a live dashboard, and the batch-vs-streaming argument changes shape.
The job finished. Nothing errored. The numbers are wrong. Stop it before anyone sees it.
One customer moves house, and ₹2,000 of last year’s revenue moves with her.
Follow one wrong figure backwards through every table it touched, in minutes rather than all day.
Drops soon · Notify meTraffic grows overnight. The job that took 4 minutes now takes 9 hours.
Drops soon · Notify meGet paged
Cases 01 to 04 are live now. Leave your email and we’ll send one short note the day each new case goes live. No spam, unsubscribe in a click.
How it works
Every episode drops you into a real pipeline failure, the kind that wakes engineers at 3 AM. You investigate, make the call, and walk away with the concept and the fix.
A dashboard shows zero revenue on a Friday night. Something upstream broke. The clock is running.
Trace the data backwards through serving, warehouse, transformation and ingestion. Pick your suspect.
The reveal explains the concept behind the failure, and you keep a one-page playbook you'll use at work.
Who it's for
You've heard Kafka, Spark and Airflow a hundred times. Here you'll see where each one actually sits.
You know SQL. Now learn the pipelines behind the tables you query, and the failures nobody tells you about.
The incidents are real. Use them as a playbook for postmortems, design reviews and interviews.
The platform
Episodes teach you how to think. The modules make you practise, in the browser, with no setup.
Cinematic episodes built on real incidents, with sources for every number.
LiveSELECT city, count(*) AS orders
FROM deliveries
GROUP BY city;Interview-grade SQL problems with instant feedback and query plans explained in plain English.
In builddf.groupBy("city")
.agg(F.count("*"))Write Spark in the browser, watch the shuffle happen, and see why the slow version is slow.
In buildAsk anything mid-episode. It knows the case you're on and answers with your own numbers.
In buildNo. The cases don’t make you write code. Basic SQL helps you get more out of them, and the SQL module is built to get you there.
The episodes are free to read. The hands-on modules will have a paid tier when they launch.
One per week during season 1. If you leave your email, we ping you when the next case is live on the site.
They're based on failures that happen in real production systems, rewritten with fictional companies. Every statistic carries its source.
Yes. The episodes are built mobile-first, with headphones recommended for the sound design.
Still deciding?