The other modules teach the platform from the documentation. This one is different: it's synthesized from ten Data + AI Summit sessions (~6.5 hours) given by the engineers and PMs who build the products — Lakeflow, Auto Loader, Unity Catalog managed tables, declarative pipelines, Spark 4.1, real-time mode — plus two production war stories (a trading-dashboard build and a 10–20-billion-events-per-day architecture).
What conference sessions give you that docs don't: the numbers ("listing 10M files takes 40 minutes; file events do it in one"), the honest caveats ("real-time mode is at-least-once — only adopt it where duplicates are acceptable"), and the decision framing ("stop asking what source am I reading; ask how much management vs. customization do I want"). This module is that material, organized, with a consolidated gotcha list in the cheat sheet.
Status caveat: GA / Preview / roadmap labels here are as stated at the conference —
verify current status in the docs before designing around anything not marked GA. Source
transcripts live in the shared workspace
(/Volumes/hd2_shared_docs/aide-shared/db-training-shared/de/transcripts/).
cleanSource, schema evolution vs VARIANT, and the observability recipe.SET MANAGED migration runbook with its DBR version matrix.Listen straight through once (about 34 minutes), then keep the cheat sheet — especially the
27-item gotcha table — as the pre-design-review checklist. The lab converts a real table with
SET MANAGED and proves the file-events claim on your own bucket.