Iceberg, Hudi, Delta, and Paimon at company scale — compaction ops, format wars, catalogs, and the S3-native future.
The freshest slice of the corpus: Adobe's Iceberg arc, Ancestry's 100-billion-row table, Kakao's operational takeaways, Whoop's schema-migration tooling — plus the deepest expert vein in the archive (Vanlightly's consistency-model series). This is also the natural partner-track cohort.
Compaction, small files, snapshot hygiene, schema migrations, and MERGE economics — the maintenance loop that open formats hand to YOU.
Your team lands a CDC stream into an Iceberg table, committing every minute. The demo is flawless. Three weeks into production, dashboards that used to scan the table in 40 seconds take four minutes, the hourly MERGE bill has doubled — and total data size has barely moved. No query changed. No schema changed. Nothing about the *data* changed. So what did?
This week is the discovery every lakehouse team makes in its first year: open table formats hand you powerful primitives and none of the scheduling. The answer to "what changed" is always in the files.
The full week 2 brief is part of LeetData Pro.