Spark · easy · ~8 min
A daily report reads one day from a year-partitioned Parquet dataset. It used to scan ~45GB; since a 'timezone fix' refactor it scans 16.4TB — the whole table. The query still returns correct results, so nobody noticed for a week. The Spark UI's scan node tells the story.