Pipeline · easy · ~10 min
A backfill kicked off 200 parallel load_dim task instances that each write to the shared Postgres warehouse. The warehouse had a brief 30-second blip. Instead of recovering, everything got *worse*: the database is now pinned at max connections, tasks are failing en masse, and even unrelated pipelines can't reach the DB.
load_dim
The blip is long over. Something about how the tasks retry turned a 30-second hiccup into an ongoing outage. Find it.