Kafka · hard · ~13 min
Your ads pipeline attributes orders to ad clicks: a Flink attribution job consumes committed orders, looks up matching ad events in a document store, and emits attribution records — each with a UUID — into Kafka. The OLAP store dedupes on that UUID via upsert. Kafka transactions + checkpoint two-phase commit protect the pipeline hops. The design doc says exactly-once, end to end.
Last Tuesday the attribution job was down for 4 hours (a dependency outage). The on-call rewound its consumer offset and replayed the window — standard runbook.
Today, finance reports advertiser invoices for Tuesday are ~6% high, concentrated in the replayed window. The pipeline dashboards show zero errors and exactly-once intact.