@N01ennn: Microsoft ran its graph system against vector RAG on 8k, 120k, and a full million token context window, and the million…

X AI KOLs Timeline Tools

Summary

Microsoft Research's LazyGraphRAG outperformed vector RAG on data-local questions across 8k, 120k, and million-token contexts, winning 92/90/91% at a tenth of the cost, and is now open-sourced on GitHub.

Microsoft ran its graph system against vector RAG on 8k, 120k, and a full million token context window, and the million token giant lost Jonathan Larson from Microsoft Research put up the numbers → LazyGraphRAG won 92, 90, and 91 percent of data-local questions, the exact place plain RAG was supposed to be strong → and did it at a tenth of the cost of the million token run the lesson buried in the benchmark → a bigger context window is not memory, structured memory is throwing more tokens at the problem was just the expensive way to be wrong open source on GitHub, the lesson a $1500 course would charge you for
Original Article
View Cached Full Text

Cached at: 08/03/26, 01:43 PM

Microsoft ran its graph system against vector RAG on 8k, 120k, and a full million token context window, and the million token giant lost

Jonathan Larson from Microsoft Research put up the numbers → LazyGraphRAG won 92, 90, and 91 percent of data-local questions, the exact place plain RAG was supposed to be strong → and did it at a tenth of the cost of the million token run

the lesson buried in the benchmark → a bigger context window is not memory, structured memory is

throwing more tokens at the problem was just the expensive way to be wrong

open source on GitHub, the lesson a $1500 course would charge you for

Similar Articles

RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM

Hugging Face Daily Papers

RAGU is an open-source multi-step GraphRAG engine that uses a compact 7B fine-tuned LLM (Meno-Lite-0.1) to achieve high-quality knowledge graph construction at a fraction of the cost of larger models, outperforming larger systems on benchmarks.

ScalableRAG: High-Quality RAG at Zero Ingestion Cost

arXiv cs.AI

This paper introduces ScalableRAG, a retrieval-augmented generation method that achieves high accuracy without any ingestion costs (no vector database or knowledge graph) by using regex-based set creation and aggregative reasoning. It outperforms baselines on multiple datasets and also presents a limited-ingestion variant for further accuracy improvements.