@N01ennn: Microsoft ran its graph system against vector RAG on 8k, 120k, and a full million token context window, and the million…
Summary
Microsoft Research's LazyGraphRAG outperformed vector RAG on data-local questions across 8k, 120k, and million-token contexts, winning 92/90/91% at a tenth of the cost, and is now open-sourced on GitHub.
View Cached Full Text
Cached at: 08/03/26, 01:43 PM
Microsoft ran its graph system against vector RAG on 8k, 120k, and a full million token context window, and the million token giant lost
Jonathan Larson from Microsoft Research put up the numbers → LazyGraphRAG won 92, 90, and 91 percent of data-local questions, the exact place plain RAG was supposed to be strong → and did it at a tenth of the cost of the million token run
the lesson buried in the benchmark → a bigger context window is not memory, structured memory is
throwing more tokens at the problem was just the expensive way to be wrong
open source on GitHub, the lesson a $1500 course would charge you for
Similar Articles
@oliviscusAI: The entire RAG industry is about to get cooked. Researchers developed a new RAG approach that bypasses almost everythin…
Researchers introduced PageIndex, a novel RAG approach that eliminates reliance on vector databases, embeddings, and chunking, achieving 98.7% accuracy on FinanceBench and outperforming existing methods while being free and open source.
Why Vector RAG fails for AI coding agents at scale (And how I used a Neo4j graph to fix it)
A new open-source tool called Writ uses a hybrid retrieval pipeline with BM25, ONNX vectors, and Neo4j graph traversals to provide context rules for AI coding agents, reducing token bloat by 726x and enforcing plan approval via bash hooks.
We open-sourced a graph-free multi-hop RAG framework — matches Graph-RAG accuracy without the rebuild cost (Apache-2.0)
MOTHRAG is a graph-free multi-hop RAG framework that matches the accuracy of graph-based systems like GraphRAG and HippoRAG on benchmarks, while avoiding costly graph rebuilds by using a dense index and query-time orchestration.
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM
RAGU is an open-source multi-step GraphRAG engine that uses a compact 7B fine-tuned LLM (Meno-Lite-0.1) to achieve high-quality knowledge graph construction at a fraction of the cost of larger models, outperforming larger systems on benchmarks.
ScalableRAG: High-Quality RAG at Zero Ingestion Cost
This paper introduces ScalableRAG, a retrieval-augmented generation method that achieves high accuracy without any ingestion costs (no vector database or knowledge graph) by using regex-based set creation and aggregative reasoning. It outperforms baselines on multiple datasets and also presents a limited-ingestion variant for further accuracy improvements.