Matching the world's top multi-hop RAG systems, with no GPU, no fine-tuning, just pip install
Summary
MOTHRAG is a multi-hop RAG system that matches the performance of top GPU-dependent systems (HippoRAG 2, CoRAG, NeocorRAG) using only commodity API calls, with no GPU, no fine-tuning, and deployment via pip install plus API keys.
View Cached Full Text
Cached at: 06/20/26, 02:29 PM
Similar Articles
We open-sourced a graph-free multi-hop RAG framework — matches Graph-RAG accuracy without the rebuild cost (Apache-2.0)
MOTHRAG is a graph-free multi-hop RAG framework that matches the accuracy of graph-based systems like GraphRAG and HippoRAG on benchmarks, while avoiding costly graph rebuilds by using a dense index and query-time orchestration.
Beyond Static RAG: An Adaptive, Tri-Metric Routing Framework for Efficient Long-Context Inference on Commodity GPUs
This paper proposes a Tri-Metric Router, a deterministic framework for adaptive routing among inference pipelines to address the Compression Paradox in long-context RAG on commodity GPUs, achieving zero OOM failures and improved performance.
RAG-Stack: Co-Optimizing RAG Serving Performance and Quality
This paper introduces RAG-Stack, a framework that co-optimizes RAG serving performance and answer quality by efficiently exploring the joint algorithm-system configuration space. It finds Pareto frontiers that cover significantly more quality-performance space than existing configuration-search methods.
ScalableRAG: High-Quality RAG at Zero Ingestion Cost
This paper introduces ScalableRAG, a retrieval-augmented generation method that achieves high accuracy without any ingestion costs (no vector database or knowledge graph) by using regex-based set creation and aggregative reasoning. It outperforms baselines on multiple datasets and also presents a limited-ingestion variant for further accuracy improvements.
@yoginth: today i'm launching http://rag.computer an open source RAG platform built on top of @turbopuffer fast ingestion, fast r…
bigRAG is an open-source RAG platform built on top of Turbopuffer for fast ingestion and retrieval, supporting multiple document formats and embedding models.