11.67% ARC-AGI-2 Local Eval on a Single 4090: The TOPAS Recursive Architecture
Summary
The authors present TOPAS, a recursive AI architecture achieving 11.67% on ARC-AGI-2 using a single RTX 4090, aiming to demonstrate that architectural efficiency can outweigh raw compute power.
Similar Articles
Cost-Effective Agent Harnesses for Abstract Reasoning and Generalization on ARC-AGI-1
This paper presents cost-effective agent harnesses for ARC-AGI-1 that achieve strong performance using DeepSeek V3.2 without fine-tuning, via an Explorer-Definer Pipeline and a Reflective Orchestrator, achieving 67.25% pass@2 at low cost.
@rohanpaul_ai: GLM-5.2 got 22.8% on ARC-AGI-2:, $0.25/task To note here, around May 2025, the best verified models on ARC-AGI-2 were o…
GLM-5.2 achieves 22.8% on ARC-AGI-2 and 77% on ARC-AGI-1 at a low cost of $0.25 per task, representing a 7.6x improvement over the best frontier score from May 2025.
Did Pathway just reveal the architecture breakthrough Andrew Curran predicted? Its 150M model sets a new ARC-AGI-1 cost-efficiency frontier
Pathway's 150M-parameter BDH-CQ model achieves 29.5% on ARC-AGI-1 at a record-low cost of $0.0007 per task, using recurrent memory and latent reasoning instead of long token chains. The architecture may be the breakthrough Andrew Curran teased, with OpenAI researcher Lukasz Kaiser as an investor and adviser.
ARC-AGI Leaderboard
The ARC-AGI leaderboard shows model performance on three versions of the benchmark, measuring fluid intelligence and efficient adaptation, with trend lines for reasoning systems and raw LLMs.
Sakana Fugu (3 minute read)
Sakana AI introduces AB-MCTS, an inference-time scaling algorithm that enables multiple frontier AI models (Gemini 2.5 Pro, o4-mini, DeepSeek-R1-0528) to cooperate, significantly outperforming individual models on the ARC-AGI-2 benchmark.