auto-research

Tag

Cards List
#auto-research

SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness

Hugging Face Daily Papers · yesterday Cached

SoL-Pi introduces a method for recursively scaling auto-research loops in coding agents, achieving significant token and cost reductions while maintaining performance on benchmarks.

0 favorites 0 likes
#auto-research

Agora: Git as Shared Memory for Collective AutoResearch

Hugging Face Daily Papers · 2d ago Cached

Agora is a shared memory system for autonomous AI research agents that uses Git to record research as an append-only DAG, enabling collaborative discovery. In a 12-day experiment, 13 agents improved a model's performance by 62% towards a trained baseline.

0 favorites 0 likes
#auto-research

LabAgent: Customize Any Research Hubs for Scientific Discoveries Using AI Agents

arXiv cs.AI · 3d ago Cached

LabAgent is an AI agent system designed to customize research hubs for scientific discoveries, enabling reproducible and continuous laboratory work across various biological domains and outperforming commercial generalist agents.

0 favorites 0 likes
#auto-research

Pi Agent Users - Nvidia Released Sol-Pi - A Pi-Extension based on AutoResearch loops to make the Harness more efficient

Reddit r/LocalLLaMA · 2026-09-10

SoL-Pi is a standalone extension for Pi agents that enhances efficiency by reducing token traffic and inference work through mechanisms like action fusion and context compaction, with all features being opt-in and preserving original evidence.

0 favorites 0 likes
#auto-research

@kirbytheodor: deslopped the tui a bit so it's actually useful at a glance ouroboros with fable and astra deleted 2.79 million lines o…

X AI KOLs Timeline · 2026-09-07 Cached

Enhanced the TUI of ouroboros, a CLI tool for self-evolving auto-research systems, and deleted 2.79 million lines of vestigial code to improve usability.

0 favorites 0 likes
#auto-research

Personalized Auto-Research: Towards a True AI Co-Scientist

arXiv cs.AI · 2026-08-18 Cached

The paper introduces a framework for personalized auto-research systems that condition every stage of the research process on individual scientist representations, arguing that personalization is essential for AI to serve as true co-scientists rather than generic instruments.

0 favorites 0 likes
#auto-research

Auto-research with codex: How I achieved a 232x Faster Kernel

Hacker News Top · 2026-08-15 Cached

A blog post detailing how the author used Codex to optimize a kernel in a GPU Mode contest, achieving a 232x speedup in QR decomposition and sharing learnings on auto-research.

0 favorites 0 likes
#auto-research

ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs

arXiv cs.AI · 2026-08-14 Cached

This paper introduces ARAC-Bench, a benchmark for evaluating the alignment, logical coherence, and completeness of Auto-Research systems' research processes against human methodology. Experiments on 11 state-of-the-art frameworks show a best alignment score of only 67.9/100, highlighting a significant gap in simulating rigorous human research behavior.

0 favorites 0 likes
#auto-research

@JinjingLiang: Hosting an auto-research hackathon this Saturday in SF It’s in a mansion (not mine). DM me or just comment if you wanna…

X AI KOLs Following · 2026-07-15 Cached

Hosting an auto-research hackathon this Saturday in SF at a mansion. Interested participants can DM or comment to join.

0 favorites 0 likes
#auto-research

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

Hugging Face Daily Papers · 2026-07-15 Cached

GigaWorld-Policy-0.5 is an enhanced World Action Model for robot control that improves training and inference efficiency through a Mixed Action-Conditioned World Modeling strategy and a Mixture-of-Transformers architecture, achieving 85ms latency on a local RTX 4090.

0 favorites 0 likes
#auto-research

@nasqret: I've been doing a lot of experiments with auto-research in the last few weeks, especially in algebra. Here are a couple…

X AI KOLs Timeline · 2026-07-10 Cached

The author shares observations from auto-research experiments in algebra, noting that AI models can generate code and discover novel abstract rules, leading to potentially alien mathematics that humans struggle to understand.

0 favorites 0 likes
#auto-research

@shiposcant: completed reading this: "the better you know something, the better you can prompt the LLMs, because you convert unknown…

X AI KOLs Timeline · 2026-07-09 Cached

A blog post describes using Codex to automatically iterate and optimize GPU kernels, achieving a 212x speedup over baseline. The post highlights how expertise amplifies AI's utility, turning unknown unknowns into known unknowns through a looped experimental workflow.

0 favorites 0 likes
#auto-research

@dejavucoder: my latest blog post "auto-research with codex: how I achieved a 212x faster kernel over baseline with codex in GPU Mode…

X AI KOLs Timeline · 2026-07-08 Cached

Blog post by Sankalp detailing how he used Codex to achieve a 232x faster GPU kernel for QR decomposition in GPU Mode's contest, outlining his auto-research methodology.

0 favorites 0 likes
#auto-research

@askalphaxiv: Introducing autoresearch for GitHub repos Change 'Github' to 'ARGithub' in any repo URL Research artifacts extend beyon…

X AI KOLs Timeline · 2026-06-24 Cached

Introduce a tool that by changing 'Github' to 'ARGithub' in any repo URL, deploys an agent to orient itself on the codebase, resolve setup issues, and run experiments.

0 favorites 0 likes
#auto-research

@teortaxesTex: Deli open sources his AutoResearch.

X AI KOLs Timeline · 2026-06-17 Cached

Deli Chen open sources his AutoResearch SKILL tool and releases a survey paper on Self-play, inspired by AlphaZero.

0 favorites 0 likes
#auto-research

SIQ-1 Qwen3.6 for autoresearch and autonomous agency

Reddit r/LocalLLaMA · 2026-06-17

SIQ-1 Qwen3.6 is a new AI model designed for automated research and autonomous agency tasks, extending the Qwen family with enhanced agentic capabilities.

0 favorites 0 likes
#auto-research

PseudoBench: Measuring How Agentic Auto-Research Fuels Pseudoscience

arXiv cs.AI · 2026-06-17 Cached

PseudoBench is a benchmark to evaluate whether LLM-based agentic auto-research systems can resist pseudoscientific narratives. Testing seven state-of-the-art agents reveals they readily produce persuasive pseudoscientific reports with near-zero refusal rates, calling for scientific alignment before deployment.

0 favorites 0 likes
#auto-research

@DrJimFan: Today, we enable AutoResearch in the physical world for the first time! Introducing ENPIRE: we give 8 Codex agents a fl…

X AI KOLs Following · 2026-06-16 Cached

NVIDIA GEAR lab introduces ENPIRE, a system that uses 8 Codex agents to autonomously control a robot fleet for physical tasks like tying zip-ties and installing GPUs, demonstrating self-improving robotics research and a new 'physical scaling' phenomenon.

0 favorites 0 likes
#auto-research

@DanKornas: If you’re trying to follow AI agents for research, the hard part is not one paper — it’s the whole lifecycle. Awesome A…

X AI KOLs Timeline · 2026-06-14 Cached

A curated GitHub resource that maps AI-assisted scientific research tools and papers across the full research lifecycle, from idea generation to dissemination.

0 favorites 0 likes
#auto-research

@mylifcc: The ceiling of Auto-Research infrastructure has arrived! Yacine's 1.5-hour in-depth interview with the two founders of Paradigma, hardcore breakdown of how DAG becomes the underlying infrastructure for autonomous research: • Why DAG is the best substrate for research (far beyond linear papers) • Ag…

X AI KOLs Timeline · 2026-05-26 Cached

Yacine conducted a 1.5-hour in-depth interview with the founders of Paradigma, discussing how to use DAG (Directed Acyclic Graph) as the underlying infrastructure for autonomous research, covering core topics such as Agent operation, building large-scale public DAGs, and avoiding bad DAGs.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback