research-automation

Tag

Cards List
#research-automation

Recursive self-improvement of AI research agents

Hugging Face Daily Papers ↗ · 4d ago Cached

This paper introduces AIDE^2, a system that enables AI research agents to autonomously improve their own code through recursive self-improvement, leading to performance gains across various AI research tasks.

0 favorites 0 likes
#research-automation

@Jiacheng_Miao: Excited to share that #paper2agent is published in @Nature today! Papers have long been the primary format for communic…

X AI KOLs Timeline ↗ · 2026-09-16 Cached

Paper2Agent, a multi-agent framework published in Nature, automatically transforms research papers into virtual authors, enabling AI agents to interact with and build upon scientific knowledge for enhanced discovery.

0 favorites 0 likes
#research-automation

@andy_matuschak: Playing with programmable highlighters, e.g. * green highlight on a citation -> finds and prints the cited paper * oran…

X AI KOLs Following ↗ · 2026-09-15 Cached

Andy Matuschak shares an experiment with programmable highlighters that automate citation finding, fact-checking, and interaction with AI agents.

0 favorites 0 likes
#research-automation

The lesson from AlphaFold for mathematicians

Reddit r/singularity ↗ · 2026-09-11

The article reflects on how AlphaFold transformed protein structure prediction and speculates that similar AI advancements could revolutionize mathematics, shifting human roles towards interpretation and verification while potentially automating key tasks.

0 favorites 0 likes
#research-automation

OpenAI: AI agents now perform 3.1 researcher-workdays for every human researcher-workday, says it has reached “automated research intern” level, and expects “automated AI researcher” by March 2028

Reddit r/singularity ↗ · 2026-09-06

OpenAI claims its AI agents now contribute 3.1 workdays for every human researcher workday, achieving an "automated research intern" level, and predicts it will reach an "automated AI researcher" by March 2028. The company says agentic systems have accelerated progress toward recursive self-improvement, effectively scaling its research workforce from roughly 1,000 to an equivalent of over 4,000.

0 favorites 0 likes
#research-automation

Leveraging Large Language Models for Systematic Literature Review of Disease Spread Models

arXiv cs.AI ↗ · 2026-08-28 Cached

This paper develops an LLM pipeline for automating systematic literature reviews in disease spread modeling, comparing the performance of GPT-4.1 and GPT-5.0 against human-conducted reviews.

0 favorites 0 likes
#research-automation

If the weights never change, is it really recursive self-improvement?

Reddit r/LocalLLaMA ↗ · 2026-08-18

The article questions whether a system with persistent memory but fixed model weights, like AQuA, qualifies as recursive self-improvement, referencing a paper that uses a narrower definition.

0 favorites 0 likes
#research-automation

Personalized Auto-Research: Towards a True AI Co-Scientist

arXiv cs.AI ↗ · 2026-08-18 Cached

The paper introduces a framework for personalized auto-research systems that condition every stage of the research process on individual scientist representations, arguing that personalization is essential for AI to serve as true co-scientists rather than generic instruments.

0 favorites 0 likes
#research-automation

@akshay_pachaar: Jeff Dean's new company already has competition. Jeff Dean's Discovery Loop targets one key bottleneck with research to…

X AI KOLs Following ↗ · 2026-08-10 Cached

Akshay Pachaar introduces Primus, an autonomous AI researcher from Transformer Lab that automates the full research loop from literature review to paper writing, demonstrating its self-correcting capabilities in experiments.

0 favorites 0 likes
#research-automation

Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review

arXiv cs.AI ↗ · 2026-08-03 Cached

This paper proposes a benchmarking protocol using automated multi-model LLM review to evaluate AI Scientist systems, comparing frameworks like Sakana AI, CycleResearcher, and Data-to-Paper, and finds that FARS benchmark papers significantly outperform other systems.

0 favorites 0 likes
#research-automation

SciForge: An AI-Native, Multimodal Workbench for Scientific Discovery

arXiv cs.AI ↗ · 2026-07-20 Cached

SciForge is an open-source, AI-native multimodal workbench for scientific discovery that integrates search, reasoning, workflow execution, and evidence governance, demonstrated through eight end-to-end use cases including gene discovery and protein design.

0 favorites 0 likes
#research-automation

@lillian_ma_: Emerging autoresearch labs worth following: @AutoScienceAI (@eliot_cowan) One of the cleanest “AI builds AI” bets: agen…

X AI KOLs Timeline ↗ · 2026-06-22 Cached

A Twitter thread highlights emerging autoresearch labs that are building AI systems to automate the full research loop, from hypothesis to experimentation.

0 favorites 0 likes
#research-automation

@_akhaliq: Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

X AI KOLs Following ↗ · 2026-06-11 Cached

This paper proposes a method for autonomous research agents using hypothesis-tree refinement to generate and test hypotheses, aiming toward generalist scientific discovery.

0 favorites 0 likes
#research-automation

AutoSci: A Memory-Centric Agentic System for the Full Scientific Research Lifecycle

arXiv cs.AI ↗ · 2026-06-01 Cached

AutoSci is a memory-centric agentic system designed to automate the full scientific research lifecycle, from literature understanding to rebuttal, using LLM-based agents with persistent memory and self-evolution capabilities.

0 favorites 0 likes
#research-automation

@Honcia13: The threshold for scientific research is being completely redefined! Previously: staying up late reading papers, repeatedly running code, writing a week-long review. Now: a single instruction is enough. The open-source AI agent Feynman compresses PhD-level research processes into fully automated execution: a single instruction can complete in-depth arXiv research, literature review, code verification …

X AI KOLs Timeline ↗ · 2026-05-28 Cached

The open-source AI agent Feynman, through the collaboration of four intelligent agents, compresses PhD-level research processes (including arXiv research, literature review, code verification) into fully automated execution, requiring only a single instruction from the user.

0 favorites 0 likes
#research-automation

@DamiDefi: Claude Code cannot read 300 files at once. So someone built a system that lets it control NotebookLM from the terminal …

X AI KOLs Timeline ↗ · 2026-05-26 Cached

A system built on Claude Code allows it to control Google's NotebookLM from the terminal, automating research by searching YouTube, uploading sources, and exporting cited answers directly into Obsidian. This workflow eliminates the need for multiple browser tabs and manual copying, with verified citation accuracy.

0 favorites 0 likes
#research-automation

@seelffff: > reads papers on arXiv autonomously > finds and checks datasets on HF Hub > writes the training script itself > genera…

X AI KOLs Timeline ↗ · 2026-05-25 Cached

Hugging Face open-sourced ml-intern, an autonomous agent that performs the entire ML post-training loop—reading papers, finding datasets, writing scripts, generating data, monitoring training, and uploading weights—achieving significant GPQA improvement with a 1.7B model in 10 hours without human intervention.

0 favorites 0 likes
#research-automation

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery

Hugging Face Daily Papers ↗ · 2026-05-22 Cached

A survey paper examining the transition of AI from task-specific assistants to workflow-level research automators, defining AutoResearch as the spectrum of AI-powered scientific workflow automation and analyzing challenges in autonomy, reproducibility, and accountability.

0 favorites 0 likes
#research-automation

@mylifcc: Conduct research while sleeping? The viral GitHub project ARIS (8.8k stars) is here! Auto-claude-code-research-in-sleep (ARIS) — a lightweight Markdown-only skill pack that enables Claude Code (or any LL…

X AI KOLs Timeline ↗ · 2026-05-12

ARIS is an open-source tool that has gone viral on GitHub (8.8k stars). It uses a lightweight Markdown skill pack to enable Claude Code or other LLM agents to autonomously complete the entire machine learning research lifecycle, including literature review, experiment execution, and paper writing.

0 favorites 0 likes
#research-automation

NanoResearch: Co-Evolving Skills, Memory, and Policy for Personalized Research Automation

Hugging Face Daily Papers ↗ · 2026-05-11 Cached

NanoResearch is a multi-agent framework designed to personalize research automation by co-evolving skills, memory, and policy to adapt to individual user preferences and research styles.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback