co-evolution

Tag

Cards List
#co-evolution

KernelZero: Co-Evolving Proposer and Coder for Continuously Improved GPU Kernel Generation

Hugging Face Daily Papers ↗ · 3d ago Cached

KernelZero introduces a co-evolution framework with Proposer and Coder models to enhance GPU kernel generation, achieving superior performance on CUDA and Triton benchmarks.

0 favorites 0 likes
#co-evolution

MACE: Memory-Agent Co-Evolution with Adaptive Memory Graphs for Multi-Agent Systems

arXiv cs.LG ↗ · 2026-09-21 Cached

MACE is a memory-agent co-evolution framework for LLM-based multi-agent systems that uses adaptive memory graphs to organize functional memory units and adapt through execution feedback, improving task performance over baselines.

0 favorites 0 likes
#co-evolution

Salesforce Finds Better Ways to Co-Evolve Agents and Their Harnesses (9 minute read)

TLDR AI ↗ · 2026-09-11 Cached

This research explores combining harness evolution with model adaptation for AI agents, discovering that direct imitation from experts harms weaker models' performance and proposing an on-policy correction method to improve performance without breaking harness fit for enterprise tasks.

0 favorites 0 likes
#co-evolution

J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data

arXiv cs.LG ↗ · 2026-08-28 Cached

J-Zero is a unified framework for co-evolving Challenger, Solver, and Judge models from zero data, enabling self-improvement in language models across both verifiable and unverifiable domains with performance surpassing baselines.

0 favorites 0 likes
#co-evolution

EnvHarness: Awakening Static Worlds for Agent Learning

Hugging Face Daily Papers ↗ · 2026-08-20 Cached

EnvHarness introduces a programmable layer to dynamically reshape static environments for reinforcement learning, improving agent performance through automated targeting of weaknesses with EnvRigger.

0 favorites 0 likes
#co-evolution

HELIX: Model-Harness Co-evolution for Recursive Self-Improvement

arXiv cs.AI ↗ · 2026-08-17 Cached

This paper proposes model-harness co-evolution as a fundamental principle for recursive self-improvement in AI agents, introducing HELIX, a source-traceable system that improves both execution and learning by generating structured training signals from verified trajectories.

0 favorites 0 likes
#co-evolution

Red queen hypothesis – a new way forward for self-improving AI

Hacker News Top ↗ · 2026-08-16 Cached

Researchers propose the Red Queen Gödel Machine, a framework for self-improving AI where agents and evaluators evolve together to overcome evaluation ceilings, showing enhanced performance in tasks like scientific paper writing and grading.

0 favorites 0 likes
#co-evolution

Evolving Parallel Algorithm Portfolios via Potential-Aware Instance Generation with LLMs

arXiv cs.AI ↗ · 2026-08-10 Cached

This paper introduces PIAC, a framework that improves LLM-based automatic construction of parallel algorithm portfolios by using a potential-gain metric that eliminates the need for reference solutions and by leveraging LLMs to generate diverse instance mutators. It consistently outperforms existing LLM-ACP baselines on TSP and CVRP, achieving up to 19.76% relative improvement.

0 favorites 0 likes
#co-evolution

Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

Hugging Face Daily Papers ↗ · 2026-08-10 Cached

A survey paper proposing a three-stage taxonomy of co-evolution in agentic systems, covering agent-agent, agent-environment, and meta co-evolution to enable open-ended improvement beyond fixed human-designed paths.

0 favorites 0 likes
#co-evolution

Scaffold-Mediated Post-Training: Co-Evolving Model Parameters and Procedural Scaffold Graphs

arXiv cs.CL ↗ · 2026-08-07 Cached

This paper proposes scaffold-mediated post-training, a paradigm where procedural scaffolds co-evolve with LLM parameters through discovery, distillation, and dynamic recompilation. On FeatureBench, automatically discovered skills improve pass rate by 8.1pp, with a 27.7% pass rate after distillation.

0 favorites 0 likes
#co-evolution

Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning

Hugging Face Daily Papers ↗ · 2026-08-05 Cached

This paper introduces CoCoEvolve, a self-supervised method that improves cross-representation understanding across charts, tables, and code by enforcing one-to-one consistency between representations, with training-time and test-time co-evolution objectives.

0 favorites 0 likes
#co-evolution

Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale

Hugging Face Daily Papers ↗ · 2026-07-30 Cached

Echoverse presents a method for generating deep, evolving synthetic environments to train computer-use agents, demonstrating substantial accuracy gains and releasing a benchmark with grounded graders.

0 favorites 0 likes
#co-evolution

DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space

Hugging Face Daily Papers ↗ · 2026-07-28 Cached

Introduces DecoEvo, a score-decoupled co-evolution method for LLM optimization in text space that jointly improves solver and rubric-generator skills without gold rubrics, achieving 2.8–5.0% relative gains over baselines across five benchmarks.

0 favorites 0 likes
#co-evolution

EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL

arXiv cs.AI ↗ · 2026-07-24 Cached

EvoSQL is a co-evolution framework for Text-to-SQL that iteratively improves SQL generation via a generator-critic pair with episodic memory, achieving gains on Spider and BIRD benchmarks.

0 favorites 0 likes
#co-evolution

Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills

Hugging Face Daily Papers ↗ · 2026-07-24 Cached

The paper introduces Skill Self-Play (Skill-SP), a co-evolutionary framework that uses a proposer, solver, and skill controller to bridge structured verification and open-ended exploration, improving LLM performance on tool-use and reasoning benchmarks.

0 favorites 0 likes
#co-evolution

From Memory to Skills: Evidence-Grounded Co-Evolution Governance for Long-Horizon LLM Agents

arXiv cs.CL ↗ · 2026-07-21 Cached

MSCE is a training-free framework that organizes LLM agent experience into three memory levels and converts them into reusable skills with evidence links, outperforming existing memory and skill-augmented baselines.

0 favorites 0 likes
#co-evolution

EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World

Hugging Face Daily Papers ↗ · 2026-07-19 Cached

Introduces EvolvingWorld, an open-schema framework and benchmark for co-evolving role-play agents and world models in interactive literary worlds, enabling long-horizon simulation with persistent character and world state updates.

0 favorites 0 likes
#co-evolution

Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents

arXiv cs.AI ↗ · 2026-07-15 Cached

This paper proposes a method for co-evolving evaluation metrics and skills in self-improving LLM agent systems, demonstrating that metrics can be evolved and that a co-evolution approach recovers most of the performance of a ground-truth-driven oracle across code generation, text-to-SQL, and report generation tasks.

0 favorites 0 likes
#co-evolution

Harness-Aware Self-Evolving: Co-Evolving Model Weights, Harness, and Task Solutions

arXiv cs.AI ↗ · 2026-07-07 Cached

HASE is a reinforcement-learning framework that co-evolves model weights, task solutions, and harness components (guidance and evaluation) in a unified agentic process, enabling a single 8B-parameter model to match the performance of much larger systems on text classification and alpha factor mining tasks.

0 favorites 0 likes
#co-evolution

@niclane7: Just in time for ICML week, we are sharing our take on a key question for recursive self-improving AI. How can AI keep …

X AI KOLs Timeline ↗ · 2026-07-05 Cached

The Red Queen Gödel Machine enables recursive self-improvement in AI by co-evolving the agent and evaluator, achieving better coding performance with fewer tokens.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback