self-evolution

Tag

Cards List
#self-evolution

@shao__meng: https://x.com/shao__meng/status/2104096547747291584

X AI KOLs Timeline ↗ · 3d ago Cached

Professor Li Hongyi from National Taiwan University has updated the 2026 spring machine learning course, focusing on AI Agents and model self-evolution in the large model era. This marks a complete transition of the course from classical machine learning to cutting-edge large model technologies.

0 favorites 0 likes
#self-evolution

Program-Verified Self-Evolution for Vision-Language Models

Hugging Face Daily Papers ↗ · 3d ago Cached

This paper introduces VQS, a method for self-evolving vision-language models that uses program verification to generate and label questions from images, significantly improving accuracy and model performance over existing baselines.

0 favorites 0 likes
#self-evolution

TimeEvo: Failure-Driven Self-Evolution of a Time Series Agent

arXiv cs.AI ↗ · 6d ago Cached

TimeEvo is a failure-driven self-evolution method for time series agents that diagnoses capability gaps, synthesizes tools, and improves accuracy across tasks and backbones.

0 favorites 0 likes
#self-evolution

@ClorisSignal: Previously shared evaluations for single-agent and multi-agent setups, so how can we prove, based on eval, that the age…

X AI KOLs Timeline ↗ · 2026-09-18 Cached

This article explores verifying the genuine improvement of AI agent self-evolution through evaluation frameworks like Meta-Harness, emphasizing the separation of powers among agent modification, evaluation, and deployment decisions to prevent false progress.

0 favorites 0 likes
#self-evolution

Self-Evolving Search Index

Hugging Face Daily Papers ↗ · 2026-09-17 Cached

The paper introduces SELF-INDEX, a framework that enables search indexes to self-evolve autonomously, improving retrieval performance and benefiting downstream applications such as search agents and agent memory systems.

0 favorites 0 likes
#self-evolution

I just solved continual learning

Reddit r/artificial ↗ · 2026-09-16

The article proposes a method for continual learning in AI where the model autonomously decides what to learn and updates its weights in real-time, potentially enabling continuous self-evolution towards AGI.

0 favorites 0 likes
#self-evolution

From State Synchronization to Cognitive Self-Evolution: An Operational Architecture for Cognitive Digital Twins

arXiv cs.AI ↗ · 2026-09-11 Cached

This paper proposes a four-layer architecture for cognitive digital twins that enables self-evolving operational loops through integrated cognitive capabilities and task-oriented feedback mechanisms.

0 favorites 0 likes
#self-evolution

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Hacker News Top ↗ · 2026-09-09

Procedural Graphs presents a framework for LLM agents to self-evolve their execution structures, enhancing adaptability and efficiency in autonomous task completion.

0 favorites 0 likes
#self-evolution

Self-Evolving Skills via Surrogate-Guided Solve-and-Reproduce

arXiv cs.AI ↗ · 2026-09-01 Cached

The paper introduces reSolve, a surrogate-guided solve-and-reproduce framework for self-evolving agent skills that achieves 74.9% performance, surpassing human-curated baselines by 14.8 points.

0 favorites 0 likes
#self-evolution

StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?

Hugging Face Daily Papers ↗ · 2026-09-01 Cached

StudyBench introduces a controlled physics benchmark to measure how efficiently self-evolution methods convert training material into transferable problem-solving ability, revealing gaps in guidance and compute.

0 favorites 0 likes
#self-evolution

HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness?

Hugging Face Daily Papers ↗ · 2026-09-01 Cached

HarnessDev evaluates LLMs by their ability to build and evolve execution harnesses, revealing significant variations in performance and poor transferability across models.

0 favorites 0 likes
#self-evolution

DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory

Hugging Face Daily Papers ↗ · 2026-09-01 Cached

DiagEvo improves language-model self-evolution by deriving training direction from internal failure history via hierarchical error-cause memory and double-confidence filtering, outperforming baselines that rely on external resources.

0 favorites 0 likes
#self-evolution

Aspire: Can Models Self-Evolve from Vague Goals?

Hugging Face Daily Papers ↗ · 2026-08-31 Cached

ASPIRE introduces a benchmark for self-evolving LLM agents from vague natural-language goals, revealing challenges in goal interpretation and stable weight-level improvement.

0 favorites 0 likes
#self-evolution

EvoUndo: Recoverability-Constrained Self-Evolution for LLM Agent Harnesses

Hugging Face Daily Papers ↗ · 2026-08-28 Cached

EvoUndo introduces a framework for evaluating and ensuring recoverability in self-modifying LLM agents, showing that reliable recovery requires co-designing verification, state grounding, and recovery language expressivity.

0 favorites 0 likes
#self-evolution

@Xudong07452910: Recently, while looking into agent self-evolution research, I've been focusing on a question: When are the trajectories left by agents worth continued learning for the next round of training? This time, I took a self-evolving agent paper from SEED and ran it entirely in Apodex…

X AI KOLs Timeline ↗ · 2026-08-26 Cached

This article discusses the importance of trajectory learning in AI agent self-evolution and introduces the Apodex 1.1 system and open-source tool FrontierAgent for executing and evaluating long-duration research tasks.

0 favorites 0 likes
#self-evolution

Rubrics as Visual-Repair Context for Self-Evolving UI-to-Code Generation

Hugging Face Daily Papers ↗ · 2026-08-25 Cached

RubSE is a framework that uses rubric-guided self-evolution to enhance the stability of UI-to-code generation by mitigating visual repair coupling and trajectory collapse.

0 favorites 0 likes
#self-evolution

Auditing Self-Evolution in Financial Agents: Capability Gains, Security Drift, and Execution-Interface Mismatch

arXiv cs.AI ↗ · 2026-08-19 Cached

This paper audits self-evolution mechanisms in financial AI agents, revealing capability improvements alongside security risks such as prompt injection drift and execution-interface mismatches, emphasizing the need for holistic auditing.

0 favorites 0 likes
#self-evolution

GenRouter: Unified Workflow Routing for Agentic Image Generation

Hugging Face Daily Papers ↗ · 2026-08-17 Cached

GenRouter is a unified routing framework for agentic image generation that adaptively directs prompts to optimal workflows, significantly reducing costs and latency while improving visual alignment through demand profiling and self-evolution.

0 favorites 0 likes
#self-evolution

Self-Evolving Embodied Agents via Skill-Harness Evolution

arXiv cs.CL ↗ · 2026-08-13 Cached

This paper introduces SHAPER, a self-evolving framework for embodied agents that keeps model parameters frozen and improves performance by evolving reusable skills and context-code harnesses through target-environment rollouts. Evaluated on VLABench and ESI-Bench, it proposes a practical alternative to fine-tuning when training is expensive or unavailable.

0 favorites 0 likes
#self-evolution

Spatial Memory Agent: Experience-Grounded Procedure Memory for Spatial Intelligence

Hugging Face Daily Papers ↗ · 2026-08-13 Cached

This paper introduces Spatial Memory Agent (SMA), a runtime framework that improves frozen vision-language models' spatial reasoning through verifier-guided reflection and reusable memory without parameter updates or external tools, achieving strong results across five benchmarks and four base VLMs.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback