self-evolving

Tag

Cards List
#self-evolving

@rohanpaul_ai: New Google paper shows LLM agents handle long tasks better when their workflow lives in an editable procedure graph tha…

X AI KOLs Timeline ↗ · 6d ago Cached

A Google paper introduces Procedural Graphs, an editable workflow structure for LLM agents that improves performance on long tasks by evolving from execution feedback, outperforming baselines in most benchmarks.

0 favorites 0 likes
#self-evolving

Position: It is Time to Virtualize Foundation Models with a Self-evolving Operating System Layer

arXiv cs.AI ↗ · 2026-09-18 Cached

This position paper proposes a Foundation Model Operating System (FMOS) to virtualize foundation model interactions, providing applications with dedicated, trustworthy instances and enabling self-evolving capabilities through orchestration and policy enforcement.

0 favorites 0 likes
#self-evolving

EvoOntology: A Self-Evolving Ontology Layer for Data Agents

Hugging Face Daily Papers ↗ · 2026-09-14 Cached

EvoOntology introduces a self-evolving ontology layer for data agents, encapsulated as an MCP server, to bridge the agent-data gap and improve performance on heterogeneous data tasks as shown in benchmarks.

0 favorites 0 likes
#self-evolving

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Hugging Face Daily Papers ↗ · 2026-09-08 Cached

This paper introduces Procedural Graph, a framework that organizes LLM agent actions into structured triplets for improved long-horizon tool use, with self-evolving topology to enhance performance.

0 favorites 0 likes
#self-evolving

@kirbytheodor: deslopped the tui a bit so it's actually useful at a glance ouroboros with fable and astra deleted 2.79 million lines o…

X AI KOLs Timeline ↗ · 2026-09-07 Cached

Enhanced the TUI of ouroboros, a CLI tool for self-evolving auto-research systems, and deleted 2.79 million lines of vestigial code to improve usability.

0 favorites 0 likes
#self-evolving

CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning

arXiv cs.AI ↗ · 2026-09-03 Cached

CHIME is a credit-aware hierarchical memory framework that separates planning and execution memory banks to improve long-horizon agentic planning by accurately attributing task outcomes and outperforming baselines.

0 favorites 0 likes
#self-evolving

J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data

arXiv cs.LG ↗ · 2026-08-28 Cached

J-Zero is a unified framework for co-evolving Challenger, Solver, and Judge models from zero data, enabling self-improvement in language models across both verifiable and unverifiable domains with performance surpassing baselines.

0 favorites 0 likes
#self-evolving

JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution

Hugging Face Daily Papers ↗ · 2026-08-26 Cached

JIT-Agent is a trainable model that synthesizes adaptive agent harnesses for off-the-shelf LLMs, improving performance across diverse models and tasks.

0 favorites 0 likes
#self-evolving

@Sa4d_k1: Among the distinguished papers recently published in the field of AI Agents. The research involved 60 researchers from …

X AI KOLs Timeline ↗ · 2026-08-15 Cached

This paper surveys agent memory in AI agents, addressing the problem of context explosion and proposing a framework for self-evolving agents through various memory types and management strategies.

0 favorites 0 likes
#self-evolving

@seclink: The barrier for autonomous-evolving code agent harnesses is genuinely low. It reminds me of the famous saying in the stand-up comedy world: Everyone can get on stage and perform for 3 minutes of stand-up comedy. Now, the code agent field is just like stand-up comedy; anyone can casually create a top-tier code agent, and everyone claims to be...

X AI KOLs Following ↗ · 2026-08-14 Cached

This article comments on the low barrier to entry for autonomous-evolving code agent tools, drawing an analogy to stand-up comedy performances, pointing out that in the current code agent field, everyone can claim to be skilled at complex tasks.

0 favorites 0 likes
#self-evolving

ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval

arXiv cs.CL ↗ · 2026-08-14 Cached

Introduces ERSkill, a retrieval-centric framework for self-evolving, skill-guided adaptive memory access in LLM agents. It co-evolves retrieval skills and a routing policy, substantially outperforming strong baselines across agent memory benchmarks.

0 favorites 0 likes
#self-evolving

MindMemOS: A Portable and Self-Evolving Memory Operating Layer for AI Agents

arXiv cs.AI ↗ · 2026-08-14 Cached

The article introduces MindMemOS, a portable and self-evolving memory operating layer for AI agents that uses a unified entity-property-time structure, with algorithms for memory refinement and skill evolution. It achieves notable accuracy on LOCOMO and PersonaMem benchmarks and improves SpreadsheetBench performance by 9.2 percentage points.

0 favorites 0 likes
#self-evolving

@BoxMrChen: The Deepseek Harness philosophy is awesome. They made everything into plugins — GUI, TUI — and oppose hardcoding workflows like plan, subagent, permissions, MCP into the core. To be honest, at first glance I felt they reinvented Pi Agent, even though it was built on Cordis...

X AI KOLs Timeline ↗ · 2026-08-13 Cached

The author comments on Deepseek Harness's plugin-based philosophy, arguing that it turns everything (GUI/TUI, etc.) into plugins and opposes hardcoding workflows. They compare it with Pi Agent and Prime-Agent, propose the idea of building a self-evolving Agent in a REPL, and have already started implementing it with Codex.

0 favorites 0 likes
#self-evolving

MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph

arXiv cs.AI ↗ · 2026-08-12 Cached

MEGA is a self-evolving infrastructure for coding-agent optimization that distills reusable wisdom from sessions, composes it via a typed Wisdom Graph, and uses operational evidence to continuously improve both the agents and the knowledge guiding their optimization.

0 favorites 0 likes
#self-evolving

GeoForge: Non-Parametric Self-Evolving Agents for Earth-Observation Reasoning

arXiv cs.AI ↗ · 2026-08-12 Cached

GeoForge is a training-free, self-evolving framework for Earth-observation reasoning that structures completed trajectories into nonparametric memories to improve LLM agent planning and tool-use without updating the backbone model.

0 favorites 0 likes
#self-evolving

Self-evolving Agentic Customer Support System at LinkedIn

arXiv cs.AI ↗ · 2026-08-12 Cached

LinkedIn presents a self-evolving agentic customer support system that integrates RAG with evolutionary auto-prompting and modular evaluation, achieving significant gains in production A/B tests including a 9.0-point increase in QA self-serve and 30.6-point improvement in routing accuracy.

0 favorites 0 likes
#self-evolving

Self-Evolving Neuro-Symbolic Skills for Tool-Augmented Spatial Reasoning

arXiv cs.AI ↗ · 2026-08-11 Cached

Presents NeSy-Spatial, a neuro-symbolic framework that self-evolves spatial reasoning skills by composing tool-use and geometry skills, improving accuracy on spatial reasoning benchmarks.

0 favorites 0 likes
#self-evolving

@AdinaYakup: BigBang-v1a self-evolving LLM from endless frontier lab in Shanghai - Self-evolving training with AI generated frontier…

X AI KOLs Timeline ↗ · 2026-08-07 Cached

BigBang-v1 is a self-evolving 36B LLM from Endless Frontier Lab in Shanghai, trained with AI-generated frontier tasks and achieving strong performance with only 10K high-quality examples across science, coding, tool use, and long context.

0 favorites 0 likes
#self-evolving

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

arXiv cs.LG ↗ · 2026-08-07 Cached

Introduces EvoHarness-RL, a framework that learns runtime harness policies for long-horizon LLM agents, enabling them to construct and update external state (belief, progress, experience) during task execution. Using Qwen3-8B on ALFWorld, it achieves 96.9% success and reveals harness annealing and evolution dynamics.

0 favorites 0 likes
#self-evolving

@Saboo_Shubham_: wtf is a dynamic agent orgs. Self-evolving agent orgs where graph rewrites itself while the work is happening.

X AI KOLs Timeline ↗ · 2026-08-05 Cached

A tweet expressing amazement at the concept of dynamic agent orgs—self-evolving multi-agent systems where the graph structure rewrites itself during execution.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback