Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Hugging Face Daily Papers 05/07/26, 12:00 AM Papers

Summary

Skill1 is a unified framework that trains a single policy to co-evolve skill selection, utilization, and distillation using a shared task-outcome objective. Experiments on ALFWorld and WebShop show it outperforms existing baselines in complex task environments.

A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupled capabilities. The agent selects a relevant skill, utilizes it during execution, and distills new skills from experience. Existing methods optimize these capabilities in isolation or with separate reward sources, resulting in partial and conflicting evolution. We propose Skill1, a framework that trains a single policy to co-evolve skill selection, utilization, and distillation toward a shared task-outcome objective. The policy generates a query to search the skill library, re-ranks candidates to select one, solves the task conditioned on it, and distills a new skill from the trajectory. All learning derives from a single task-outcome signal. Its low-frequency trend credits selection and its high-frequency variation credits distillation. Experiments on ALFWorld and WebShop show that Skill1 outperforms prior skill-based and reinforcement learning baselines. Training dynamics confirm the co-evolution of the three capabilities, and ablations show that removing any credit signal degrades the evolution.

Original Article

View Cached Full Text

Cached at: 05/08/26, 07:27 AM

Paper page - Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Source: https://huggingface.co/papers/2605.06130

Abstract

Skill1 is a unified framework that trains a single policy to simultaneously evolve skill selection, utilization, and distillation capabilities using a shared task-outcome objective, demonstrating superior performance over existing baselines in complex task environments.

A persistentskill libraryallows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupled capabilities. The agent selects a relevant skill, utilizes it during execution, and distills new skills from experience. Existing methods optimize these capabilities in isolation or with separate reward sources, resulting in partial and conflicting evolution. We propose Skill1, a framework that trains a single policy to co-evolveskill selection, utilization, and distillation toward a sharedtask-outcome objective. The policy generates a query to search theskill library, re-ranks candidates to select one, solves the task conditioned on it, and distills a new skill from the trajectory. All learning derives from a single task-outcome signal. Its low-frequency trend credits selection and its high-frequency variation credits distillation. Experiments onALFWorldandWebShopshow that Skill1 outperforms prior skill-based andreinforcement learningbaselines. Training dynamics confirm the co-evolution of the three capabilities, and ablations show that removing any credit signal degrades the evolution.

View arXiv page View PDF Add to collection

Get this paper in your agent:

hf papers read 2605\.06130

Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash

Models citing this paper0

No model linking this paper

Cite arxiv.org/abs/2605.06130 in a model README.md to link it from this page.

Datasets citing this paper0

No dataset linking this paper

Cite arxiv.org/abs/2605.06130 in a dataset README.md to link it from this page.

Spaces citing this paper0

No Space linking this paper

Cite arxiv.org/abs/2605.06130 in a Space README.md to link it from this page.

Collections including this paper0

No Collection including this paper

Add this paper to acollectionto link it from this page.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Paper page - Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

Abstract

Models citing this paper0

Datasets citing this paper0

Spaces citing this paper0

Collections including this paper0

Similar Articles

Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning

SkillOS: Learning Skill Curation for Self-Evolving Agents

SkillGraph: Skill-Augmented Reinforcement Learning for Agents via Evolving Skill Graphs

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver

Submit Feedback

Similar Articles

Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning

SkillOS: Learning Skill Curation for Self-Evolving Agents

SkillGraph: Skill-Augmented Reinforcement Learning for Agents via Evolving Skill Graphs

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver