agent-learning

Tag

Cards List
#agent-learning

EpiCon: Collective Agent Learning through Co-Evolving Multimodal Memory

Hugging Face Daily Papers ↗ · yesterday Cached

EpiCon presents a shared multimodal memory framework for collective learning among AI agents, enhancing performance across eleven benchmarks without updating host model parameters.

0 favorites 0 likes
#agent-learning

ACLArena: Agent Continue Learning in Multi-stage Post-training

Hugging Face Daily Papers ↗ · 2026-09-21 Cached

The paper presents ACLArena, a framework for evaluating Agent Continual Learning in multi-stage post-training, analyzing forgetting and generalization mechanisms, and proposing an improved ACL recipe using offline replay and LoRA experts.

0 favorites 0 likes
#agent-learning

ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments

Hugging Face Daily Papers ↗ · 2026-09-16 Cached

ScienceIDE introduces infrastructure for converting scientific code repositories into programmable environments for scientific agents, enabling task generation, execution, and verification, with trained models showing improvements in scientific code repair and general capabilities.

0 favorites 0 likes
#agent-learning

Training Needs Trustworthy Worlds: Verified Synthetic Web Environments for Agent Learning

arXiv cs.AI ↗ · 2026-08-25 Cached

This paper introduces a framework for constructing verified synthetic web environments to improve the training of web agents, demonstrating enhanced performance and transferability across benchmarks.

0 favorites 0 likes
#agent-learning

EnvHarness: Awakening Static Worlds for Agent Learning

Hugging Face Daily Papers ↗ · 2026-08-20 Cached

EnvHarness introduces a programmable layer to dynamically reshape static environments for reinforcement learning, improving agent performance through automated targeting of weaknesses with EnvRigger.

0 favorites 0 likes
#agent-learning

Has anyone else tried teaching agent networks w/o fine-tuning?

Reddit r/AI_Agents ↗ · 2026-07-29

The author shares a deterministic learning harness that lets multi-agent systems improve across episodes without fine-tuning or prompt edits, by promoting successful strategies into persistent playbooks. On the Mini Amusement Park benchmark, reward improved from 12,121 to 483,019, reaching #1 on the leaderboard.

0 favorites 0 likes
#agent-learning

Sample-Efficient Learning from Agent Experience

Hugging Face Daily Papers ↗ · 2026-07-23 Cached

Proposes Experience Distillation, a method that internalizes in-context learning gains from agent interaction histories into model weights without requiring additional environment interaction, achieving significant sample efficiency improvements on software engineering and text-adventure tasks.

0 favorites 0 likes
#agent-learning

@akshay_pachaar: Hermes /learn explained. agents usually learn the hard way. they struggle through a task live, fail a few times, find t…

X AI KOLs Timeline ↗ · 2026-06-25 Cached

Hermes Agent by Nous Research introduces /learn, a command that lets the agent deliberately create skills from documentation, code, or instructions without needing to first fail at the task, turning any source into a reusable skill.

0 favorites 0 likes
← Back to home

Submit Feedback