agentic-training

Tag

Cards List
#agentic-training

@maximelabonne: Such a cool collab, very happy to see this kind of recipe getting open-sourced Train LFM2.5-2.6B on all the harnesses!

X AI KOLs Timeline ↗ · 2d ago Cached

Maximilien Labonne highlights an open-sourced guide to training models with RL inside real agent harnesses like Claude Code, Codex, and OpenCode, enabling training of any model (e.g., LFM2.5-2.6B) on any task set across harnesses.

0 favorites 0 likes
#agentic-training

DeepSeek Elastic Compute (DSec): Sandbox Infrastructure for Effective Agentic Training at Scale

Lobsters Hottest ↗ · 2026-09-23 Cached

DeepSeek Elastic Compute (DSec) is a sandbox infrastructure for effective agentic training of large language models at scale, featuring elastic execution with unified SDK and integration with reinforcement learning frameworks.

0 favorites 0 likes
#agentic-training

@Azaliamirh: Check out TRACE, a new self-improvement approach where the agent identifies the missing capabilities behind its own fai…

X AI KOLs Timeline ↗ · 2026-07-09 Cached

TRACE is a new self-improvement approach where an AI agent identifies the missing capabilities behind its own failures and trains itself to address them. TRACE-trained Qwen3.6-27B achieves 73.2% on SWE-bench Verified, outperforming much larger models with fewer training rollouts.

0 favorites 0 likes
← Back to home

Submit Feedback