agentic-learning

Tag

Cards List
#agentic-learning

Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning

Hugging Face Daily Papers · 6d ago Cached

Agent-G^2 introduces a Gaussian guidance framework for hint depth in reinforcement learning, enhancing performance on long-horizon agentic tasks without extra probing rollouts, with superior results on ALFWorld and WebShop benchmarks.

0 favorites 0 likes
#agentic-learning

Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis

arXiv cs.AI · 2026-08-17 Cached

This paper proposes a framework for evaluating agentic learning harnesses in cybersecurity without labeled benchmarks, using a teacher-student model based on the scaling hypothesis to proxy performance improvements.

0 favorites 0 likes
#agentic-learning

SERL-SQL: Selective Hindsight Distillation for Text-to-SQL Reinforcement Agentic Learning

arXiv cs.CL · 2026-08-04 Cached

SERL-SQL proposes a selective execution-grounded reinforcement learning framework for multi-turn Text-to-SQL agents, using teacher-student likelihood gaps to reweight GRPO advantages on SQL action tokens. It achieves strong results on BIRD and Spider benchmarks.

0 favorites 0 likes
#agentic-learning

@rohanpaul_ai: Can LLM agents actually discover hidden rules by interacting? The answer is uncomfortable. The more complicated the hid…

X AI KOLs Following · 2026-06-22 Cached

This paper investigates whether LLM agents can infer hidden world models through interaction, finding that they struggle to build stable internal models as complexity increases.

0 favorites 0 likes
← Back to home

Submit Feedback