demonstrations

Tag

Cards List
#demonstrations

Learning Implicit Causal World Models from Multi-Agent Demonstrations

arXiv cs.LG ↗ · 2026-07-30 Cached

This paper presents a method for learning implicit causal world models from multi-agent demonstrations, enabling agents to infer causal structures from observed behavior.

0 favorites 0 likes
#demonstrations

AI agents are widely demonstrated, but their commercialization still seems to be not fully mature.

Reddit r/AI_Agents ↗ · 2026-07-09

The article discusses how AI agents are being widely demonstrated for tasks like browsing, coding, and automating workflows, but their commercialization remains immature, with unclear answers about who pays and how.

0 favorites 0 likes
#demonstrations

When Correct Demonstrations Hurt: Rethinking the Role of Exemplars in In-Context Learning

arXiv cs.LG ↗ · 2026-05-27 Cached

This paper reveals a counterintuitive phenomenon where correct demonstrations in in-context learning can degrade model accuracy, introducing task preserving perturbations to study the gap between exemplar correctness and utility.

0 favorites 0 likes
#demonstrations

Self-Distillation Enables Continual Learning [pdf]

Hacker News Top ↗ · 2026-05-17 Cached

Introduces Self-Distillation Fine-Tuning (SDFT), a method that enables on-policy learning from demonstrations to achieve continual learning without catastrophic forgetting, outperforming supervised fine-tuning.

0 favorites 0 likes
← Back to home

Submit Feedback