student-teacher

Tag

Cards List
#student-teacher

What Drives Interactive Improvement from Feedback?

arXiv cs.AI ↗ · 2026-07-01 Cached

This paper investigates whether natural-language feedback leads to improvement beyond repeated attempts alone in multi-turn language agent settings. Using a controlled student-teacher protocol across multiple benchmarks, the authors find that self-generated feedback adds little, while strong external teachers yield larger gains, and that the student's ability to act on feedback is a key bottleneck.

0 favorites 0 likes
#student-teacher

ScholarSum: Student-Teacher Abstractive Summarization via Knowledge Graph Reasoning and Reflective Refinement

arXiv cs.CL ↗ · 2026-06-18 Cached

ScholarSum is a hierarchical reflective graph-based framework for scientific abstractive summarization that emulates a student–teacher writing process. It uses a hierarchical knowledge graph to capture global structure, generates an initial draft, and iteratively refines it via evidence retrieval and teacher-like review to improve both fluency and factual faithfulness.

0 favorites 0 likes
#student-teacher

@blc_16: MIT just released a new RL method called Pedagogical RL. The main lesson -> correct reasoning traces can still be bad t…

X AI KOLs Following ↗ · 2026-05-18 Cached

MIT introduces Pedagogical RL, a method that trains a teacher to produce trajectories that are learnable for a student by penalizing surprising steps, improving RL training efficiency.

0 favorites 0 likes
#student-teacher

@SOURADIPCHAKR18: Two things make this work. 1. Spike-aware pedagogy rewards: only reward the model for being correct AND plausible. Puni…

X AI KOLs Following ↗ · 2026-05-14 Cached

Describes a training technique involving spike-aware pedagogy rewards that penalize implausible jumps, and surprisal-gated imitation where the student learns easy tokens quickly and hard ones slowly.

0 favorites 0 likes
← Back to home

Submit Feedback