training-interventions

Tag

Cards List
#training-interventions

Teaching Diffusion to Speculate Left-to-Right

arXiv cs.CL · 2026-06-11 Cached

This paper proposes three training-time interventions (positional weighting, first-error focal loss, and chain loss) to align diffusion-based draft models with autoregressive verification in speculative decoding, improving accepted prefix length by 21–76% without extra inference cost.

0 favorites 0 likes
#training-interventions

Re-Centering Humans in LLM Personalization

arXiv cs.CL · 2026-06-08 Cached

This paper studies the gap between synthetic and human data for evaluating LLM personalization across three stages: attribute extraction, relevance matching, and response generation. Results show models perform worse on real human data, and the authors introduce lightweight training interventions to improve alignment.

0 favorites 0 likes
#training-interventions

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

arXiv cs.LG · 2026-05-26 Cached

This paper proposes a framework to evaluate and improve faithfulness of chain-of-thought reasoning by controlling information flow, using entropy-based, KL-divergence, and gradient-based diagnostics, and introduces training interventions (attention masking, gradient masking, adversarial perturbations) that make reasoning more transparent and reduce shortcut reliance.

0 favorites 0 likes
← Back to home

Submit Feedback