distilled-rl

Tag

Cards List
#distilled-rl

Distilled Reinforcement Learning for LLM Post-training

Hugging Face Daily Papers · 2026-07-19 Cached

Introduces Distilled Reinforcement Learning, a method that uses a teacher model to provide fine-grained token-level gradient signals for LLM post-training, combining reinforcement learning with knowledge distillation.

0 favorites 0 likes
← Back to home

Submit Feedback