@neural_avb: If yall are interested in On Policy Distillation, check this specific repo. Somebody put together a curated collection …

X AI KOLs Timeline Tools

Summary

A curated collection of papers and tools for On Policy Distillation, organized and annotated with a getting-started section, shared via a GitHub repo.

If yall are interested in On Policy Distillation, check this specific repo. Somebody put together a curated collection of papers and tools categorized and annotated. Comes with a "Get Started" section too. https://t.co/nwVgFdoLDY
Original Article
View Cached Full Text

Cached at: 05/29/26, 02:10 PM

If yall are interested in On Policy Distillation, check this specific repo.

Somebody put together a curated collection of papers and tools categorized and annotated. Comes with a “Get Started” section too. https://t.co/nwVgFdoLDY

pradheep (@pradheepraop): starting a proper deep dive into opd/opsd now.

thanks @neural_avb and @chrisliu298 for consolidating some really useful resources.

recommendations are welcome 🙂

Similar Articles

On-policy distillation: one of the hottest terms on PapersWithCode [R]

Reddit r/MachineLearning

Hugging Face's Niels introduces On-policy Distillation (OPD), a key post-training technique used in models like Qwen 3.6/3.7, GLM-5.1, and DeepSeek-V4, now featured on PapersWithCode with a linked whiteboard explanation by Sasha Rush and Dwarkesh Patel.