rl-for-llms

Tag

Cards List
#rl-for-llms

@ickma2311: CMU Advanced NLP: Reinforcement Learning I had been curious about how RL works on top of LLMs, and this CMU lecture mad…

X AI KOLs Timeline · 2026-04-21 Cached

CMU Advanced NLP lecture clarifies how reinforcement learning optimizes whole-output rewards (correctness, helpfulness, safety) rather than next-token prediction used in pretraining/fine-tuning.

0 favorites 0 likes
← Back to home

Submit Feedback