Tag
Introduces QDOS, a unified pipeline for offline-to-online reinforcement learning that uses advantage-weighted quality-diversity pretraining to extract diverse and high-value skills, significantly improving performance in manipulation and locomotion tasks.