d4rl

Tag

Cards List
#d4rl

CODS: Iterative Bellman-Residual Data Selection for Reusable Offline Reinforcement Learning

arXiv cs.LG ↗ · 2026-08-11 Cached

Introduces CODS, an iterative critic-guided data selection method for offline reinforcement learning that retains task performance at low data budgets by selecting high-residual transitions over multiple rounds.

0 favorites 0 likes
#d4rl

Neuro-Inspired Inverse Learning for Planning and Control

arXiv cs.AI ↗ · 2026-05-26 Cached

This paper introduces a neuro-inspired framework called Inverter that uses Inverse Learning (IL) for fast and efficient planning and control, achieving significant improvements on D4RL benchmarks and quantum gate synthesis with orders of magnitude less inference computation.

0 favorites 0 likes
← Back to home

Submit Feedback