multi-objective-rl

Tag

Cards List
#multi-objective-rl

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

arXiv cs.AI · 2026-08-03 Cached

This paper introduces LEMUR, a framework that combines multi-objective reinforcement learning with preference-based learning from multiple human feedback to learn Pareto-optimal policies without predefined reward functions.

0 favorites 0 likes
#multi-objective-rl

AETDICE: Unified Framework and Offline Optimization for Nonlinear Multi-Objective RL

arXiv cs.LG · 2026-07-01 Cached

AETDICE proposes a unified framework for nonlinear multi-objective reinforcement learning in offline settings, bridging SER and ESR paradigms via density-ratio estimation.

0 favorites 0 likes
← Back to home

Submit Feedback