multi-objective-alignment

Tag

Cards List
#multi-objective-alignment

Multi-Objective Exploration and Preference Optimization via Mutual Information

arXiv cs.CL · 2026-07-03 Cached

Proposes MI-EPO, an information-theoretic framework for multi-objective alignment of large language models that uses mutual information to enhance exploration and ensure generated responses are distinguishable and aligned with different preference vectors, achieving stable trade-offs across conflicting objectives.

0 favorites 0 likes
← Back to home

Submit Feedback