token-weighting

Tag

Cards List
#token-weighting

Token-weighted Direct Preference Optimization with Attention

arXiv cs.CL · 2026-05-22 Cached

Proposes AttentionPO, a token-weighted direct preference optimization method that uses attention from the LLM itself to estimate token weights, improving alignment performance on AlpacaEval, MT-Bench, and ArenaHard without requiring a separate reward model.

0 favorites 0 likes
← Back to home

Submit Feedback