pairwise-preference

Tag

Cards List
#pairwise-preference

WorldReward: Reward Modeling for Camera-Conditioned World Models

Hugging Face Daily Papers · 5d ago Cached

WorldReward introduces a vision-language reward model for camera-conditioned world models that unifies action-consistency and visual-quality evaluation through chunk decomposition and preference aggregation, outperforming existing methods like GPT-5.5 on benchmarks.

0 favorites 0 likes
#pairwise-preference

Preferred, Not Safer: Pairwise Preference Is a Poor Proxy for Clinical Safety

arXiv cs.CL · 2026-08-05 Cached

This paper evaluates whether clinician pairwise preferences reliably indicate clinical safety in LLMs, using 26,804 judgments from 736+ clinicians across 13 models. It finds that preference rankings poorly track safety-critical failures and proposes a clinically adjusted ranking that better incorporates rubric-based safety signals.

0 favorites 0 likes
#pairwise-preference

Pairwise Reference Alignment as a Model-Level Ordinal Observable

arXiv cs.CL · 2026-06-01 Cached

This paper formalizes pairwise reference alignment as a model-level ordinal observable, defining a statistic to measure agreement between a model's scoring and a reference preference distribution, with finite-sample estimators and an empirical study on Qwen2.5 models and RewardBench.

0 favorites 0 likes
← Back to home

Submit Feedback