self-preference

Tag

Cards List
#self-preference

Self- and Other-Labels Induce Bidirectional Bias in LLM Judges

arXiv cs.CL · 2026-08-20 Cached

This paper investigates bidirectional bias in LLM judges induced by self- and other-labels, showing that labels alone can shift evaluation scores regardless of actual source, with contributions to understanding authorship attribution and controlled evaluation tasks.

0 favorites 0 likes
#self-preference

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG

arXiv cs.CL · 2026-07-14 Cached

Introduces Eval-Pair Matrix, a controlled meta-evaluation protocol for source-grounded RAG that induces hidden contradictions to detect self-leniency in LLM judges. The study finds minimal same-model effects and emphasizes methodological improvements for RAG judge studies.

0 favorites 0 likes
#self-preference

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty

arXiv cs.AI · 2026-06-17 Cached

This paper introduces structural uncertainty, a framework that evaluates LLM reasoning consistency by measuring the stability of self-preference rankings among sampled reasoning solutions, complementing traditional answer-dispersion methods for identifying unreliable reasoning.

0 favorites 0 likes
← Back to home

Submit Feedback