pairwise-comparison

Tag

Cards List
#pairwise-comparison

Adaptive KappaSharp: Condition-Number Shaping for Preferential Bayesian Optimization

arXiv cs.LG · 4d ago Cached

This paper introduces KappaSharp, a method for condition-number shaping in Preferential Bayesian Optimization that corrects the ill-conditioned Hessian from isolated pairwise comparisons, showing up to +10.9% improvement over the standard PairedGP/EUBO baseline on 11 benchmarks including plasma medicine controller tuning.

0 favorites 0 likes
#pairwise-comparison

(Towards) Scalable Reliable Automated Evaluation with Large Language Models

arXiv cs.CL · 2026-07-31 Cached

This paper proposes a scalable, domain-agnostic framework for automated LLM evaluation that uses pairwise comparisons by multiple LLMs and an Elo rating system to approximate expert judgments, reducing the need for human intervention.

0 favorites 0 likes
#pairwise-comparison

Finding the Best Dog Treat with Statistics

Hacker News Top · 2026-06-22 Cached

Uses the Bradley-Terry model and Elo rating system to statistically determine a dog's favorite treat through pairwise comparison experiments.

0 favorites 0 likes
#pairwise-comparison

Prompt Perturbation for Reliable LLM Evaluation over Comparison Graphs

arXiv cs.CL · 2026-06-17 Cached

Proposes a prompt perturbation framework that generates perturbed prompt variants, filters out structurally inconsistent comparison patterns using graph-level consistency checks, then applies standard ranking methods to yield more reliable LLM rankings.

0 favorites 0 likes
#pairwise-comparison

When it comes to predicting people’s preferences, it pays to consider “the power of three”

MIT News — Artificial Intelligence · 2026-06-11 Cached

MIT researchers present a paper showing that using three-way comparisons instead of pairwise comparisons can significantly improve the accuracy of random utility models for predicting human preferences.

0 favorites 0 likes
← Back to home

Submit Feedback