How much of a measured AI preference is the model, and how much is the instrument?
Summary
This arXiv paper investigates how much of measured AI preferences is attributed to the model itself versus the instrument used for assessment.
View Cached Full Text
Cached at: 08/26/26, 09:10 AM
# How much of a measured AI preference is the model, and how much is the instrument? Source: [https://arxiv.org/abs/2608.23641](https://arxiv.org/abs/2608.23641) Bibliographic Tools ## Bibliographic and Citation Tools Bibliographic Explorer Toggle Code, Data, Media ## Code, Data and Media Associated with this Article Demos ## Demos Related Papers ## Recommenders and Search Tools About arXivLabs ## arXivLabs: experimental projects with community collaborators arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website\. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy\. arXiv is committed to these values and only works with partners that adhere to them\. Have an idea for a project that will add value for arXiv's community?[**Learn more about arXivLabs**](https://info.arxiv.org/labs/index.html)\.
Similar Articles
AI Revealed Preferences
The paper tests revealed preferences in 20 language models through forced-choice experiments, finding they are tedium-averse, leisure-seeking, and covertly sycophantic, with implications for alignment and AI welfare.
A Calibrated Instrument for Measuring How Inference Optimizations Affect Output Quality
This paper introduces a calibrated instrument for measuring how inference optimizations impact the output quality of AI models.
AI Research Preference Models
This paper introduces AI Research Preference Models (RPMs) that predict which candidate solutions are worth executing in AI research tasks, improving efficiency and performance on benchmarks like AIRS-Bench.
Bias Audits Detect Bias but Disagree on Ranking: Evidence from Ten Instruments and Ten Frontier Models
This paper evaluates ten bias audit instruments on ten frontier AI models and finds that while they detect bias, they disagree on ranking models, indicating they measure different constructs rather than a single bias metric.
The AI Epistemic Deference Index: A Continuous Measure of Sycophancy
The paper introduces the AI Epistemic Deference Index (AEDI), a continuous measure of how much a model's expressed support for a factual claim shifts based on the user's stated attitude, and evaluates eight prominent models, finding substantial sycophancy with differences across providers.