automated-metrics

Tag

Cards List
#automated-metrics

Response drift across frontier large language models

arXiv cs.CL · 4d ago Cached

A large-scale human evaluation of 10 frontier LLMs across 62 questions finds that all models exhibit response drift, with most converging to a 78-81% deviation ceiling, while two achieve lower deviation. Drift varies by domain and question, and automated metrics explain little of human judgments, highlighting the need for human evaluation.

0 favorites 0 likes
#automated-metrics

Fully Automated Identification of Lexical Alignment and Preference-Stage Shifts in Large Language Models

arXiv cs.CL · 2026-06-03 Cached

This paper introduces two automated metrics, Lexical Alignment Score and Triangulated Preference Shift, to identify lexical overuse in LLMs and attribute it to preference learning stages. The method is tested on six model families using PubMed abstracts, replicating prior findings without manual intervention.

0 favorites 0 likes
← Back to home

Submit Feedback