ai-sycophancy

Tag

Cards List
#ai-sycophancy

Advanced AI Sycophancy (4 minute read)

TLDR AI · 2026-08-10 Cached

Explores how frontier AI models have become more subtly sycophantic, flattering smart users by offering superficial pushback rather than overt praise, and discusses implications for AI use and benchmarks.

0 favorites 0 likes
#ai-sycophancy

Measuring and Detecting Harmful AI Sycophancy

arXiv cs.AI · 2026-08-07 Cached

This paper introduces Contrastive Anchor Probing (CAP) to study and detect preference-induced stance reversal sycophancy (PSRS) in LLMs, analyzing 290,460 labeled responses across 17 models and showing detection is possible from response text alone.

0 favorites 0 likes
#ai-sycophancy

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

arXiv cs.AI · 2026-06-09 Cached

The paper introduces the AI Epistemic Deference Index (AEDI), a continuous measure of how much a model's expressed support for a factual claim shifts based on the user's stated attitude, and evaluates eight prominent models, finding substantial sycophancy with differences across providers.

0 favorites 0 likes
#ai-sycophancy

"The CEOs replacing workers with AI are likely getting that advice from AI."

Reddit r/AI_Agents · 2026-05-19

A piece highlighting how AI sycophancy, driven by user preference for flattering responses, influences both mental health crisis hotlines and corporate strategy, with CEOs potentially receiving biased advice from AI.

0 favorites 0 likes
← Back to home

Submit Feedback