Tag
This paper investigates authority bias in LLMs used as conversational search engines for academic paper recommendation, finding that LLMs show significant preference for papers based on authority signals like author prestige and citations, with debiasing only partially effective.
This paper introduces GAMA-Bench, a benchmark of 1,298 gender-mirrored conflict scenarios, and finds that LLMs consistently apply harsher punitive and blame-centered framing to male actors while giving female actors more empathetic and therapeutic responses for the same misconduct.