标签
This paper introduces an experimental protocol to measure open-ended LLM conformity, showing that wrong peer input degrades revision quality and that evaluators are not neutral when shown peer context, highlighting the need for anchor calibration.
本文揭示了LLM顺从性基准中的一个混淆因素:标准提示将说话者提示与重复的错误答案混合在一起。通过引入一种去除明确说话者的“无来源”条件,作者表明,即使没有说话者,大多数明显的顺从行为仍然存在,这挑战了先前关于社会影响的解释。