credibility

标签

Cards List
#credibility

Is This Your Final Answer? Cross-Contextual Consistency as a Measure of LLM Credibility

arXiv cs.CL · 2天前 缓存

This paper introduces Cross-Contextual Consistency (C3), a behavioral property for measuring LLM credibility by checking whether answers remain stable under topic-aligned, content-neutral perturbations. Across 26 models and six benchmarks, they find that higher consistency correlates with correctness, offering a complementary evaluation axis.

0 人收藏 0 人点赞
#credibility

面向检索增强生成的源感知重排序:一种可靠性先验方法

arXiv cs.AI · 2026-07-28 缓存

本文介绍了一种面向RAG的源感知重排序方法,该方法引入了基于领域信息的源可靠性先验,在一个由120份文档组成的健康领域语料库上将Precision@5从0.48提升至0.72,并减少了对抗性文档的检索。

0 人收藏 0 人点赞
#credibility

研究预览协助请求:CALM WINS 关于LLM对虐待情境中依据情绪和脏话使用判断两个发言者可信度的回应

Reddit r/artificial · 2026-07-27

一位研究人员发现,当受害者使用情绪化语言和脏话时,LLM系统性地认为虐待受害者比其跟踪者可信度更低,模型在90.8%的回复中指责受害者,并在57%的情况下指导跟踪者。作者寻求对这些发现的验证和发表指导。

0 人收藏 0 人点赞
#credibility

How AI Slop Is Damaging Our News Media

Reddit r/ArtificialInteligence · 2026-06-10 缓存

文章分析了AI生成的虚假专家、抗议海报和特朗普的合成照片如何侵蚀新闻媒体的可信度,指出媒体在核实信息源上的缺失导致错误信息扩散。

0 人收藏 0 人点赞
← 返回首页

提交意见反馈