bias-testing

标签

Cards List
#bias-testing

LLM驱动的自动驾驶汽车在行人礼让中继承人类驾驶员偏见:新基准测试的结果与启示

arXiv cs.AI ↗ · 2026-09-02 缓存

本文提出了针对自动驾驶汽车中LLM和VLM的新偏见测试方法,表明这些模型在行人礼让决策中继承了基于性别、种族和年龄等属性的人类偏见。

0 人收藏 0 人点赞
#bias-testing

@no_stp_on_snek: Qwen3.8 lands in tomorrw, so I went back and finished the 3.6-27B card first. No point measuring a successor against a …

X AI KOLs Following ↗ · 2026-08-13 缓存

The author presents an off-label evaluation card for Qwen3.6-27B, covering quantization, reasoning mode effects, bias probes, and jailbreak resistance, and compares reasoning effects with Nemotron 3.5 Lightning, finding that thinking mode is net-negative for Qwen but positive for Nemotron.

0 人收藏 0 人点赞
#bias-testing

对GLM 5.2的政治偏见测试

Reddit r/ArtificialInteligence ↗ · 2026-07-10

测试GLM 5.2语言模型的政治偏见,以评估其公正性和中立性。

0 人收藏 0 人点赞
#bias-testing

AI聊天机器人存在政治偏见?《华盛顿邮报》测试发现:

Reddit r/singularity ↗ · 2026-06-24

《华盛顿邮报》对主流AI聊天机器人进行了测试,发现其回答中存在政治偏见,引发对AI系统客观性的担忧。

0 人收藏 0 人点赞
#bias-testing

文档草垛中的语义针:LLM-as-a-Judge 相似度评分的敏感性测试

arXiv cs.CL ↗ · 2026-04-22 缓存

PNNL 与华盛顿大学的研究人员提出一套系统化框架,测试五种大语言模型在文档中捕捉细微语义变化的能力,揭示位置偏差、上下文连贯效应及模型特有的评分“指纹”。

0 人收藏 0 人点赞
← 返回首页

提交意见反馈