persona-conditioning

标签

Cards List
#persona-conditioning

@yoheinakajima: there's an increasing amount of research replacing LLMs as human subjects, so I was curious how well this actually work…

X AI KOLs Following · 2026-08-13 缓存

Yohei Nakajima discusses a new paper testing whether LLMs can replace human subjects in behavioral experiments, finding that a GPT-4.1 persona panel passed coarse marginal checks but failed to provide precise treatment-response estimates, so human substitutability is not established.

0 人收藏 0 人点赞
#persona-conditioning

多未必佳:LLM观点多样性的关键因素是什么?

arXiv cs.CL · 2026-07-24 缓存

一项析因实验表明,角色细节并不会单调增加LLM观点多样性;不同的交互架构探索了互不重叠的观点区域;而温度缩放等低成本干预措施效果甚微。

0 人收藏 0 人点赞
#persona-conditioning

低宜人性人格条件化实现安全LLM微调

arXiv cs.CL · 2026-06-29 缓存

本文介绍了一种基于人格驱动的重写流水线,通过将LLM微调条件化为低宜人性,以降低越狱敏感性和有害输出,同时保持对话温度,无需安全标签或改变训练目标。

0 人收藏 0 人点赞
#persona-conditioning

单一策略,无限NPC:面向可扩展游戏角色的角色追溯共享强化学习策略

arXiv cs.AI · 2026-05-25 缓存

提出PCSP,一种基于冻结LLM角色描述嵌入的单一强化学习策略,可在生活模拟游戏中实现可扩展、实时的角色可追溯NPC控制。实验表明,该方法实现了零样本角色识别和行为对齐,推理速度比LLM基线快。

0 人收藏 0 人点赞
← 返回首页

提交意见反馈