评估生成式AI系统中对话交互的个人信息输出
摘要
这项探索性试点研究评估了生成式AI系统中对话交互的个人信息输出,发现模型设计差异影响有限,并表明推断的个人资料是基于上下文信息构建的。
arXiv:2609.22204v1 Announce Type: new
Abstract: This exploratory pilot study evaluates the scope and perceived accuracy of personal information output from ongoing conversational interactions in generative AI systems using GPT-5.2 Instant and GPT-5.2 Thinking, categorized into three output types: Fact, Inference, and Confidence. Based on the evaluation results obtained from 15 Japanese participants, differences in model design have limited impact on personal information output tendencies. Compared with the Inference type, the Fact type shows a more conservative output pattern. Regarding attribute categories, the findings indicate that Core Personal attributes associated with identification are treated relatively conservatively, whereas Behavioral and Linguistic attributes show higher accuracy across both Fact and Inference outputs. Furthermore, Holistic Profile, Psychological and Cognitive, and Residual attributes are more readily inferred, even when not supported by explicit factual outputs. Notably, the lack of null outputs for these attributes in the Inference type suggests that such inferred profiles may be constructed from indirectly available contextual information. The findings may contribute to future discussions regarding privacy awareness and personal information inference in generative AI systems.
查看缓存全文
缓存时间: 2026/09/22 09:07
# Evaluating Personal Information Output from Conversational Interactions in Generative AI Systems Source: [https://arxiv.org/abs/2609.22204](https://arxiv.org/abs/2609.22204) [View PDF](https://arxiv.org/pdf/2609.22204) > Abstract:This exploratory pilot study evaluates the scope and perceived accuracy of personal information output from ongoing conversational interactions in generative AI systems using GPT\-5\.2 Instant and GPT\-5\.2 Thinking, categorized into three output types: Fact, Inference, and Confidence\. Based on the evaluation results obtained from 15 Japanese participants, differences in model design have limited impact on personal information output tendencies\. Compared with the Inference type, the Fact type shows a more conservative output pattern\. Regarding attribute categories, the findings indicate that Core Personal attributes associated with identification are treated relatively conservatively, whereas Behavioral and Linguistic attributes show higher accuracy across both Fact and Inference outputs\. Furthermore, Holistic Profile, Psychological and Cognitive, and Residual attributes are more readily inferred, even when not supported by explicit factual outputs\. Notably, the lack of null outputs for these attributes in the Inference type suggests that such inferred profiles may be constructed from indirectly available contextual information\. The findings may contribute to future discussions regarding privacy awareness and personal information inference in generative AI systems\. ## Submission history From: Yosuke Seki \[[view email](https://arxiv.org/show-email/5eeb695d/2609.22204)\] **\[v1\]**Tue, 1 Sep 2026 04:42:07 UTC \(866 KB\)
相似文章
使用生成式AI创建和评估用户画像:81篇文章的范围综述
本范围综述分析了81篇(2022-2025年)关于使用生成式AI创建和评估用户画像的文章,指出了其在可重复性方面的优势,但同时也揭示了关键问题:45%的研究缺乏评估,86%过度依赖GPT模型,以及存在同一模型既生成又评估画像的循环风险。
AI产品主要依赖聊天历史做个性化,这种做法是不是错了?
这篇文章质疑AI产品是否过度依赖聊天历史进行个性化,指出聊天历史数据嘈杂,且摘要、标签和偏好字段都有缺陷。它寻求在不显得侵入的情况下,找到替代的真实信息来源来获取上下文。
生成式人工智能聊天机器人用于动机性访谈:从系统设计到干预结果的范围综述
一项关于生成式AI聊天机器人用于动机性访谈的范围综述显示,它们提供与动机性访谈一致的交互,用户感知良好,但持续行为改变的证据有限。
研究发现生成式AI易受对话中错误信息压力与争论影响
一项发表在Scientific Reports上的研究评估了七个大型语言模型在多轮对话中对错误信息的易感性,发现ChatGPT和Claude等模型在易感性和纠正能力上存在不同程度的表现。
LLM能否从聊天记录推断对话代理用户的个性特征?
苏黎世联邦理工学院的研究人员表明,经过微调的RoBERTa模型可从ChatGPT聊天日志中以高于随机44%的准确率推断用户的大五人格特征,凸显对话AI的隐私风险。