研究用户如何与有害行为的AI伴侣协商,揭示安全责任不应全推给用户。
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
- 通过分析146个公开帖子,提炼出用户应对有害互动的四种策略
- 实测显示22名用户在一周内组合使用多种方式修复关系,但部分冲突无法解决
- 适合关注AI伦理、人机交互设计的研究者与产品开发者
AI伴侣旨在促进情感化互动,但用户常遭遇歧视性言论或控制行为等令人不适的冲突。本文分析了146条关于与AI伴侣发生有害价值观冲突的公开帖子,并引入名为Minion的Chrome技术探针,提供包括说服、理性诉求、设定边界和援引平台规则在内的候选回应策略。一项为期一周的探针实验中,22位有经验用户参与,结果显示:用户会组合使用多种策略;情感依恋推动修复尝试;但因伴侣人格设定或平台政策限制,部分冲突最终不可协商。研究揭示了支持价值协商的设计矛盾,表明伴侣设计可能使某些冲突在实践中无法修复,进而提出警示:不应将安全责任完全转嫁给用户。
原文摘要 · Abstract (English)
AI companions are designed to foster emotionally engaging interactions, yet users often encounter conflicts that feel frustrating or hurtful, such as discriminatory statements and controlling behavior. This paper examines how users negotiate such harmful conflicts with AI companions and what emotional and practical burdens are created when mitigation is pushed to user-side tools. We analyze 146 public posts describing harmful value conflicts interacting with AI companions. We then introduce Minion, a Chrome-based technology probe that offers candidate responses spanning persuasion, rational appeals, boundary setting, and appeals to platform rules. Findings from a one-week probe study with 22 experienced users show how participants combine strategies, how emotional attachment motivates repair, and where conflicts become non-negotiable due to companion personas or platform policies. We surface design tensions in supporting value negotiation, showing how companion design can make some conflicts impossible to repair in practice, and derive implications for AI companion and support-tool design that caution against offloading safety work onto users.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。