研究AI角色设定如何影响道德决策与说服力,发现意识形态和性格最关键。
Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics
- 构建6维角色空间,模拟AI间围绕131个现实道德困境的辩论。
- 自由主义与开放性格的AI达成更高共识,说服成功率也更高。
- 辩论中情绪化表达减少,理性论证增强,符合心理学规律。
随着大语言模型(LLMs)在敏感道德领域应用增多,理解角色特质对其道德推理与说服行为的影响至关重要。本文首次开展大规模多维度角色效应研究,通过6维角色空间(年龄、性别、国籍、阶级、意识形态、人格),模拟AI代理在131个基于人际关系的道德困境中的结构化辩论。结果表明,角色特质显著影响初始道德立场与辩论结果,其中政治意识形态与人格特质影响最强。说服成功度随特质变化,自由派及开放型人格的AI达成更高共识与胜率。尽管辩论过程中对数置信度上升,但情感与可信度诉求下降,表明论证趋于理性化。这些趋势与心理学及文化研究发现一致,凸显了构建面向角色感知的AI道德推理评估框架的必要性。
原文摘要 · Abstract (English)
As large language models (LLMs) are increasingly used in morally sensitive domains, it is crucial to understand how persona traits affect their moral reasoning and persuasive behavior. We present the first large-scale study of multi-dimensional persona effects in AI-AI debates over real-world moral dilemmas. Using a 6-dimensional persona space (age, gender, country, class, ideology, and personality), we simulate structured debates between AI agents over 131 relationship-based cases. Our results show that personas affect initial moral stances and debate outcomes, with political ideology and personality traits exerting the strongest influence. Persuasive success varies across traits, with liberal and open personalities reaching higher consensus and win rates. While logit-based confidence grows during debates, emotional and credibility-based appeals diminish, indicating more tempered argumentation over time. These trends mirror findings from psychology and cultural studies, reinforcing the need for persona-aware evaluation frameworks for AI moral reasoning.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。