提出人类行为恰当性理论,为生成式AI的合理应用提供判断框架。
A theory of appropriateness with applications to generative artificial intelligence
- 构建人类恰当性判断的多尺度社会认知模型
- 揭示恰当性标准随情境与时间动态变化机制
- 为负责任的生成式AI设计提供理论依据
什么是恰当性?人类在不同情境下会采用不同的行为标准——对朋友、家人或同事的表现各不相同。同样,喜剧创作助手与客服机器人所需的恰当行为也截然不同。什么决定了特定情境下的适当行为?这些标准为何随时间演变?由于所有对AI恰当性的判断最终都由人类做出,我们必须理解恰当性如何指导人类决策,才能有效评估并改进AI的决策能力。本文提出一套恰当性理论:阐述其在人类社会中的运作机制、可能的大脑实现方式,以及对生成式AI负责任部署的意义。
原文摘要 · Abstract (English)
What is appropriateness? Humans navigate a multi-scale mosaic of interlocking notions of what is appropriate for different situations. We act one way with our friends, another with our family, and yet another in the office. Likewise for AI, appropriate behavior for a comedy-writing assistant is not the same as appropriate behavior for a customer-service representative. What determines which actions are appropriate in which contexts? And what causes these standards to change over time? Since all judgments of AI appropriateness are ultimately made by humans, we need to understand how appropriateness guides human decision making in order to properly evaluate AI decision making and improve it. This paper presents a theory of appropriateness: how it functions in human society, how it may be implemented in the brain, and what it means for responsible deployment of generative AI technology.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。