arXiv:2411.10613cs.AIcs.CY2024-11被引 5

让智能体学会体谅他人,实现多元价值观对齐

Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI

  • 通过考虑他人未来福祉与自主性来设计行为策略
  • 体谅行为能有效促进多元价值共存的系统对齐
  • 适合关注伦理对齐与多主体协作的研究者

多元对齐关注确保人工智能系统的目标与行为与人类多元的价值观和视角相一致。本文研究了在代理型AI背景下多元对齐的实现路径,特别是当一个智能体在学习策略时,能够顾及环境中其他主体的价值观与视角。为此,我们证明了考虑其他(人类)代理未来的福祉与自主性,能够促成一种多元对齐的形式。

原文摘要 · Abstract (English)

Pluralistic alignment is concerned with ensuring that an AI system's objectives and behaviors are in harmony with the diversity of human values and perspectives. In this paper we study the notion of pluralistic alignment in the context of agentic AI, and in particular in the context of an agent that is trying to learn a policy in a manner that is mindful of the values and perspective of others in the environment. To this end, we show how being considerate of the future wellbeing and agency of other (human) agents can promote a form of pluralistic alignment.

AI对齐代理智能体价值观对齐

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。