arXiv:2506.09656cs.AI2025-06综述被引 10

系统梳理多智能体AI中的多层级价值对齐方法与挑战

Multi-level Value Alignment in Agentic AI Systems: Survey and Perspectives

  • 构建宏观-中观-微观三级价值层级框架
  • 涵盖从通用到具体场景的多类应用与评估体系
  • 适合关注AI伦理与多智能体协同的研究者

随着AI范式的演进,研究已进入代理型AI阶段。大型语言模型(LLMs)的发展推动其应用向复杂环境下的多智能体自主决策与任务协作演进,随之而来的局势与系统性风险日益凸显。因此,代理型AI的价值对齐问题受到广泛关注,旨在确保智能体的目标、偏好与行为与人类价值观和社会规范保持一致。本文以基于LLM的多智能体系统为典型代表,提出多层级价值框架,系统综述了该领域的三大维度:第一,通过自上而下的方式,在宏观、中观和微观三个层面结构化价值原则;第二,应用场景按从通用到具体的连续谱进行分类,对应上述价值层级;第三,通过系统分析基准数据集与相关方法,将价值对齐方法与评估手段映射至该分层框架。此外,还深入探讨了代理型AI系统中多智能体间的价值协调机制。最后,提出了若干潜在研究方向。

原文摘要 · Abstract (English)

The ongoing evolution of AI paradigms has propelled AI research into the agentic AI stage. Consequently, the focus of research has shifted from single agents and simple applications towards multi-agent autonomous decision-making and task collaboration in complex environments. As Large Language Models (LLMs) advance, their applications become more diverse and complex, leading to increasing situational and systemic risks. This has brought significant attention to value alignment for agentic AI systems, which aims to ensure that an agent's goals, preferences, and behaviors align with human values and societal norms. Addressing socio-governance demands through a Multi-level Value framework, this study comprehensively reviews value alignment in LLM-based multi-agent systems as the representative archetype of agentic AI systems. Our survey systematically examines three interconnected dimensions: First, value principles are structured via a top-down hierarchy across macro, meso, and micro levels. Second, application scenarios are categorized along a general-to-specific continuum explicitly mirroring these value tiers. Third, value alignment methods and evaluation are mapped to this tiered framework through systematic examination of benchmarking datasets and relevant methodologies. Additionally, we delve into value coordination among multiple agents within agentic AI systems. Finally, we propose several potential research directions in this field.

价值对齐多智能体伦理LLM

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。