让AI理解并回应人类价值观,构建更可信的智能系统。
Value-Aware Multiagent Systems
- 用形式化语义学习和表示人类价值观
- 确保单个与多智能体系统的价值对齐
- 提供基于价值观的行为可解释性
本文提出人工智能中的价值意识概念,超越传统价值对齐问题。价值意识定义为工程化价值感知AI提供了简洁清晰的路线图,包含三大核心支柱:(1)利用形式化语义学习与表征人类价值观;(2)确保个体智能体及多智能体系统的价值对齐;(3)提供基于价值观的行为可解释性。论文展示了我们在部分议题上的持续研究工作,并探讨了其在真实场景中的应用。
原文摘要 · Abstract (English)
This paper introduces the concept of value awareness in AI, which goes beyond the traditional value-alignment problem. Our definition of value awareness presents us with a concise and simplified roadmap for engineering value-aware AI. The roadmap is structured around three core pillars: (1) learning and representing human values using formal semantics, (2) ensuring the value alignment of both individual agents and multiagent systems, and (3) providing value-based explainability on behaviour. The paper presents a selection of our ongoing work on some of these topics, along with applications to real-life domains.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。