arXiv:2609.03920cs.AI2026-09

用架构设计保障AI系统中的隐私、公平与安全。

Value-Preserving Architectures for Agentic AI Systems

论文配图:Value-Preserving Architectures for Agentic AI Systems
图 1 · 摘自论文原文
  • 提出三种支持人类价值的架构模式:联邦拓扑、分布式设计、守卫代理。
  • 实证表明架构选择能有效提升隐私保护、多样性和公平性。
  • 适合关注AI伦理与可信系统设计的研究者与工程师。

智能体化AI和基于大模型的多智能体系统(MAS)为自动化复杂任务带来前所未有的机遇,同时也引发了对隐私、公平与安全等核心人类价值观保护的担忧。传统软件工程关注功能正确性,而大模型与智能体融入社会技术系统后,亟需负责任的软件工程和稳健的价值对齐。在多智能体系统中,协调机制、通信协议与系统拓扑等架构决策直接影响系统行为与产出结果。本文认为,架构选择不仅影响功能与性能,还能促进价值导向的行为。因此,我们研究不同架构设计如何支持人类中心价值观,提出三类价值保持架构模式:(i) 基于联邦拓扑的隐私感知架构,(ii) 分布式架构以促进多元与多样性,(iii) 守卫代理架构用于检测与缓解不公平现象。最后,通过代表性应用场景展示其实际应用。本工作将架构设计与人类中心价值相联结,为可信赖多智能体系统的构建奠定统一的架构模式与指南基础。

原文摘要 · Abstract (English)

The emergence of agentic AI and LLM-based multi-agent systems (MAS) presents unprecedented opportunities for automating complex tasks, while simultaneously raising critical concerns about the preservation of fundamental human-centered values, such as privacy, fairness, and safety. Although software engineering has traditionally focused on functional correctness, the adoption of LLMs and AI agents into complex socio-technical systems has intensified the need for responsible software engineering and robust value alignment. In MAS, architectural design decisions, such as coordination mechanisms, communication protocols, and system topologies, play a central role in shaping system behavior and the outcomes they produce. This paper argues that architectural choices influence not only the functionality and performance of MAS but can also promote value-oriented system behavior. Therefore, we investigate how different architectural designs support different human-centered values, discussing the following value-preserving architectural patterns: (i) a privacy-aware architecture with a federated topology, (ii) a distributed architecture to promote pluralism and diversity, and (iii) a guard-agent architecture to detect and mitigate unfairness. Finally, we introduce representative use cases to illustrate the proposed architectures in real-world scenarios. By linking architectural design with human-centered values, this work lays the foundation for a unified set of architectural patterns and guidelines towards the design of trustworthy MAS.

多智能体架构设计价值观对齐AI伦理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。