arXiv:2505.05197cs.AIcs.CY2025-05被引 8

AI对齐不应追求统一道德,而应管理多元分歧。

Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt

  • 提出'适宜性框架',以多元文化与持续适应为基础设计AI伦理
  • 强调持久分歧是常态,需通过社区定制与多中心治理应对
  • 适合关注AI伦理、制度设计的政策制定者与技术团队

人工智能系统日益影响在线空间管理、科研决策和政策建议等关键领域,确保其安全且符合伦理至关重要。然而,当前多数解决方案采用‘一刀切’式的对齐策略,忽视了长期存在的道德多样性,可能引发抵制、削弱信任并动摇制度。本文指出问题根源在于‘理性趋同公理’——即理想条件下理性主体将趋于单一伦理。我们质疑该前提的必要性与可靠性,提出‘适宜性框架’:基于冲突理论、文化演化、多智能体系统与制度经济学,将持久分歧视为常态,通过四项原则设计系统:(1)情境化基础,(2)社区定制,(3)持续适应,(4)多中心治理。主张将对齐的隐喻从道德统一转向冲突管理,认为此举兼具必要性与紧迫性。

原文摘要 · Abstract (English)

Artificial Intelligence (AI) systems are increasingly placed in positions where their decisions have real consequences, e.g., moderating online spaces, conducting research, and advising on policy. Ensuring they operate in a safe and ethically acceptable fashion is thus critical. However, most solutions have been a form of one-size-fits-all "alignment". We are worried that such systems, which overlook enduring moral diversity, will spark resistance, erode trust, and destabilize our institutions. This paper traces the underlying problem to an often-unstated Axiom of Rational Convergence: the idea that under ideal conditions, rational agents will converge in the limit of conversation on a single ethics. Treating that premise as both optional and doubtful, we propose what we call the appropriateness framework: an alternative approach grounded in conflict theory, cultural evolution, multi-agent systems, and institutional economics. The appropriateness framework treats persistent disagreement as the normal case and designs for it by applying four principles: (1) contextual grounding, (2) community customization, (3) continual adaptation, and (4) polycentric governance. We argue here that adopting these design principles is a good way to shift the main alignment metaphor from moral unification to a more productive metaphor of conflict management, and that taking this step is both desirable and urgent.

AI伦理对齐框架多元治理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。