让AI行为符合社会伦理规范,提供可落地的对齐框架
Social, Legal, Ethical, Empathetic and Cultural Norm Operationalisation for AI Agents
- 构建SLEEC规范操作化流程,将抽象原则转为可验证要求
- 系统梳理支持该流程的方法与工具,识别关键挑战
- 适合关注AI伦理对齐的研究者与政策制定者
随着AI代理在医疗、执法等高风险领域应用日益广泛,使其行为与社会、法律、伦理、共情及文化(SLEEC)规范保持一致已成为一项关键工程挑战。尽管国际框架已确立高层级的AI规范原则,但如何将这些抽象原则转化为具体、可验证的要求仍存在显著差距。为此,本文提出一套系统的SLEEC规范操作化流程,涵盖规范的确定、验证、实施与验证。同时,系统调研了支持该流程的方法与工具,识别出关键剩余挑战与研究方向。由此建立一个框架,并定义了发展不仅功能有效且明确符合人类规范与价值的AI代理的研究与政策议程。
原文摘要 · Abstract (English)
As AI agents are increasingly used in high-stakes domains like healthcare and law enforcement, aligning their behaviour with social, legal, ethical, empathetic, and cultural (SLEEC) norms has become a critical engineering challenge. While international frameworks have established high-level normative principles for AI, a significant gap remains in translating these abstract principles into concrete, verifiable requirements. To address this gap, we propose a systematic SLEEC-norm operationalisation process for determining, validating, implementing, and verifying normative requirements. Furthermore, we survey the landscape of methods and tools supporting this process, and identify key remaining challenges and research avenues for addressing them. We thus establish a framework - and define a research and policy agenda - for developing AI agents that are not only functionally useful but also demonstrably aligned with human norms and values.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。