为复杂道德场景设计可解释的AI决策框架,兼顾深度思考与实时响应。
Normative Moral Pluralism for AI: A Framework for Deliberation in Complex Moral Contexts
- 基于规范性多元主义构建分层推理架构,融合上下文与文化因素。
- 双层混合结构实现道德阈值控制与情境化权衡,支持高风险决策。
- 适合追求可解释、负责任AI的开发者与伦理研究者参考。
本文提出一种以规范性道德多元主义为基础的审议式道德推理框架,旨在处理复杂的道德情境。系统通过生成、筛选和权衡来自不同伦理视角的规范性论据,实现透明且有原则的理性审议。该框架不仅服务于机器伦理,还通过结构化道德推理与行动的衔接,为价值对齐提供实质性贡献。其核心是双层混合架构:通用层通过自上而下与自下而上的学习定义道德阈值;局部层在不突破阈值的前提下,学习情境化权衡并整合文化特异的规范内容。框架将道德复杂性扩展至冲突信念、多因素困境、多方利益相关者及非道德考量的整合,致力于在真实高风险场景中实现有道德依据的决策。此外,该系统还可作为新型两层架构的道德教师,训练快速响应的模型,使其在无需完整审议结构的情况下实现及时行动。
原文摘要 · Abstract (English)
The conceptual framework proposed in this paper centers on the development of a deliberative moral reasoning system - one designed to process complex moral situations by generating, filtering, and weighing normative arguments drawn from diverse ethical perspectives. While the framework is rooted in Machine Ethics, it also makes a substantive contribution to Value Alignment by outlining a system architecture that links structured moral reasoning to action under time constraints. Grounded in normative moral pluralism, this system is not constructed to imitate behavior but is built on reason-sensitive deliberation over structured moral content in a transparent and principled manner. Beyond its role as a deliberative system, it also serves as the conceptual foundation for a novel two-level architecture: functioning as a moral reasoning teacher envisioned to train faster models that support real-time responsiveness without reproducing the full structure of deliberative reasoning. Together, the deliberative and intuitive components are designed to enable both deep reflection and responsive action. A key design feature is the dual-hybrid structure: a universal layer that defines a moral threshold through top-down and bottom-up learning, and a local layer that learns to weigh competing considerations in context while integrating culturally specific normative content, so long as it remains within the universal threshold. By extending the notion of moral complexity to include not only conflicting beliefs but also multifactorial dilemmas, multiple stakeholders, and the integration of non-moral considerations, the framework aims to support morally grounded decision-making in realistic, high-stakes contexts.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。