让AI在复杂环境中自主生成伦理行为,不依赖人类道德直觉。
NAEL: Non-Anthropocentric Ethical Logic
- 基于主动推理与符号推理的神经符号架构
- 实现自保、学习与集体福利的动态平衡
- 适合多智能体系统中的伦理决策研究
我们提出NAEL(非人类中心伦理逻辑),一种基于主动推理和符号推理的人工智能伦理框架。与传统以人为中心的AI伦理方法不同,NAEL将伦理行为形式化为智能体在动态多智能体环境中最小化全局期望自由能的涌现属性。我们设计了一种神经符号架构,使智能体能在不确定环境下评估自身行为的伦理后果。该系统克服了现有伦理模型的局限,使智能体无需预设人类道德直觉,即可发展出情境敏感、可适应且具有关系性的伦理行为。一个关于资源分配的案例研究展示了NAEL在自保、认知学习与集体福祉之间的动态权衡能力。
原文摘要 · Abstract (English)
We introduce NAEL (Non-Anthropocentric Ethical Logic), a novel ethical framework for artificial agents grounded in active inference and symbolic reasoning. Departing from conventional, human-centred approaches to AI ethics, NAEL formalizes ethical behaviour as an emergent property of intelligent systems minimizing global expected free energy in dynamic, multi-agent environments. We propose a neuro-symbolic architecture to allow agents to evaluate the ethical consequences of their actions in uncertain settings. The proposed system addresses the limitations of existing ethical models by allowing agents to develop context-sensitive, adaptive, and relational ethical behaviour without presupposing anthropomorphic moral intuitions. A case study involving ethical resource distribution illustrates NAEL's dynamic balancing of self-preservation, epistemic learning, and collective welfare.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。