arXiv:2607.04710cs.AI2026-07

通过融合利他与公平偏好,让智能体在合作困境中实现更高收益的互惠合作。

Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas

论文配图:Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas
图 1 · 摘自论文原文
  • 设计融合利他与公平偏好的新效用函数,驱动合作行为
  • 在两个序列社会困境游戏中,合作收益和公平性均优于基线方法
  • 利他促进公共品贡献,公平促成双方互惠,机制清晰可解释

在多智能体强化学习中,如何在社会困境场景下诱导分布式智能体合作仍是难题。此类情境下个体利益与集体最优相悖,理性行为常导致次优群体结果。人类却能在类似情境中实现合作,常归因于社会偏好。为此,本文借鉴社会心理学与行为经济学中的利他与公平偏好,提出一种新型效用函数——利他与公平偏好(AFP),通过奖励共享机制将自身与他人奖励转化为合作激励。在两个具有挑战性的序列社会困境游戏上,与标准强化学习及不平等厌恶智能体对比实验表明,AFP智能体实现了更优的集体收益与更高公平性。进一步分析显示,利他偏好促使智能体主动贡献公共品,公平偏好则激发智能体间的相互协作行为。

原文摘要 · Abstract (English)

Inducing cooperation among distributed agents is still a difficult problem in the field of multi-agent reinforcement learning (MARL), particularly in social dilemma situations. There, individual interests are misaligned with the common good and individual rationality leads to suboptimal group outcomes. In contrast, humans are able to achieve cooperation with one another in such situations. A common explanation for such cooperative behavior is that individuals have social preferences. In order to achieve cooperation in MARL, we design a new utility function integrating altruistic preferences (incentive for other's reward) and fairness preferences (incentive for equality) from social psychology and behavioral economics, namely, Altruistic and Fairness Preference (AFP), a reward-sharing mechanism which converts one's own and other's rewards to incentives for cooperative behavior. We performed comparative experiments with standard RL and inequity aversion agents in two challenging sequential social dilemma games, and showed that AFP agents successfully achieved mutual cooperation with more collective rewards and higher equity than the baselines. To further understand the progression of AFP during training, we subsequently explore the effects of altruistic preferences and fairness preferences on agents' behavior. The results suggest that altruistic preferences encourage agents to contribute to the public goods, and fairness preferences induce mutual behavior between agents.

多智能体合作博弈社会偏好强化学习

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。