arXiv:2605.31330cs.GTcs.AI2026-05

研究制度激励如何最大化社会总收益,发现最优激励有明确规律。

Social welfare optimisation under institutional reward and punishment

论文配图:Social welfare optimisation under institutional reward and punishment
图 1 · 摘自论文原文
  • 从社会福利角度设计奖惩机制,分析其在合作博弈中的表现
  • 发现激励效率与选择强度共同决定福利峰值,存在多极优化点
  • 证明最优激励要么为零,要么集中在特定公式解附近,适合政策制定者

制度性激励被广泛用于促进自主个体间的合作,涵盖人类社会到多智能体与人工智能系统。现有研究通常将激励设计视为双目标问题:最小化制度成本并实现高频率合作。但此类方案是否真正最大化社会福利——即总群体收益减去制度支出——仍缺乏深入探讨。本文针对有限、完全混合群体中进行社会困境博弈(捐赠博弈与公共品博弈)的情形,同时考虑对合作者的奖励和对背叛者的惩罚,构建以社会福利为核心的激励框架。对每种机制,推导出期望社会福利的显式表达式,并刻画其随激励效率与选择强度的变化规律。理论分析揭示:在某些参数区间内,社会福利仅有一个最优激励水平;而在另一些区间则出现定性相变,导致福利非单调且存在多个局部最优。进一步证明,任何福利最大化的激励要么为零,要么集中于一个简单的闭式目标值,并提供高效算法求解该最优值。通过比较奖励与惩罚机制,我们推导出在任意给定预算下,奖励优于惩罚的闭式条件。总体而言,结果揭示了以成本或合作频率为目标的激励设计与真正最大化社会福利之间存在系统性差距。

原文摘要 · Abstract (English)

Institutional incentives are widely used to promote cooperation among autonomous, self-regarding agents, from human societies to multi-agent and AI systems. Existing work typically treats incentive design as a bi-objective problem: minimise institutional cost while achieving a high long-run frequency of cooperation. Whether such schemes also maximise social welfare - total population payoff net of institutional expenditure - has remained largely unexplored. We develop a welfare-centric framework for institutional incentives in finite, well-mixed populations playing a social dilemma (Donation Game and Public Goods Game), considering both rewards for cooperators and punishments for defectors. For each mechanism, we derive explicit expressions for expected social welfare and characterise how it depends on incentive efficiency and selection intensity. Analytically, we identify parameter regimes where social welfare has a single optimal incentive level and regimes with qualitative phase transitions, in which welfare becomes non-monotonic with multiple local optima. We prove that any welfare-maximising incentive is either zero or concentrated around a simple closed-form target, and we provide an efficient algorithm to compute these optima. Comparing reward and punishment, we further derive close-formed conditions under which reward outperform punishment in terms of social welfare for any given budget. Overall, our results reveal a systematic gap between incentives optimised for cost or cooperation frequency and those that maximise welfare.

制度激励社会福利博弈论多智能体

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。