arXiv:2606.00826cs.LG2026-06AAAI被引 2

提出部分公平透明机制,让智能体逐步猜对系统真实公平标准。

Partial Fairness Awareness: Belief-Guided Strategic Mechanism for Strategic Agents

论文配图:Partial Fairness Awareness: Belief-Guided Strategic Mechanism for Strategic Agents
图 1 · 摘自论文原文
  • 用信念更新机制让代理在互动中推测系统公平规则
  • 相比完全公开或隐藏规则,显著降低群体公平差距
  • 适合关注公平性与策略行为平衡的研究者

战略机器学习研究代理通过操纵特征以获取有利决策的场景。为应对战略分类中的公平性问题,现有方法引入了群体特定的公平约束。然而,当前公平感知方法面临公平暴露的根本困境:公开约束会引发策略操纵导致公平反转,而隐藏约束则可能降低社会福利并抑制真实改进。为此,我们提出部分公平意识(PFA)问题,理论分析表明,通过发布公平约束候选集但隐藏实际使用约束可缓解该困境。具体地,我们引入信念引导的战略机制,代理与决策系统迭代交互,并在候选约束集中维护信念分布。该信念引导过程使代理通过持续反馈逐步更新信念,最终逼近系统所采用的真实公平约束。在真实和合成数据集上的大量实验表明,相较于完全公开或私有公平制度,PFA实现了更低的群体公平差距、更高的合格个体接受率以及更稳定的预测结果。

原文摘要 · Abstract (English)

Strategic machine learning investigates scenarios where agents manipulate their features to receive favorable decisions from predictive models. To address fairness concerns intrinsic to strategic classification, recent work has introduced group-specific fairness constraints. However, current fairness-aware approaches face a fundamental dilemma in the issue of fairness exposure: making these constraints public enables strategic manipulation and can lead to fairness reversal, while keeping them hidden may reduce social welfare and discourage genuine improvement. To fill this gap, we subsequently propose the problem of partial fairness awareness (PFA), as our theoretical analysis informs that such a dilemma can be mitigated by releasing the candidate set of fairness constraints and concealing the grounding constraint. To be specific, we introduce a belief-guided strategic mechanism, wherein agents iteratively interact with the decision system and maintain a belief distribution over the candidate set of fairness constraints. This belief-guided process enables agents, through iterative interaction and feedback, to update their belief distribution over the candidate set, thereby gradually aligning their belief with the grounding fairness constraint employed by the system. Extensive experiments on real-world and synthetic datasets demonstrate that PFA achieves lower group fairness gaps, higher acceptance of truly qualified individuals, and more stable outcomes compared to fully public or private fairness regimes.

公平性策略学习信念机制

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。