基于真实谈判数据构建多方协商基准,支持分阶段承诺评估。
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
- 用可配置生成器与真实气候谈判文档构造协商游戏
- 小规模游戏精确评估,大规模游戏对比表现无统一优解
- 适合研究动态协商、强化学习与博弈策略的学者
许多现实中的多方协商以一系列具有约束力的行动级承诺逐步展开,而非单一最终结果,但现有基准对此类场景研究不足。本文提出一个基准与评估框架,结合可配置的协商游戏生成器与来自气候谈判实践的文档驱动实例,并提供多个基线求解器。在小规模游戏上进行精确评估,在大规模实例上进行对比评估,结果显示无一求解器在所有情境下占优,性能取决于游戏的结构特性。该发现推动了对能稳健处理部分承诺的新协商方法的研究。代码与数据已公开于:https://anonymous.4open.science/r/negotiation_MARL-46B8
原文摘要 · Abstract (English)
Many real-world multi-party negotiations unfold as sequences of binding, action-level commitments rather than a single final outcome, yet this regime remains under-studied in existing benchmarks. We introduce a benchmark and evaluation framework for this setting, combining a configurable negotiation game generator with document-grounded instances derived from a climate negotiation exercise. We also provide several baseline solvers. Exact evaluation on small games and comparative evaluation on larger instances show that no solver dominates across regimes; performance depends on the structural properties of the game. These results motivate the creation of novel negotiation methods that value partial commitments robustly across diverse strategic regimes. Code and data for the benchmark are available at: https://anonymous.4open.science/r/negotiation_MARL-46B8
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。