arXiv:2606.00804cs.MAcs.AI2026-06

企业多智能体系统应按任务类型动态选择协作策略,而非固定模式。

Dynamic Coordination Strategy Selection for Enterprise Multi-Agent Systems

  • 根据任务类别动态切换协作方式,而非全局固定策略。
  • 预测策略始终在最优方案0.10分内,表现稳定可靠。
  • 适合需要灵活协作的企业级AI系统设计者参考。

企业多智能体系统面临多种协作模式,但部署时缺乏依据判断何时使用共识、辩论、融合或简单单智能体流程。本文评估是否应基于问题类别动态选择策略而非全局固定。实验涵盖30项跨六行业任务,五类问题、四种执行条件,每组重复三次,共四类模型:qwen_local、sonnet、gemma_openrouter及辅助openai云验证臂,总计1,440份输出由固定Sonnet评分标准评判。主要发现为有界且可操作,但不支持原假设H1:精确胜者身份在不同模型间不稳定,部分预测策略仅接近而非超越最优替代方案。弱化后的近最优路由主张被强烈支持:所有预注册模型臂与问题类别中,预测策略均在0.10分内接近最佳表现;结构合规性验证是唯一例外,各模型均偏好single_agent而非consensus。预注册的Kendall's W检验显示越南语与英语任务在协作条件排名一致性上无显著差异(两组均值W=0.20,符号秩检验p=.85),故H2不成立。结论:企业协作策略应以动态路由为校准默认,而非确定性胜者选择法则。

原文摘要 · Abstract (English)

Enterprise multi-agent systems increasingly expose multiple coordination patterns, but deployments often lack evidence for when to use consensus, debate, synthesis, or a simpler single-agent workflow. This paper evaluates whether coordination strategy should be selected dynamically by problem class rather than fixed globally. We run a frozen matrix of 30 enterprise tasks spanning six industries, five problem classes, four execution conditions, three replications per cell, and four model arms: qwen_local, sonnet, gemma_openrouter, and an auxiliary openai cloud-validation arm. All 1,440 generated outputs are judged by a fixed Sonnet rubric. The main finding is bounded and operationally useful, but it is not the original strict H1. The pre-registered exact-winner/CI criterion is not supported: exact winner identity is unstable across model arms, and several predicted strategies are close to, but not above, the best observed alternative. A weaker near-best routing claim is strongly supported. In every pre-registered model arm and problem class, and again in the auxiliary OpenAI validation arm, the predicted strategy is within 0.10 quality-score points of the best observed condition. Structured compliance verification is the clearest exception to the original mapping: all arms favor single_agent rather than consensus. A pre-registered Kendall's W test finds no reliable difference between Vietnamese-domain and English-domain tasks in how consistently the four coordination conditions are ranked (mean W of 0.20 in both strata; signed-rank p = .85), so H2 is not supported. We conclude that enterprise coordination policy should use dynamic routing as a calibrated default, not as a deterministic winner-selection law.

多智能体动态路由企业AI协作策略

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。