arXiv:2603.27490cs.CLcs.AI2026-03被引 9

动态调整网页搜索中上下文管理策略,提升长时序任务效率与精度

AgentSwing: Adaptive Parallel Context Management Routing for Long-Horizon Web Agents

  • 基于状态感知的并行分支路由,动态选择最优上下文管理路径
  • 在多个基准上减少3倍交互轮次,同时提升最终任务完成率
  • 适合需要长期规划的智能代理研究者与开发者参考

随着大语言模型演变为执行长时序信息检索的自主代理,有限的上下文容量已成为关键瓶颈。现有方法通常在整个任务过程中采用单一固定策略,但此类静态设计在不同阶段的表现可能不均衡,难以适应累积上下文的效用与可靠性变化。为此,我们提出一个概率框架,从搜索效率和终端精度两个互补维度定义长时序成功。在此基础上,提出AgentSwing:一种状态感知的自适应并行上下文管理路由框架。在每个触发点,AgentSwing并行展开多个上下文管理分支,并通过前瞻路由选择最有望的延续路径。在多种基准与代理主干上的实验表明,AgentSwing始终优于强基线静态管理方法,常以最多减少3倍交互轮次达到或超越其性能上限,同时提升长时序网页代理的最终表现天花板。除实证收益外,该概率框架为未来长时序代理的上下文管理策略分析与设计提供了原则性视角。

原文摘要 · Abstract (English)

As large language models (LLMs) evolve into autonomous agents for long-horizon information-seeking, managing finite context capacity has become a critical bottleneck. Existing context management methods typically commit to a single fixed strategy throughout the entire trajectory. Such static designs may work well in some states, but they cannot adapt as the usefulness and reliability of the accumulated context evolve during long-horizon search. To formalize this challenge, we introduce a probabilistic framework that characterizes long-horizon success through two complementary dimensions: search efficiency and terminal precision. Building on this perspective, we propose AgentSwing, a state-aware adaptive parallel context management routing framework. At each trigger point, AgentSwing expands multiple context-managed branches in parallel and uses lookahead routing to select the most promising continuation. Experiments across diverse benchmarks and agent backbones show that AgentSwing consistently outperforms strong static context management methods, often matching or exceeding their performance with up to $3\times$ fewer interaction turns while also improving the ultimate performance ceiling of long-horizon web agents. Beyond the empirical gains, the proposed probabilistic framework provides a principled lens for analyzing and designing future context management strategies for long-horizon agents.

长时序代理上下文管理自适应路由LLM应用

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。