大模型时代查询扩展方法全解析,揭示新范式与实用权衡
Query Expansion in the Age of Pre-trained and Large Language Models: A Comprehensive Survey
- 基于大模型的上下文感知与指令跟随能力实现精准扩展
- 提出四维设计框架:注入位置、语料对齐、学习方式、知识融合
- 适合检索系统开发者参考,尤其关注效果与成本平衡
现代信息检索需应对短而模糊的查询与日益多样动态的语料库。查询扩展(QE)仍是缓解词汇不匹配的核心技术,但预训练模型和大语言模型(PLMs/LLMs)重塑了其设计空间。本文综述了大模型时代的QE方法,提供统一视角以理解新兴格局。首先总结不同模型家族如何催生新的扩展行为,包括更强的上下文理解、更可控的生成和指令遵循能力。接着沿四个互补维度组织近期技术:扩展在流程中的注入位置、与语料证据的锚定与交互方式、学习或对齐机制,以及结构化知识(如知识图谱)的融合方式。超越分类,我们提炼代表性检索场景中的应用模式与部署考量,强调有效性、可控性、锚定质量与运行成本之间的实际权衡。最后,指出未来在真实约束下实现更可靠、安全、高效且持续自适应的QE所面临的开放挑战与方向。
原文摘要 · Abstract (English)
Modern information retrieval must reconcile short, ambiguous queries with increasingly diverse and dynamic corpora. Query expansion (QE) remains a core technique for mitigating vocabulary mismatch, but its design space has been reshaped by pre-trained and large language models (PLMs/LLMs). This survey reviews QE methods in the PLM/LLM era and provides a unified view of the emerging landscape. We first summarize how different model families enable new expansion behaviors, including stronger contextualization, more controllable generation, and instruction-following. We then organize recent techniques along four complementary design dimensions: where expansion is injected in the pipeline, how it is grounded and interacts with corpus evidence, how it is learned or aligned, and how structured knowledge such as knowledge graphs is incorporated. Beyond taxonomy, we synthesize application patterns and deployment considerations across representative retrieval settings, highlighting practical trade-offs among effectiveness, controllability, grounding quality, and operating cost. Finally, we outline open challenges and future directions toward more reliable, safe, efficient, and continually adaptive QE under real-world constraints.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。