arXiv:2604.24468cs.CRcs.CL2026-04综述

让资源有限方安全微调大模型,避免数据外泄风险

A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations

论文配图:A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations
图 1 · 摘自论文原文
  • 将大模型拆分在客户端与服务器间协同训练
  • 支持低资源机构安全使用大模型,保护数据隐私
  • 系统梳理模型、效率、隐私三类关键技术进展

微调使大语言模型(LLMs)适用于特定场景,但其高昂的计算成本常使资源受限机构难以承担。尽管云平台可提供所需算力,但数据隐私担忧使向第三方共享敏感信息变得风险重重。一种有前景的解决方案是用于大模型微调的分割学习,该方法将模型在客户端与服务器之间拆分,通过交换中间数据实现协作且安全的训练,使资源受限参与者能安全地适配大模型。为此,相关研究不断涌现,涵盖多样化的模型方法、系统优化及隐私攻防技术。为厘清该领域发展方向,亟需一份全面综述以分类、比较并评析这些多元方法。本文首次系统性地调研了用于大模型微调的分割学习,提出统一且细粒度的训练流程,识别关键操作组件,并从模型级优化、系统级效率、隐私保护三个核心维度对前沿工作进行系统性回顾。通过这一结构化分类体系,为构建可扩展、鲁棒且安全的大模型协同适配奠定基础。

原文摘要 · Abstract (English)

Fine-tuning unlocks large language models (LLMs) for specialized applications, but its high computational cost often puts it out of reach for resource-constrained organizations. While cloud platforms could provide the needed resources, data privacy concerns make sharing sensitive information with third parties risky. A promising solution is split learning for LLM fine-tuning, which divides the model between clients and a server, allowing collaborative and secure training through exchanged intermediate data, thus enabling resource-constrained participants to adapt LLMs safely. % In light of this, a growing body of literature has emerged to advance this paradigm, introducing varied model methods, system optimizations, and privacy defense-attack techniques for split learning. To bring clarity and direction to the field, a comprehensive survey is needed to classify, compare, and critique these diverse approaches. This paper fills the gap by presenting the first extensive survey dedicated to split learning for LLM fine-tuning. We propose a unified, fine-grained training pipeline to pinpoint key operational components and conduct a systematic review of state-of-the-art work across three core dimensions: model-level optimization, system-level efficiency, and privacy preservation. Through this structured taxonomy, we establish a foundation for advancing scalable, robust, and secure collaborative LLM adaptation.

大模型微调分割学习隐私保护协同训练

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。