arXiv:2603.11768cs.AI2026-03被引 14

提出SSGM框架,解决大模型代理记忆动态演化中的安全与稳定问题。

Governing Evolving Memory in LLM Agents: Risks, Mechanisms, and the Stability and Safety Governed Memory (SSGM) Framework

  • 通过一致性验证、时间衰减建模和动态访问控制治理记忆演化
  • 可缓解敏感信息固化导致的知识泄露,防止知识因迭代总结而退化
  • 适合关注大模型长期记忆安全与可靠性的研究者和开发者

长期记忆已成为自主大型语言模型(LLM)代理的基础组件,支持持续适应、终身多模态学习和复杂推理。然而,当记忆系统从静态检索数据库转向动态、自主的机制时,记忆治理、语义漂移和隐私漏洞等关键问题浮现。尽管近期综述聚焦于记忆检索效率,却普遍忽视了高度动态环境中记忆污染的新兴风险。为此,我们提出稳定性与安全性治理记忆(SSGM)框架,一种概念性治理架构。SSGM通过在记忆固化前强制执行一致性验证、时间衰减建模和动态访问控制,将记忆演化与执行解耦。通过形式化分析与架构分解,我们证明了SSGM可缓解拓扑引发的知识泄露问题——即敏感上下文被固化至长期存储;同时有助于防止语义漂移——即知识通过迭代摘要逐渐退化。本工作最终构建了记忆污染风险的全面分类体系,并确立了部署安全、持久、可靠代理记忆系统的稳健治理范式。

原文摘要 · Abstract (English)

Long-term memory has emerged as a foundational component of autonomous Large Language Model (LLM) agents, enabling continuous adaptation, lifelong multimodal learning, and sophisticated reasoning. However, as memory systems transition from static retrieval databases to dynamic, agentic mechanisms, critical concerns regarding memory governance, semantic drift, and privacy vulnerabilities have surfaced. While recent surveys have focused extensively on memory retrieval efficiency, they largely overlook the emergent risks of memory corruption in highly dynamic environments. To address these emerging challenges, we propose the Stability and Safety-Governed Memory (SSGM) framework, a conceptual governance architecture. SSGM decouples memory evolution from execution by enforcing consistency verification, temporal decay modeling, and dynamic access control prior to any memory consolidation. Through formal analysis and architectural decomposition, we show how SSGM can mitigate topology-induced knowledge leakage where sensitive contexts are solidified into long-term storage, and help prevent semantic drift where knowledge degrades through iterative summarization. Ultimately, this work provides a comprehensive taxonomy of memory corruption risks and establishes a robust governance paradigm for deploying safe, persistent, and reliable agentic memory systems.

大模型代理记忆治理安全框架

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。