AI具身记忆可能带来新风险,需提前研究应对
Episodic memory in AI agents poses risks that should be studied and mitigated
- 给AI添加可回溯的事件记忆能力
- 记忆功能虽提升可控性,但引出新安全隐患
- 提出四条原则指导安全开发
当前多数AI模型缺乏存储与回溯自身行为记录的能力。在人类认知中,情景记忆对回忆过去和规划未来至关重要。赋予AI类似能力将显著提升其在现实世界交互中的表现。尽管这有助于用户更好监控、理解与控制AI行为,但作为一项广泛应用的新能力,也将引入显著新风险。本文分析潜在风险与收益,提出四项指导原则,以确保情景记忆技术增强而非削弱AI的安全性与可信度。
原文摘要 · Abstract (English)
Most current AI models have little ability to store and later retrieve a record or representation of what they do. In human cognition, episodic memories play an important role in both recall of the past as well as planning for the future. The ability to form and use episodic memories would similarly enable a broad range of improved capabilities in an AI agent that interacts with and takes actions in the world. Researchers have begun directing more attention to developing memory abilities in AI models. It is therefore likely that models with such capability will be become widespread in the near future. This could in some ways contribute to making such AI agents safer by enabling users to better monitor, understand, and control their actions. However, as a new capability with wide applications, we argue that it will also introduce significant new risks that researchers should begin to study and address. We outline these risks and benefits and propose four principles to guide the development of episodic memory capabilities so that these will enhance, rather than undermine, the effort to keep AI safe and trustworthy.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。