从大脑启发到安全协作,系统梳理智能体的演进与挑战
Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems

- 构建类脑模块化架构,整合记忆、目标、情感等认知组件
- 提出自主进化机制,支持持续学习与环境自适应能力
- 聚焦多智能体协同与安全对齐,适合研究具身智能与可信AI者
大型语言模型的兴起推动了人工智能的变革,催生出具备复杂推理、强感知与多功能行动能力的先进智能体。随着智能体在科研与应用中日益关键,其设计、评估与持续改进面临多重挑战。本书系统性地将智能体置于模块化、类脑架构中,融合认知科学、神经科学与计算研究思想。内容分为四部分:第一,系统映射智能体的认知、感知与操作模块至人类大脑功能,阐明记忆、世界建模、奖励处理、目标与情绪等核心组件;第二,探讨自我增强与自适应演化机制,实现智能体在动态环境中的持续学习与自动优化;第三,研究多智能体系统的集体智能,包括交互、合作与社会结构;第四,强调安全与有益性,涵盖内在与外在威胁、伦理对齐、鲁棒性及实际缓解策略,以保障可信部署。通过跨学科整合,本文识别关键挑战与机遇,倡导技术进步与社会价值的协同发展。
原文摘要 · Abstract (English)
The advent of large language models (LLMs) has catalyzed a transformative shift in artificial intelligence, paving the way for advanced intelligent agents capable of sophisticated reasoning, robust perception, and versatile action across diverse domains. As these agents increasingly drive AI research and practical applications, their design, evaluation, and continuous improvement present intricate, multifaceted challenges. This book provides a comprehensive overview, framing intelligent agents within modular, brain-inspired architectures that integrate principles from cognitive science, neuroscience, and computational research. We structure our exploration into four interconnected parts. First, we systematically investigate the modular foundation of intelligent agents, systematically mapping their cognitive, perceptual, and operational modules onto analogous human brain functionalities and elucidating core components such as memory, world modeling, reward processing, goal, and emotion. Second, we discuss self-enhancement and adaptive evolution mechanisms, exploring how agents autonomously refine their capabilities, adapt to dynamic environments, and achieve continual learning through automated optimization paradigms. Third, we examine multi-agent systems, investigating the collective intelligence emerging from agent interactions, cooperation, and societal structures. Finally, we address the critical imperative of building safe and beneficial AI systems, emphasizing intrinsic and extrinsic security threats, ethical alignment, robustness, and practical mitigation strategies necessary for trustworthy real-world deployment. By synthesizing modular AI architectures with insights from different disciplines, this survey identifies key research challenges and opportunities, encouraging innovations that harmonize technological advancement with meaningful societal benefit.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。