首个公开的智能体系统数据库,记录其能力、应用与安全措施。
The AI Agent Index
- 构建首个公开智能体系统数据库,收录技术组件与应用领域
- 发现开发者多披露功能应用,但安全机制信息不足
- 适合关注AI安全、开发透明度的研究者与从业者
领先的AI开发者和初创公司正越来越多地部署能够自主规划并执行复杂任务的智能体系统,人类干预极少。然而,目前尚无系统框架来记录智能体系统的技术构成、应用场景及安全特性。为此,我们推出了首个公开的智能体系统数据库——AI Agent Index。对符合纳入标准的每个系统,我们基于公开信息及与开发者的沟通,记录其组件(如基础模型、推理实现、工具使用)、应用领域(如计算机操作、软件工程)以及风险管控实践(如评估结果、防护机制)。研究发现,尽管开发者普遍提供丰富的功能与应用信息,但在安全与风险管理方面的披露仍十分有限。该数据库已上线:https://aiagentindex.mit.edu/
原文摘要 · Abstract (English)
Leading AI developers and startups are increasingly deploying agentic AI systems that can plan and execute complex tasks with limited human involvement. However, there is currently no structured framework for documenting the technical components, intended uses, and safety features of agentic systems. To fill this gap, we introduce the AI Agent Index, the first public database to document information about currently deployed agentic AI systems. For each system that meets the criteria for inclusion in the index, we document the system's components (e.g., base model, reasoning implementation, tool use), application domains (e.g., computer use, software engineering), and risk management practices (e.g., evaluation results, guardrails), based on publicly available information and correspondence with developers. We find that while developers generally provide ample information regarding the capabilities and applications of agentic systems, they currently provide limited information regarding safety and risk management practices. The AI Agent Index is available online at https://aiagentindex.mit.edu/
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。