提出六级智能度框架,明确GUI代理的自主能力边界。
How Smart Is Your GUI Agent? A Framework for the Future of Software Interaction
- 构建六级自主性分级体系,定义代理能力范围
- 统一评估标准,便于比较不同代理性能
- 助力可信软件交互发展,适合研究者与开发者参考
GUI代理正迅速成为人机交互的新范式,使用户无需逐点击操作即可完成网页、桌面和移动应用的导航。然而,当前对“代理”的描述在自主程度上差异巨大,导致能力、责任与风险难以界定。本文提出GUI代理自主性等级(GAL)框架,涵盖六个层级,使自主性显式化,并为可信软件交互的发展提供基准衡量标准,推动该领域的规范化与可比性。
原文摘要 · Abstract (English)
GUI agents are rapidly becoming a new interaction to software, allowing people to navigate web, desktop and mobile rather than execute them click by click. Yet ``agent'' is described with radically different degrees of autonomy, obscuring capability, responsibility and risk. We call for conceptual clarity through GUI Agent Autonomy Levels (GAL), a six-level framework that makes autonomy explicit and helps benchmark progress toward trustworthy software interaction.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。