arXiv:2605.06651cs.AI2026-05被引 10

AI助手助数学家探索难题,能自动推理论证并发现新方向。

AI co-mathematician: Accelerating mathematicians with agentic AI

论文配图:AI co-mathematician: Accelerating mathematicians with agentic AI
图 1 · 摘自论文原文
  • 构建可交互的智能工作台,支持数学研究全流程
  • 在前沿数学测试中达48%正确率,创AI新纪录
  • 适合科研人员、数学爱好者及人工智能辅助研究者

我们提出AI co-mathematician,一个供数学家与AI代理协作进行开放式研究的工作平台。该系统针对数学研究的探索性与迭代性特点,提供从想法生成、文献检索、计算探索、定理证明到理论构建的全流程支持。通过异步、有状态的工作空间,系统能够管理不确定性、细化用户意图、追踪失败假设,并输出原生数学成果,模拟人类协作流程。初步测试显示,该系统帮助研究人员解决开放问题、发现新研究方向、识别被忽视的文献。除展示高度交互的AI辅助数学发现范式外,其在硬核问题求解基准上达到领先水平,于FrontierMath Tier 4得分48%,是当前所有评估AI系统中的最高分。

原文摘要 · Abstract (English)

We introduce the AI co-mathematician, a workbench for mathematicians to interactively leverage AI agents to pursue open-ended research. The AI co-mathematician is optimized to provide holistic support for the exploratory and iterative reality of mathematical workflows, including ideation, literature search, computational exploration, theorem proving and theory building. By providing an asynchronous, stateful workspace that manages uncertainty, refines user intent, tracks failed hypotheses, and outputs native mathematical artifacts, the system mirrors human collaborative workflows. In early tests, the AI co-mathematician helped researchers solve open problems, identify new research directions, and uncover overlooked literature references. Besides demonstrating a highly interactive paradigm for AI-assisted mathematical discovery, the AI co-mathematician also achieves state of the art results on hard problem-solving benchmarks, including scoring 48% on FrontierMath Tier 4, a new high score among all AI systems evaluated.

AI数学智能助手定理证明

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。