arXiv:2602.03828cs.AIcs.CL2026-02中稿 · ICLR被引 24

自动生成学术论文级插图,让科研绘图告别手动繁琐。

AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations

  • 构建智能流程,通过思考、重组与验证生成高质量插图
  • 在3300对图文数据上超越所有基线方法,产出可直接发表的插图
  • 适合需要高效制图的科研人员和学术出版团队

高质量科学插图对于有效传达复杂科学概念至关重要,但其手工制作在学术界和工业界仍是一个公认的瓶颈。我们提出了FigureBench,首个基于长篇科学文本生成科学插图的大规模基准数据集,包含3,300对高质量的文本-插图配对,覆盖来自科学论文、综述、博客和教科书的多样化图文任务。此外,我们提出AutoFigure,首个能基于长篇科学文本自动生成高质量科学插图的代理式框架。该框架在生成最终图像前,进行深度思考、内容重组与验证,确保布局结构严谨且视觉美观,输出兼具结构完整性和美学吸引力的科学插图。利用FigureBench的高质量数据,我们开展了大量实验,对比AutoFigure与多种基线方法的表现。结果表明,AutoFigure始终优于所有基线方法,能够生成可直接发表的科学插图。代码、数据集及Hugging Face空间已开源:https://github.com/ResearAI/AutoFigure。

原文摘要 · Abstract (English)

High-quality scientific illustrations are crucial for effectively communicating complex scientific and technical concepts, yet their manual creation remains a well-recognized bottleneck in both academia and industry. We present FigureBench, the first large-scale benchmark for generating scientific illustrations from long-form scientific texts. It contains 3,300 high-quality scientific text-figure pairs, covering diverse text-to-illustration tasks from scientific papers, surveys, blogs, and textbooks. Moreover, we propose AutoFigure, the first agentic framework that automatically generates high-quality scientific illustrations based on long-form scientific text. Specifically, before rendering the final result, AutoFigure engages in extensive thinking, recombination, and validation to produce a layout that is both structurally sound and aesthetically refined, outputting a scientific illustration that achieves both structural completeness and aesthetic appeal. Leveraging the high-quality data from FigureBench, we conduct extensive experiments to test the performance of AutoFigure against various baseline methods. The results demonstrate that AutoFigure consistently surpasses all baseline methods, producing publication-ready scientific illustrations. The code, dataset and huggingface space are released in https://github.com/ResearAI/AutoFigure.

科学绘图自动化图文生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。