PPTAgent用编辑思路生成更高质量的演示文稿,兼顾内容、设计与逻辑。
PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides
- 分两阶段生成:先分析参考稿提取结构,再迭代生成编辑动作
- 在内容、设计、连贯性三方面均显著优于现有方法
- 适合需要高效制作专业幻灯片的研究者和职场人士
从文档自动生成演示文稿是一项挑战性任务,需兼顾内容质量、视觉吸引力和结构连贯性。现有方法主要孤立提升内容质量,忽视视觉美感与结构一致性,限制了实际应用。为此,我们提出PPTAgent,采用受人类工作流程启发的两阶段、基于编辑的方法,全面改进演示文稿生成。PPTAgent首先分析参考演示文稿,提取每页的功能类型与内容模板,再据此草拟大纲,并基于选定参考页迭代生成编辑指令以创建新幻灯片。为全面评估生成效果,我们进一步提出PPTEval评估框架,从内容、设计、连贯性三个维度进行评价。实验结果表明,PPTAgent在三项指标上均显著优于现有自动演示文稿生成方法。
原文摘要 · Abstract (English)
Automatically generating presentations from documents is a challenging task that requires accommodating content quality, visual appeal, and structural coherence. Existing methods primarily focus on improving and evaluating the content quality in isolation, overlooking visual appeal and structural coherence, which limits their practical applicability. To address these limitations, we propose PPTAgent, which comprehensively improves presentation generation through a two-stage, edit-based approach inspired by human workflows. PPTAgent first analyzes reference presentations to extract slide-level functional types and content schemas, then drafts an outline and iteratively generates editing actions based on selected reference slides to create new slides. To comprehensively evaluate the quality of generated presentations, we further introduce PPTEval, an evaluation framework that assesses presentations across three dimensions: Content, Design, and Coherence. Results demonstrate that PPTAgent significantly outperforms existing automatic presentation generation methods across all three dimensions.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。