用多智能体LLM自动生成符合设计美学的学术海报,减少人工修改。
PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs
- 四类专业智能体协同:内容提取、布局规划、视觉风格、最终渲染。
- 生成海报在内容准确率上持平,视觉设计显著优于现有方法。
- 适合需要快速制作高质量会议海报的研究人员使用。
基于大语言模型的多智能体系统在处理复杂组合任务方面表现出色。本文将此范式应用于论文转海报这一实际但耗时的任务,该任务常困扰准备会议的研究人员。尽管已有方法尝试自动化该过程,多数仍忽视核心设计与美学原则,导致生成的海报需大量人工修正。为此,我们提出PosterGen,一个模拟专业海报设计师工作流程的多智能体框架。其包含四个协作的专业智能体:(1) 解析与整理智能体从论文中提取内容并组织故事板;(2) 布局智能体将内容映射为连贯的空间布局;(3) 风格智能体应用色彩与排版等视觉元素;(4) 渲染智能体生成最终海报。这些智能体共同产出语义准确且视觉美观的海报。为评估设计质量,我们引入基于视觉-语言模型(VLM)的评分标准,衡量布局平衡性、可读性和美学一致性。实验表明,PosterGen在内容保真度上保持一致,且在视觉设计上显著优于现有方法,生成的海报基本无需人工修改即可用于展示。
原文摘要 · Abstract (English)
Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In this work, we apply this paradigm to the paper-to-poster generation problem, a practical yet time-consuming process faced by researchers preparing for conferences. While recent approaches have attempted to automate this task, most neglect core design and aesthetic principles, resulting in posters that require substantial manual refinement. To address these design limitations, we propose PosterGen, a multi-agent framework that mirrors the workflow of professional poster designers. It consists of four collaborative specialized agents: (1) Parser and Curator agents extract content from the paper and organize storyboard; (2) Layout agent maps the content into a coherent spatial layout; (3) Stylist agents apply visual design elements such as color and typography; and (4) Renderer composes the final poster. Together, these agents produce posters that are both semantically grounded and visually appealing. To evaluate design quality, we introduce a vision-language model (VLM)-based rubric that measures layout balance, readability, and aesthetic coherence. Experimental results show that PosterGen consistently matches in content fidelity, and significantly outperforms existing methods in visual designs, generating posters that are presentation-ready with minimal human refinements.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。