让论文自动生成可编辑的海报,还能自由调整设计风格。
PosterMELD: Multi-Agent Paper-to-Poster Generation for Controllable Design Diversity with Editable Print-Ready Outputs

- 用多智能体系统按模板生成海报,先写后渲染,失败可修复。
- 81.3%的请求通过打印就绪检测,是现有方法的3.4到5.2倍。
- 输出可直接编辑的PPTX和PNG文件,适合科研人员快速定制海报。
科学海报需将长篇多模态论文浓缩为可读、可编辑的视觉呈现。现有系统仅评估完整输出,隐藏请求级失败;直接图像生成不可编辑,编码代理流程成本高。PosterMELD是一种模板引导的多智能体流程:容量感知槽位指导内容撰写,确定性门控与视觉语言模型(VLM)审查实现故障定向修复。每个成功请求输出可编辑的PowerPoint(PPTX)和便携式网络图形(PNG)文件;显式设计控制支持同一篇论文的多种变体。在621篇论文上,打印就绪率(PRR)统计通过几何布局、可读性、资产完整性及明显事实错误检查的请求数,独立报告原生可编辑性。冻结的VLM对打印就绪输出赋予条件性的工艺-和谐-表现力(CHE)评分。PosterMELD达成81.3%的PRR,是P2P的3.4倍、PosterGen的5.2倍,且在生成方法中拥有最高条件CHE评分。原生可编辑性与显式设计控制的平均成本为每请求0.38美元,仅为Codex+Skill的3.5%。代码与资源见https://github.com/Shannon4Science/PosterMELD。
原文摘要 · Abstract (English)
Scientific poster construction compresses a long multimodal paper into a readable, editable canvas. Existing systems hide request-level failures by scoring only completed outputs; direct image generation is not element-editable, while coding-agent workflows are costly. PosterMELD is a template-conditioned multi-agent pipeline: capacity-aware slots guide writing before rendering, and deterministic gates plus vision-language model (VLM) review route failures to bounded repair. Each accepted request exports editable PowerPoint (PPTX) and Portable Network Graphics (PNG) artifacts; explicit design controls yield same-paper variants. Across 621 papers, Print-Ready Rate (PRR) counts requests passing geometric, readability, asset-integrity, and obvious-factual-error checks, with native editability reported separately. A frozen VLM assigns conditional Craftsmanship-Harmony-Expressiveness (CHE) scores to print-ready outputs. PosterMELD attains 81.3% PRR, 3.4 times P2P's rate and 5.2 times PosterGen's, and the highest conditional CHE among generated methods with multiple print-ready outputs. Native editability and explicit design controls are retained at a mean cost of USD 0.38 per request, 3.5% of Codex+Skill's. Code and resources are available at https://github.com/Shannon4Science/PosterMELD.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。