用纯文本生成多格日漫,支持布局可控与角色一致
Manga Generation via Layout-controllable Diffusion
- 基于扩散模型构建跨面板信息交互机制
- 生成结果保持面板数与多样布局合理
- 适合漫画创作、故事可视化应用
纯文本生成漫画研究广泛,但基于纯文本生成多格日漫(Manga)的研究较少。日漫具有页面内多面板、叙事连贯、布局合理多样、角色一致及画面与文字语义对应等特征,生成难度大。本文提出漫画生成任务,并构建Manga109Story数据集以支持纯文本生成研究。同时提出MangaDiffusion模型,促进生成过程中的面板内与面板间信息交互。实验表明,该方法能有效保证面板数量、生成合理且多样的页面布局。基于此方法,有望将大量文本故事转化为更具吸引力的漫画阅读形式,具备广泛应用前景。
原文摘要 · Abstract (English)
Generating comics through text is widely studied. However, there are few studies on generating multi-panel Manga (Japanese comics) solely based on plain text. Japanese manga contains multiple panels on a single page, with characteristics such as coherence in storytelling, reasonable and diverse page layouts, consistency in characters, and semantic correspondence between panel drawings and panel scripts. Therefore, generating manga poses a significant challenge. This paper presents the manga generation task and constructs the Manga109Story dataset for studying manga generation solely from plain text. Additionally, we propose MangaDiffusion to facilitate the intra-panel and inter-panel information interaction during the manga generation process. The results show that our method particularly ensures the number of panels, reasonable and diverse page layouts. Based on our approach, there is potential to converting a large amount of textual stories into more engaging manga readings, leading to significant application prospects.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。