生成连贯的过肩对话视频,解决角色一致性与空间连续性难题。
ShoulderShot: Generating Over-the-Shoulder Dialogue Videos
- 采用双镜头生成加循环视频技术,支持长对话
- 在多镜头布局和空间连贯性上优于现有方法
- 适合影视、广告等需要自然对话场景的创作者
过肩对话视频在电影、短剧和广告中至关重要,能提升视觉多样性和观众情感共鸣。然而,这类对话场景在视频生成研究中仍被严重忽视。主要挑战包括跨镜头角色一致性维持、空间连续性构建,以及在有限计算资源下生成长时序多轮对话。本文提出 ShoulderShot 框架,结合双镜头生成与循环视频机制,实现持续对话生成的同时保持角色一致性和空间连贯性。实验结果表明,该方法在镜头切换布局、空间连续性及对话长度灵活性方面均超越现有技术,为实际对话视频生成开辟新路径。视频与对比结果见 https://shouldershot.github.io。
原文摘要 · Abstract (English)
Over-the-shoulder dialogue videos are essential in films, short dramas, and advertisements, providing visual variety and enhancing viewers' emotional connection. Despite their importance, such dialogue scenes remain largely underexplored in video generation research. The main challenges include maintaining character consistency across different shots, creating a sense of spatial continuity, and generating long, multi-turn dialogues within limited computational budgets. Here, we present ShoulderShot, a framework that combines dual-shot generation with looping video, enabling extended dialogues while preserving character consistency. Our results demonstrate capabilities that surpass existing methods in terms of shot-reverse-shot layout, spatial continuity, and flexibility in dialogue length, thereby opening up new possibilities for practical dialogue video generation. Videos and comparisons are available at https://shouldershot.github.io.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。