用粗略3D场景+AI生成视频,快速制作电影预演样片
PrevizWhiz: Combining Rough 3D Scenes and 2D Video to Guide Generative Video Previsualization
- 结合粗糙3D场景与生成模型,实现风格化视频预演
- 支持帧级图像重绘、运动路径编辑,生成高保真视频片段
- 降低影视创作门槛,适合导演与动画师快速迭代创意
在影视前期制作中,导演与3D动画专家需快速原型化创意以探索影片可能性,但传统方法存在效率与表现力的权衡:手绘分镜缺乏复杂运镜所需的空间精度,而3D预演则依赖专业技能和高质量绑定资源。为此,我们提出PrevizWhiz系统,利用粗略3D场景结合生成式图像与视频模型,生成风格化视频预览。工作流包含帧级图像重绘(可调节相似度)、基于运动路径或外部视频输入的时间编辑,以及生成高保真视频片段的优化。对导演的实证研究表明,该系统降低了技术门槛,加速了创意迭代,有效弥合了沟通鸿沟,同时也揭示了连续性、作者权属及人工智能辅助创作中的伦理挑战。
原文摘要 · Abstract (English)
In pre-production, filmmakers and 3D animation experts must rapidly prototype ideas to explore a film's possibilities before fullscale production, yet conventional approaches involve trade-offs in efficiency and expressiveness. Hand-drawn storyboards often lack spatial precision needed for complex cinematography, while 3D previsualization demands expertise and high-quality rigged assets. To address this gap, we present PrevizWhiz, a system that leverages rough 3D scenes in combination with generative image and video models to create stylized video previews. The workflow integrates frame-level image restyling with adjustable resemblance, time-based editing through motion paths or external video inputs, and refinement into high-fidelity video clips. A study with filmmakers demonstrates that our system lowers technical barriers for film-makers, accelerates creative iteration, and effectively bridges the communication gap, while also surfacing challenges of continuity, authorship, and ethical consideration in AI-assisted filmmaking.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。