arXiv:2603.23246cs.CV2026-03

用3D模型控制扩散模型,实现任意视角和光照下的高质量物体渲染。

GO-Renderer: Generative Object Rendering with 3D-aware Controllable Video Diffusion Models

  • 用重建的3D代理引导扩散模型生成视频,实现精准视角控制。
  • 在新光照下生成逼真图像,无需显式建模材质与光照。
  • 适合需要高保真物体渲染与可控生成的研究者和开发者。

从图像重建可渲染的3D模型是一项有用但具有挑战性的任务。近期的前馈式3D重建方法在高效恢复几何结构方面取得显著进展,但仍难以准确建模复杂外观。基于扩散的生成模型可通过参考图像合成逼真的物体图像或视频,而无需显式建模其外观,为物体渲染提供了有前景的方向,但缺乏对视角的精确控制。本文提出GO-Renderer,一个统一框架,将重建的3D代理作为指导,驱动视频生成模型,在任意视角和任意光照条件下实现高质量物体渲染。该方法不仅利用3D代理实现精准视角控制,还能通过扩散生成模型在不同光照环境下生成高质量渲染结果,无需显式建模复杂材质与光照。大量实验表明,GO-Renderer在物体渲染任务中达到领先性能,包括在新视角合成图像、在新光照环境中渲染物体,以及将物体插入现有视频。

原文摘要 · Abstract (English)

Reconstructing a renderable 3D model from images is a useful but challenging task. Recent feedforward 3D reconstruction methods have demonstrated remarkable success in efficiently recovering geometry, but still cannot accurately model the complex appearances of these 3D reconstructed models. Recent diffusion-based generative models can synthesize realistic images or videos of an object using reference images without explicitly modeling its appearance, which provides a promising direction for object rendering, but lacks accurate control over the viewpoints. In this paper, we propose GO-Renderer, a unified framework integrating the reconstructed 3D proxies to guide the video generative models to achieve high-quality object rendering on arbitrary viewpoints under arbitrary lighting conditions. Our method not only enjoys the accurate viewpoint control using the reconstructed 3D proxy but also enables high-quality rendering in different lighting environments using diffusion generative models without explicitly modeling complex materials and lighting. Extensive experiments demonstrate that GO-Renderer achieves state-of-the-art performance across the object rendering tasks, including synthesizing images on new viewpoints, rendering the objects in a novel lighting environment, and inserting an object into an existing video.

3D重建扩散模型视频生成可控生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。