arXiv:2604.09231cs.CV2026-04被引 1

用多视角引导生成更完整对齐的3D纹理,解决纹理缺失和不一致问题。

Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation

论文配图:Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation
图 1 · 摘自论文原文
  • 融合多视角图像先验与原生3D纹理表示,提升生成质量。
  • 在真实场景数据集上,纹理完整性与跨视角一致性显著优于现有方法。
  • 适合需要高质量3D内容生成的工业级应用,如数字孪生、游戏建模。

尽管近期研究提升了3D纹理生成质量,但现有方法仍面临纹理覆盖不全、跨视角不一致以及几何与纹理错位等问题。为此,我们提出Hitem3D 2.0,一种基于多视角引导的原生3D纹理生成框架,通过整合2D多视角生成先验与原生3D纹理表示,显著提升纹理质量。该框架包含两个核心组件:多视角合成模块与原生3D纹理生成模型。多视角合成基于预训练图像编辑骨干网络,引入即插即用模块,显式强化几何对齐、跨视角一致性与光照均匀性,实现高保真多视角图像生成。在生成视图与3D几何条件下,原生3D纹理生成模型将多视角纹理投影至3D表面,并合理补全未见区域纹理。通过融合多视角一致性约束与原生3D纹理建模,Hitem3D 2.0大幅改善纹理完整性、跨视角连贯性与几何对齐度。实验表明,该方法在纹理细节、保真度、一致性、连贯性及对齐性方面均优于现有方法。

原文摘要 · Abstract (English)

Although recent advances have improved the quality of 3D texture generation, existing methods still struggle with incomplete texture coverage, cross-view inconsistency, and misalignment between geometry and texture. To address these limitations, we propose Hitem3D 2.0, a multi-view guided native 3D texture generation framework that enhances texture quality through the integration of 2D multi-view generation priors and native 3D texture representations. Hitem3D 2.0 comprises two key components: a multi-view synthesis framework and a native 3D texture generation model. The multi-view generation is built upon a pre-trained image editing backbone and incorporates plug-and-play modules that explicitly promote geometric alignment, cross-view consistency, and illumination uniformity, thereby enabling the synthesis of high-fidelity multi-view images. Conditioned on the generated views and 3D geometry, the native 3D texture generation model projects multi-view textures onto 3D surfaces while plausibly completing textures in unseen regions. Through the integration of multi-view consistency constraints with native 3D texture modeling, Hitem3D 2.0 significantly improves texture completeness, cross-view coherence, and geometric alignment. Experimental results demonstrate that Hitem3D 2.0 outperforms existing methods in terms of texture detail, fidelity, consistency, coherence, and alignment.

3D纹理多视角生成模型

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。