用文本生成可远距离漫游的全景3D世界
SphericalDreamer: Generating Navigable Immersive 3D Worlds with Panorama Fusion

- 通过生成多张全景图并融合成3D场景
- 实现完整360度水平视野与180度垂直视野
- 适合虚拟现实、游戏等需要沉浸式环境的场景
随着虚拟现实和3D内容的普及,生成沉浸式可漫游的3D环境愈发重要。然而,现有方法存在根本局限:无法同时实现(i)长距离空间范围内的可导航性,以及(ii)完整的全向视野(360°水平,180°垂直)。为此,我们提出SphericalDreamer,一种从文本提示生成完整沉浸式、长距离3D户外环境的方法。该方法基于生成多张全景图像,随后将其提升为3D并融合,同时保持视觉与几何一致性。SphericalDreamer生成高度细节化、完全沉浸的3D环境,在尺度与可导航性上显著优于先前方法。
原文摘要 · Abstract (English)
The generation of immersive and navigable 3D environments is increasingly prevalent with the growing adoption of virtual reality and 3D content. However, recent methods face a fundamental limitation: they cannot produce 3D worlds that simultaneously (i) are navigable over long-range spatial extents and (ii) cover the complete omnidirectional field of view ($360^\circ$ horizontally and $180^\circ$ vertically). To address this challenge, we introduce SphericalDreamer, a method for generating fully immersive and long-range 3D outdoor environments from textual prompts. Our approach is built on the generation of multiple panoramic images, which are subsequently lifted into 3D and fused together while maintaining visual and geometric consistency. SphericalDreamer produces highly detailed, fully immersive 3D environments, while substantially improving scale and navigability compared to prior approaches.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。