arXiv:2504.12356eess.IVcs.CV2025-04中稿 · ACM Multimedia 202…被引 10

用立体基础模型实现高效增量3D重建,支持千视角大规模场景

Regist3R: Incremental Registration with Stereo Foundation Model

  • 采用增量重建范式,避免全局对齐累积误差
  • 在公共数据集上达到优化方法水平,计算效率显著提升
  • 首次实现超千视角点云重建,适合城市建模等应用

多视角3D重建是计算机视觉中的核心挑战。尽管DUSt3R及其后续方法在无姿态图像的3D重建上取得突破,但在多视角场景中仍存在计算成本高、全局对齐引发累积误差等问题。为此,我们提出Regist3R,一种专为高效可扩展增量重建设计的立体基础模型。Regist3R采用增量重建范式,可从无序多视角图像集合中实现大规模3D重建。我们在公开数据集上评估了其相机位姿估计与3D重建性能,结果表明,Regist3R在保持优化方法相当精度的同时显著提升计算效率,并优于现有多视角重建模型。此外,为验证实际应用能力,我们引入一个具有长空间跨度和数百视角的难测斜视航空数据集,结果证明了Regist3R的有效性。我们还首次展示了通过基于点云的基础模型重建超过千视角的大规模场景,展现出其在城市建模、航拍测绘等任务中的实用潜力。

原文摘要 · Abstract (English)

Multi-view 3D reconstruction has remained an essential yet challenging problem in the field of computer vision. While DUSt3R and its successors have achieved breakthroughs in 3D reconstruction from unposed images, these methods exhibit significant limitations when scaling to multi-view scenarios, including high computational cost and cumulative error induced by global alignment. To address these challenges, we propose Regist3R, a novel stereo foundation model tailored for efficient and scalable incremental reconstruction. Regist3R leverages an incremental reconstruction paradigm, enabling large-scale 3D reconstructions from unordered and many-view image collections. We evaluate Regist3R on public datasets for camera pose estimation and 3D reconstruction. Our experiments demonstrate that Regist3R achieves comparable performance with optimization-based methods while significantly improving computational efficiency, and outperforms existing multi-view reconstruction models. Furthermore, to assess its performance in real-world applications, we introduce a challenging oblique aerial dataset which has long spatial spans and hundreds of views. The results highlight the effectiveness of Regist3R. We also demonstrate the first attempt to reconstruct large-scale scenes encompassing over thousands of views through pointmap-based foundation models, showcasing its potential for practical applications in large-scale 3D reconstruction tasks, including urban modeling, aerial mapping, and beyond.

3D重建增量重建立体模型大场景

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。