arXiv:2603.27300cs.CV2026-03

端到端重建动态场景4D几何,补全遮挡区域

Complet4R: Geometric Complete 4D Reconstruction

  • 用纯解码器变压器全局处理视频序列,直接积累完整上下文
  • 在新基准上实现4D重建与3D点追踪的最新性能
  • 适合做动态场景三维重建、视觉跟踪的研究者使用

我们提出Complet4R,一种用于几何完整4D重建的新端到端框架,旨在恢复动态场景中时间连贯且几何完整的三维结构。该方法将几何完整4D重建任务统一为重建与补全的联合框架,通过直接将完整上下文累积到每一帧来实现。与依赖成对重建或局部运动估计的现有方法不同,Complet4R采用仅解码器的Transformer,直接从序列视频输入中全局操作,为每个时间戳重建完整的几何结构,包括其他帧中可见但当前遮挡的区域。该方法在我们提出的几何完整4D重建基准和3D点追踪任务上均达到当前最优表现。代码将开源以支持后续研究。

原文摘要 · Abstract (English)

We introduce Complet4R, a novel end-to-end framework for Geometric Complete 4D Reconstruction, which aims to recover temporally coherent and geometrically complete reconstruction for dynamic scenes. Our method formalizes the task of Geometric Complete 4D Reconstruction as a unified framework of reconstruction and completion, by directly accumulating full contexts onto each frame. Unlike previous approaches that rely on pairwise reconstruction or local motion estimation, Complet4R utilizes a decoder-only transformer to operate all context globally directly from sequential video input, reconstructing a complete geometry for every single timestamp, including occluded regions visible in other frames. Our method demonstrates the state-of-the-art performance on our proposed benchmark for Geometric Complete 4D Reconstruction and the 3D Point Tracking task. Code will be released to support future research.

4D重建几何补全视频理解

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。