arXiv:2604.12837cs.RO2026-04中稿 · ICRA

用通用运动模型提升动态环境下的单目3D高斯映射精度

GGD-SLAM: Monocular 3DGS SLAM Powered by Generalizable Motion Model for Dynamic Environments

论文配图:GGD-SLAM: Monocular 3DGS SLAM Powered by Generalizable Motion Model for Dynamic Environments
图 1 · 摘自论文原文
  • 通过先进先出队列与序列注意力,分离静态与动态特征
  • 在真实动态数据集上实现领先的位姿估计与稠密重建效果
  • 无需语义标注或深度输入,适合复杂动态场景应用

视觉SLAM算法借助3D高斯点阵(3DGS)表示显著提升了稠密地图的保真度,但其依赖静态环境假设,在动态环境中性能严重下降。本文提出GGD-SLAM框架,采用可泛化的运动模型应对动态环境中的定位与稠密建图挑战——无需预定义语义标注或深度输入。系统使用先进先出(FIFO)队列管理输入帧,结合序列注意力机制实现动态语义特征提取,并通过动态特征增强器分离静态与动态成分。为降低动态干扰对静态部分的影响,设计基于静态信息采样的遮挡区域填补方法,以及针对动态环境定制的干扰自适应结构相似性指数(SSIM)损失,显著增强系统鲁棒性。在真实动态数据集上的实验表明,该系统在相机位姿估计与稠密重建方面达到当前最优性能。

原文摘要 · Abstract (English)

Visual SLAM algorithms achieve significant improvements through the exploration of 3D Gaussian Splatting (3DGS) representations, particularly in generating high-fidelity dense maps. However, they depend on a static environment assumption and experience significant performance degradation in dynamic environments. This paper presents GGD-SLAM, a framework that employs a generalizable motion model to address the challenges of localization and dense mapping in dynamic environments - without predefined semantic annotations or depth input. Specifically, the proposed system employs a First-In-First-Out (FIFO) queue to manage incoming frames, facilitating dynamic semantic feature extraction through a sequential attention mechanism. This is integrated with a dynamic feature enhancer to separate static and dynamic components. Additionally, to minimize dynamic distractors' impact on the static components, we devise a method to fill occluded areas via static information sampling and design a distractor-adaptive Structure Similarity Index Measure (SSIM) loss tailored for dynamic environments, significantly enhancing the system's resilience. Experiments conducted on real-world dynamic datasets demonstrate that the proposed system achieves state-of-the-art performance in camera pose estimation and dense reconstruction in dynamic scenes.

3D高斯动态建图视觉SLAM单目

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。