arXiv:2506.06757cs.CV2025-06被引 2

从单张SAR图像中恢复飞机目标的三维语义结构,实现可解释的智能解译。

SAR2Struct: Extracting 3D Semantic Structural Representation of Aircraft Targets from Single-View SAR Image

  • 基于结构描述符的两阶段框架,融合真实与仿真数据学习结构映射。
  • 首次实现仅用一张SAR图就提取飞机目标的3D语义结构关系。
  • 适用于军事侦察、目标识别等需理解复杂结构的场景。

将合成孔径雷达(SAR)图像转化为人类可理解的形式,是高级信息检索的终极目标。现有方法多聚焦于三维表面重建或局部几何特征提取,忽视了结构建模在捕捉语义信息中的作用。本文提出新任务:从单视图SAR图像中恢复目标结构,旨在推断目标组件及其间的结构关系(如对称性与邻接性)。通过学习同类型目标在不同SAR图像中呈现的结构一致性与几何多样性,直接从二维SAR图像中获取目标的语义表示。为此,设计基于结构描述符的两步算法框架:训练阶段先从真实SAR图像检测2D关键点,再利用仿真数据学习关键点到3D层次结构的映射;测试阶段整合两步,从真实图像推断3D结构。实验验证了各步骤有效性,并首次证明可直接从单视图SAR图像中恢复飞机目标的3D语义结构表示。

原文摘要 · Abstract (English)

To translate synthetic aperture radar (SAR) image into interpretable forms for human understanding is the ultimate goal of SAR advanced information retrieval. Existing methods mainly focus on 3D surface reconstruction or local geometric feature extraction of targets, neglecting the role of structural modeling in capturing semantic information. This paper proposes a novel task: SAR target structure recovery, which aims to infer the components of a target and the structural relationships between its components, specifically symmetry and adjacency, from a single-view SAR image. Through learning the structural consistency and geometric diversity across the same type of targets as observed in different SAR images, it aims to derive the semantic representation of target directly from its 2D SAR image. To solve this challenging task, a two-step algorithmic framework based on structural descriptors is developed. Specifically, in the training phase, it first detects 2D keypoints from real SAR images, and then learns the mapping from these keypoints to 3D hierarchical structures using simulated data. During the testing phase, these two steps are integrated to infer the 3D structure from real SAR images. Experimental results validated the effectiveness of each step and demonstrated, for the first time, that 3D semantic structural representation of aircraft targets can be directly derived from a single-view SAR image.

SAR图像三维结构语义理解目标识别

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。