arXiv:2512.02780cs.CV2025-12AAAI

针对腹腔镜手术烟雾类型差异,提出首个感知烟雾类型的去烟方法。

Rethinking Surgical Smoke: A Smoke-Type-Aware Laparoscopic Video Desmoking Method and Dataset

  • 区分扩散型与环境型烟雾,用注意力机制联合预测烟雾类型与掩码
  • 通过粗到细解耦模块提升烟雾掩码精度,实现针对性去烟重建
  • 构建首个带烟雾类型标注的合成数据集,适合医疗视觉研究者使用

电刀或激光手术不可避免产生手术烟雾,干扰腹腔镜视频的视觉引导。根据运动模式,烟雾可分为不同类型,导致视频中具有不同的时空特征。然而现有去烟方法未考虑烟雾类型的差异。为此,我们提出首个烟雾类型感知的腹腔镜视频去烟网络STANet,区分两类烟雾:扩散烟雾与环境烟雾。设计烟雾掩码分割子网络,基于注意力加权掩码聚合,联合预测烟雾掩码与类型;并提出烟雾去除视频重建子网络,依据两类烟雾掩码对烟雾特征进行针对性去烟。为解决两类烟雾的纠缠问题,引入粗到细解耦模块,通过烟雾类型感知的跨区域注意力,在非纠缠与纠缠区域间实现更精准的解耦掩码。此外,我们还构建了首个大规模带烟雾类型标注的合成去烟数据集。大量实验表明,该方法在质量评估上优于当前最优方法,并在多个下游手术任务中展现更强泛化能力。

原文摘要 · Abstract (English)

Electrocautery or lasers will inevitably generate surgical smoke, which hinders the visual guidance of laparoscopic videos for surgical procedures. The surgical smoke can be classified into different types based on its motion patterns, leading to distinctive spatio-temporal characteristics across smoky laparoscopic videos. However, existing desmoking methods fail to account for such smoke-type-specific distinctions. Therefore, we propose the first Smoke-Type-Aware Laparoscopic Video Desmoking Network (STANet) by introducing two smoke types: Diffusion Smoke and Ambient Smoke. Specifically, a smoke mask segmentation sub-network is designed to jointly conduct smoke mask and smoke type predictions based on the attention-weighted mask aggregation, while a smokeless video reconstruction sub-network is proposed to perform specially desmoking on smoky features guided by two types of smoke mask. To address the entanglement challenges of two smoke types, we further embed a coarse-to-fine disentanglement module into the mask segmentation sub-network, which yields more accurate disentangled masks through the smoke-type-aware cross attention between non-entangled and entangled regions. In addition, we also construct the first large-scale synthetic video desmoking dataset with smoke type annotations. Extensive experiments demonstrate that our method not only outperforms state-of-the-art approaches in quality evaluations, but also exhibits superior generalization across multiple downstream surgical tasks.

医学图像去烟烟雾分类视频重建

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。