arXiv:2603.01767cs.CVeess.IV2026-03中稿 · publication in IEE…被引 10

针对水下图像识别任务,提出感知驱动的增强框架。

Downstream Task Inspired Underwater Image Enhancement: A Perception-Aware Study from Dataset Construction to Network Design

  • 基于任务感知注意力模块设计双分支网络,融合特征。
  • 在多个下游任务上显著提升分割与检测性能。
  • 构建任务导向数据集,适配水下视觉应用开发。

真实水下环境中,语义分割、目标检测等下游图像识别任务常受模糊和色彩失真影响。水下图像增强(UIE)作为预处理方法,旨在提升目标可识别性。然而,现有方法多聚焦于人眼视觉感知,难以恢复对任务关键的高频细节。为此,本文提出下游任务启发的水下图像增强框架(DTI-UIE),利用人类视觉感知模型,优化图像以支持水下视觉任务。设计高效双分支网络,引入任务感知注意力模块进行特征融合;采用多阶段训练与任务驱动感知损失;并基于多种任务网络自动构建任务导向的水下图像增强数据集(TI-UIED)。实验表明,该方法生成的预处理图像显著提升语义分割、目标检测及实例分割等任务性能。代码已开源:https://github.com/oucailab/DTIUIE。

原文摘要 · Abstract (English)

In real underwater environments, downstream image recognition tasks such as semantic segmentation and object detection often face challenges posed by problems like blurring and color inconsistencies. Underwater image enhancement (UIE) has emerged as a promising preprocessing approach, aiming to improve the recognizability of targets in underwater images. However, most existing UIE methods mainly focus on enhancing images for human visual perception, frequently failing to reconstruct high-frequency details that are critical for task-specific recognition. To address this issue, we propose a Downstream Task-Inspired Underwater Image Enhancement (DTI-UIE) framework, which leverages human visual perception model to enhance images effectively for underwater vision tasks. Specifically, we design an efficient two-branch network with task-aware attention module for feature mixing. The network benefits from a multi-stage training framework and a task-driven perceptual loss. Additionally, inspired by human perception, we automatically construct a Task-Inspired UIE Dataset (TI-UIED) using various task-specific networks. Experimental results demonstrate that DTI-UIE significantly improves task performance by generating preprocessed images that are beneficial for downstream tasks such as semantic segmentation, object detection, and instance segmentation. The codes are publicly available at https://github.com/oucailab/DTIUIE.

水下图像图像增强任务感知视觉任务

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。