arXiv:2412.01493cs.CVeess.IV2024-12ICML被引 7

通过通道感知机制统一处理多种光照问题,提升视觉一致性与效率。

Learning Adaptive Lighting via Channel-Aware Guidance

  • 引入通道感知注意力机制,分离并融合彩色通道的光照特性。
  • 在四个光照任务上超越现有方法,且计算开销更低。
  • 适合需要多任务光照自适应的计算机视觉应用开发者。

光照自适应是实现良好视觉感知和支撑下游视觉任务的关键步骤。现有研究常孤立处理高动态范围成像、曝光校正等单一光照问题。本文发现这些任务共享两个基本属性:不同颜色通道具有不同的光照特性,且通道差异在空间与频率域表现各异。基于此,提出通道感知光照自适应网络(LALNet),一种可高效处理多类光照任务的多任务框架。LALNet通过光引导注意力(LGA)模块,将色彩分离特征与传统混合特征融合,使混合特征聚焦于通道差异并保持全通道视觉一致性。同时采用双域通道调制生成色彩分离特征,以及混合通道调制与光照状态空间模块生成混合特征。在四个代表性光照任务上的大量实验表明,LALNet显著优于当前最优方法,且所需计算资源更少。演示链接:https://xxxxxx2025.github.io/LALNet/

原文摘要 · Abstract (English)

Learning lighting adaptation is a crucial step in achieving good visual perception and supporting downstream vision tasks. Current research often addresses individual light-related challenges, such as high dynamic range imaging and exposure correction, in isolation. However, we identify shared fundamental properties across these tasks: i) different color channels have different light properties, and ii) the channel differences reflected in the spatial and frequency domains are different. Leveraging these insights, we introduce the channel-aware Learning Adaptive Lighting Network (LALNet), a multi-task framework designed to handle multiple light-related tasks efficiently. Specifically, LALNet incorporates color-separated features that highlight the unique light properties of each color channel, integrated with traditional color-mixed features by Light Guided Attention (LGA). The LGA utilizes color-separated features to guide color-mixed features focusing on channel differences and ensuring visual consistency across all channels. Additionally, LALNet employs dual domain channel modulation for generating color-separated features and a mixed channel modulation and light state space module for producing color-mixed features. Extensive experiments on four representative light-related tasks demonstrate that LALNet significantly outperforms state-of-the-art methods on benchmark tests and requires fewer computational resources. We provide an anonymous online demo at https://xxxxxx2025.github.io/LALNet/.

光照自适应多任务学习通道感知视觉一致性

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。