arXiv:2605.21440cs.CV2026-05

轻量级框架ReMATF仅用两帧视频就有效缓解湍流失真,适合实时部署。

ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes

论文配图:ReMATF: Recurrent Motion-Adaptive Multi-scale Turbulence Mitigation for Dynamic Scenes
图 1 · 摘自论文原文
  • 基于多尺度编码器-解码器与运动自适应融合,每帧仅依赖前后两帧
  • 在合成与真实数据集上实现更高PSNR/SSIM和更低LPIPS,速度远超多帧模型
  • 适合嵌入式设备等资源受限场景,兼顾清晰度与时间一致性

大气湍流会引入几何扭曲、模糊和时间闪烁等失真,严重影响视频的视觉清晰度与时间一致性。现有最先进的方法多基于变换器或3D架构,需多帧输入,但计算量大、内存占用高,难以实时部署,尤其在资源受限场景下。本文提出ReMATF,一种轻量级递归框架,仅需两帧输入即可恢复视频,同时保持空间细节与时间稳定性。ReMATF结合多尺度编码器-解码器、时空形变与运动自适应时间融合模块,通过逐像素融合前一帧输出与当前预测结果,在不扩大时间窗口的前提下提升一致性。该设计有效抑制闪烁、增强细节,并保持高效。在合成与真实湍流数据集上的实验表明,ReMATF在PSNR/SSIM和感知质量(LPIPS)上均取得一致提升,推理速度显著快于多帧变压器基线,适用于资源受限场景下的湍流抑制。

原文摘要 · Abstract (English)

Atmospheric turbulence severely degrades video quality by introducing distortions such as geometric warping, blur, and temporal flickering, posing significant challenges to both visual clarity and temporal consistency. Current state-of-the-art methods are based on transformer, 3D architectures and require multi-frame input, but their large computational cost and memory usage limit real-time deployment, especially in resource-constrained scenarios. In this work, we propose ReMATF, a lightweight recurrent framework that restores videos using only two frames at a time while preserving spatial detail and temporal stability. ReMATF combines a multi-scale encoder-decoder with temporal warping and a motion-adaptive temporal fusion module that performs per-pixel fusion between the warped previous output and the current prediction to enhance coherence without enlarging the temporal window. This design reduces flicker, sharpens details, and remains efficient. Experiments on synthetic and real turbulence datasets show consistent improvements in PSNR/SSIM and perceptual quality (LPIPS), along with substantially faster inference than multi-frame transformer baselines, making ReMATF suitable turbulence mitigation in resource-constrained scenarios.

视频去湍流轻量化模型时序稳定实时处理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。