arXiv:2509.21265cs.CVcs.AI2025-09ICCV被引 3

医学视频超分新框架,解决模糊抖动和结构失真问题

MedVSR: Medical Video Super-Resolution with Cross State-Space Propagation

  • 用跨状态空间传播对齐远距离帧,提升运动估计精度
  • 在四类医疗场景中,重建质量超越现有模型,尤其改善组织结构清晰度
  • 适合内窥镜、白内障手术等高精度医疗视频增强任务

高分辨率医学视频对精准诊断至关重要,但受硬件与生理限制难以获取。临床采集的低分辨率医学视频存在相机抖动、噪声及帧间突变等挑战,导致光流估计误差大、帧对齐困难。同时,人体组织结构连续细腻,现有超分辨率模型易引入伪影和特征扭曲,误导医生判断。为此,我们提出专用于医学视频超分辨率的MedVSR框架。该框架首先采用跨状态空间传播(CSSP),将远距离帧作为控制矩阵嵌入状态空间模型,实现一致且信息丰富的特征选择性传播,有效缓解对齐偏差。此外,设计内部状态空间重建模块(ISSR),通过联合长程空间特征学习与大核短程信息聚合,增强组织结构表现并减少伪影。在包括内窥镜和白内障手术在内的四个不同医疗数据集上的实验表明,MedVSR在重建性能与效率上显著优于现有模型。代码已开源:https://github.com/CUHK-AIM-Group/MedVSR。

原文摘要 · Abstract (English)

High-resolution (HR) medical videos are vital for accurate diagnosis, yet are hard to acquire due to hardware limitations and physiological constraints. Clinically, the collected low-resolution (LR) medical videos present unique challenges for video super-resolution (VSR) models, including camera shake, noise, and abrupt frame transitions, which result in significant optical flow errors and alignment difficulties. Additionally, tissues and organs exhibit continuous and nuanced structures, but current VSR models are prone to introducing artifacts and distorted features that can mislead doctors. To this end, we propose MedVSR, a tailored framework for medical VSR. It first employs Cross State-Space Propagation (CSSP) to address the imprecise alignment by projecting distant frames as control matrices within state-space models, enabling the selective propagation of consistent and informative features to neighboring frames for effective alignment. Moreover, we design an Inner State-Space Reconstruction (ISSR) module that enhances tissue structures and reduces artifacts with joint long-range spatial feature learning and large-kernel short-range information aggregation. Experiments across four datasets in diverse medical scenarios, including endoscopy and cataract surgeries, show that MedVSR significantly outperforms existing VSR models in reconstruction performance and efficiency. Code released at https://github.com/CUHK-AIM-Group/MedVSR.

医学图像视频超分状态空间内窥镜

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。