针对高压缩视频设计的超分辨率方法,可将180p/270p视频提升至720p/1080p。
Compressed Video Super-Resolution based on Hierarchical Encoding
- 采用分层编码变压器块,专治H.265/HEVC压缩伪影。
- 在多种码率下训练,实现4倍超分辨率,恢复细粒度细节。
- 适用于视频会议场景,已提交ICME 2025挑战赛测试。
本文提出一种通用视频超分辨率方法VSR-HE,专为提升高压缩内容的感知质量而设计。针对重压缩场景,该方法可将低分辨率视频放大4倍,从180p升至720p,或从270p升至1080p。VSR-HE采用分层编码变压器块,在不同量化参数(QP)水平下,有效消除H.265/HEVC编码引入的广泛压缩伪影。为确保鲁棒性和泛化能力,模型在多样的压缩设置下进行训练与评估,能有效恢复精细细节并保持视觉保真度。该方法已正式提交至ICME 2025视频会议视频超分辨率挑战赛(团队BVI-VSR),涵盖通用真实世界视频内容(Track 1)与说话人头部视频(Track 2)两个赛道。
原文摘要 · Abstract (English)
This paper presents a general-purpose video super-resolution (VSR) method, dubbed VSR-HE, specifically designed to enhance the perceptual quality of compressed content. Targeting scenarios characterized by heavy compression, the method upscales low-resolution videos by a ratio of four, from 180p to 720p or from 270p to 1080p. VSR-HE adopts hierarchical encoding transformer blocks and has been sophisticatedly optimized to eliminate a wide range of compression artifacts commonly introduced by H.265/HEVC encoding across various quantization parameter (QP) levels. To ensure robustness and generalization, the model is trained and evaluated under diverse compression settings, allowing it to effectively restore fine-grained details and preserve visual fidelity. The proposed VSR-HE has been officially submitted to the ICME 2025 Grand Challenge on VSR for Video Conferencing (Team BVI-VSR), under both the Track 1 (General-Purpose Real-World Video Content) and Track 2 (Talking Head Videos).
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。