面向MIMO信道的视频语义传输框架,支持变长变率编码。
Context Video Semantic Transmission with Variable Length and Rate Coding over MIMO Channels
- 构建上下文-信道相关映射,实现特征与子信道的显式关联。
- 提出多参考熵编码,支持信道状态感知的变长编码。
- 基于棋盘调制实现单模型多码率,提升部署灵活性。
语义通信的发展深刻影响了无线视频传输,其应用是现代带宽消耗的主要驱动力。然而,现有方案大多针对加性高斯白噪声或瑞利衰落信道优化,忽略了普遍存在且严重影响实际部署的多输入多输出(MIMO)环境。为此,本文提出面向MIMO信道的上下文视频语义传输(CVST)框架。基于高效的上下文视频传输骨干网络,CVST有效学习上下文-信道相关映射,显式建模特征组与MIMO子信道之间的关系。利用这些信道感知特征,设计多参考熵编码机制,实现信道状态感知的变长编码。此外,引入基于棋盘的特征调制策略,在单一训练模型内实现多个码率点,增强部署灵活性。上述创新构成多参考变长变率编码(MR-VLRC)方案。通过融合上下文传输与MR-VLRC,CVST在多种标准化分离编码方法及近期无线视频语义通信方案上均表现出显著性能提升。代码已开源:https://github.com/xie233333/CVST。
原文摘要 · Abstract (English)
The evolution of semantic communications has profoundly impacted wireless video transmission, whose applications dominate driver of modern bandwidth consumption. However, most existing schemes are predominantly optimized for simple additive white Gaussian noise or Rayleigh fading channels, neglecting the ubiquitous multiple-input multiple-output (MIMO) environments that critically hinder practical deployment. To bridge this gap, we propose the context video semantic transmission (CVST) framework under MIMO channels. Building upon an efficient contextual video transmission backbone, CVST effectively learns a context-channel correlation map to explicitly formulate the relationships between feature groups and MIMO subchannels. Leveraging these channel-aware features, we design a multi-reference entropy coding mechanism, enabling channel state-aware variable length coding. Furthermore, CVST incorporates a checkerboard-based feature modulation strategy to achieve multiple rate points within a single trained model, thereby enhancing deployment flexibility. These innovations constitute our multi-reference variable length and rate coding (MR-VLRC) scheme. By integrating contextual transmission with MR-VLRC, CVST demonstrates substantial performance gains over various standardized separated coding methods and recent wireless video semantic communication approaches. The code is available at https://github.com/xie233333/CVST.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。