arXiv:2505.10946cs.ITcs.AI2025-05被引 17

用大模型提升海量文本/图像令牌通信的效率与质量。

ToDMA: Large Model-Driven Massive Token Communications for Semantic Multiple Access

  • 基于大模型驱动,通过共享调制码字实现多设备无协调传输。
  • 在图像和文本任务中降低接入延迟,保持高恢复准确率。
  • 适合大规模低时延语义通信场景,如物联网智能终端。

令牌通信(TokenCom)是一种新兴的生成式语义通信范式,其中令牌作为跨模态的紧凑表示单元,其上下文依赖性可通过预训练大模型进行语义恢复。本文提出令牌域多址接入(ToDMA),一种面向海量令牌通信的大模型驱动语义多址方案。ToDMA融合无源随机接入与上下文感知令牌处理,使大量非协作设备可共用上行资源传输令牌化源表示。具体而言,每个令牌索引关联一个共享调制码字,向接收端暴露令牌级结构以实现上下文感知恢复。接收端首先采用压缩感知联合检测活跃令牌并估计其对应的信道状态信息(CSI)。随后,通过利用多个令牌位置间令牌相关CSI的一致性重建源令牌序列。当存在令牌碰撞时,部分活跃令牌可能未被分配,导致重建序列中出现缺失项。为此,采用候选受限的掩码令牌预测,借助预训练上下文模型的上下文能力缓解碰撞影响。图像与文本传输任务的仿真结果表明,ToDMA在降低接入延迟的同时,维持了良好的令牌恢复与语义重建质量,展现出语义多址的可扩展性。

原文摘要 · Abstract (English)

Token communications (TokenCom) is an emerging generative semantic communication paradigm, where tokens serve as compact representation units across modalities. Their contextual dependencies can be exploited by pretrained large models for semantic recovery. In this paper, we propose token-domain multiple access (ToDMA), a large-model-driven semantic multiple access scheme for massive token communications. ToDMA integrates unsourced random access with context-aware token processing. It enables massive uncoordinated devices to transmit tokenized source representations over common uplink resources. Specifically, each token index is associated with a shared modulation codeword, exposing token-level structure to the receiver for context-aware recovery. At the receiver, compressed sensing is first employed to jointly detect active tokens and estimate their corresponding channel state information (CSI) from the superposed signals. The source token sequences are then reconstructed by exploiting the consistency of token-associated CSI across multiple token positions. In the presence of token collisions, some active tokens may remain unassigned, leading to missing entries in the reconstructed token sequences. To recover these tokens, candidate-restricted masked-token prediction is performed using pretrained contextual models, thereby leveraging token-level context to mitigate collision effects. Simulation results on both image and text transmission tasks demonstrate that ToDMA reduces access latency while maintaining favorable token recovery and semantic reconstruction quality, showing its scalability for semantic multiple access.

语义通信大模型多址接入令牌通信

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。