79
4

Token Communications: A Unified Framework for Cross-modal Context-aware Semantic Communications

Abstract

In this paper, we introduce token communications (TokCom), a unified framework to leverage cross-modal context information in generative semantic communications (GenSC). TokCom is a new paradigm, motivated by the recent success of generative foundation models and multimodal large language models (GFM/MLLMs), where the communication units are tokens, enabling efficient transformer-based token processing at the transmitter and receiver. In this paper, we introduce the potential opportunities and challenges of leveraging context in GenSC, explore how to integrate GFM/MLLMs-based token processing into semantic communication systems to leverage cross-modal context effectively, present the key principles for efficient TokCom at various layers in future wireless networks. We demonstrate the corresponding TokCom benefits in a GenSC setup for image, leveraging cross-modal context information, which increases the bandwidth efficiency by 70.8% with negligible loss of semantic/perceptual quality. Finally, the potential research directions are identified to facilitate adoption of TokCom in future wireless networks.

View on arXiv
@article{qiao2025_2502.12096,
  title={ Token Communications: A Unified Framework for Cross-modal Context-aware Semantic Communications },
  author={ Li Qiao and Mahdi Boloursaz Mashhadi and Zhen Gao and Rahim Tafazolli and Mehdi Bennis and Dusit Niyato },
  journal={arXiv preprint arXiv:2502.12096},
  year={ 2025 }
}
Comments on this paper