跳到主要导航 跳到搜索 跳到主要内容

Reliable Multimodal Semantic Communication for Audio-Visual Event Localization

  • Yuandi Li
  • , Zhe Xiang
  • , Fei Yu*
  • , Zhuoran Zhang
  • , Yanhao Wang*
  • , Zhangshuang Guan
  • , Hui Ji
  • , Zhiguo Wan
  • *此作品的通讯作者
  • Jiangsu University
  • Liaoning University of Technology
  • Zhejiang Lab
  • Zhejiang University
  • Wuxi Taihu University

科研成果: 期刊稿件文章同行评审

摘要

The widespread adoption of smart mobile devices and applications has driven an exponential growth in wireless data traffic, posing significant challenges to modern communication systems. Ensuring reliable task-oriented multimodal semantic communication has become increasingly critical. In this letter, we propose RMMSC, a novel framework designed to enhance the effectiveness and reliability of Audio-Visual Event (AVE) localization-driven multimodal semantic communication. Specifically, RMMSC improves the accuracy of multimodal semantic information through advanced semantic encoding and cross-modal feature integration. It employs a two-level coding scheme that combines error-correcting codes with semantic encoders to enhance the reliability of multimodal semantic transmission. As an optional design choice, RMMSC supports a hybrid encryption mechanism to protect transmitted data if required by the application context. Simulation results validate the effectiveness of RMMSC, demonstrating significant improvements in accuracy and reliability for the AVE task.

源语言英语
页(从-至)317-321
页数5
期刊IEEE Communications Letters
30
DOI
出版状态已出版 - 2026

学术指纹

探究 'Reliable Multimodal Semantic Communication for Audio-Visual Event Localization' 的科研主题。它们共同构成独一无二的学术指纹。

引用此