计算机科学
编码器
语义学(计算机科学)
人工智能
编码(社会科学)
解码方法
人机交互
保险丝(电气)
接头(建筑物)
融合机制
通信系统
延迟(音频)
传输(电信)
编码
传感器融合
雷达
实时计算
机器学习
利用
语义数据模型
建筑
多通道交互
无线
编码(内存)
作者
Yubo Peng,Luping Xiang,Kun Yang,Feibo Jiang,Kezhi Wang,Dapeng Wu
标识
DOI:10.1109/jsac.2025.3610398
摘要
Traditional unimodal sensing faces limitations in accuracy and capability, and its decoupled implementation with communication systems increases latency in bandwidth-constrained environments. Additionally, single-task-oriented sensing systems fail to address users’ diverse demands. To overcome these challenges, we propose a semantic-driven integrated multimodal sensing and communication (SIMAC) framework. This framework leverages a joint source-channel coding architecture to achieve simultaneous sensing, decoding, and transmission of sensing results. Specifically, SIMAC first introduces a multimodal semantic fusion (MSF) network, which employs two extractors to extract semantic information from radar signals and images, respectively. MSF then applies cross-attention mechanisms to fuse these unimodal features and generate multimodal semantic representations. Secondly, we present a large language model (LLM)-based semantic encoder (LSE), where relevant communication parameters and multimodal semantics are mapped into a unified latent space and input to the LLM, enabling channel-adaptive semantic encoding. Thirdly, a task-oriented sensing semantic decoder (SSD) is proposed, in which different decoded heads are designed according to the specific needs of tasks. Simultaneously, a multi-task learning strategy is introduced to train the SIMAC framework, achieving diverse sensing services. Finally, experimental simulations demonstrate that the proposed framework achieves diverse and higher-accuracy sensing services.
科研通智能强力驱动
Strongly Powered by AbleSci AI