6.3MMMay 3
Contextual Wireless Video Semantic Communication in MIMO-OFDM SystemsBingyan Xie, Cong Zhou, Yuxuan Shi et al.
This paper proposes a MIMO-OFDM-based context video semantic transmission framework, namely M-CVST, for robust video communication over multi-path multiple-input multiple-output (MIMO) channels. It introduces a context-subcarrier correlation map that aligns video feature context with groups of MIMO subcarriers. To leverage the time-correlated nature of multi-path channels, a recursive subcarrier sampling method paired with time-correlated reference embedding is designed, enabling the use of previously sampled MIMO subcarrier CSI to enhance channel state awareness in the entropy coding model. Numerical results verify the superiority of proposed M-CVST over MIMO multi-path channels compared to other semantic schemes and traditional separated schemes.
5.1ITJun 29
Semantic Noise Aided Secure Image Transmission over MIMO Fading ChannelsXue Han, Biqian Feng, Ting Zhou et al.
Existing semantic communications have exhibited satisfactory performance in many tasks, but secure image transmission remains insufficiently explored. We propose a novel secure image semantic communication (SISC) framework over multiple-input multiple-output (MIMO) fading channels. To ensure high-quality image reconstruction for the legitimate semantic user (SU) and simultaneously interfere with the eavesdropper (Eve), we design a semantic noise generation (SNG) network. This network generates a beneficial semantic noise map based on both the source features and the SU channel state information (CSI). An efficient channel estimation enhanced network is incorporated to obtain the accurate CSI and enhance the system performance. Furthermore, to improve the secure image reconstruction quality, we develop an efficient transceiver beamformer optimization algorithm, where the formulated problem is solved using the constrained stochastic successive convex approximation method. In the proposed SISC framework, semantic noise generation and beamforming optimization work together to ensure secure and high-quality image transmission. Numerical results demonstrate that the proposed semantic noise aided transmission scheme effectively protects image information from leakage to Eve while maintaining high-fidelity image reconstruction at SU.
7.9ITJun 29
Towards World Model-Empowered Integrated Sensing, Communication, and Decision for Complex Unmanned SystemsXue Han, Yongpeng Wu, Meng Shen et al.
Complex unmanned systems comprising satellites, unmanned aerial vehicles (UAVs), unmanned ground vehicles (UGVs), and quadruped robots are increasingly deployed to perform large-scale sensing and autonomous operations. We propose a world model-empowered sensing, communication, decision (SCD) integration framework for complex unmanned communication networks. The proposed architecture establishes a closed-loop system where a unified world model jointly optimizes time-sensitive sensing, wireless communication, and intelligent decision-making. To regulate sensing freshness and reduce redundant data generation, we propose a time-sensitive age of information (AoI)-driven sensing mechanism that dynamically schedules sensing updates based on task urgency and predictive uncertainty. Furthermore, a predictive world model is developed to jointly represent environmental dynamics, wireless channel evolution, and agent mobility within a hybrid deterministic-stochastic latent space. This enables proactive communication scheduling and decision evaluation via latent rollout. To support large-scale heterogeneous coordination, a multi-granularity knowledge graph is further designed to organize cross-population relationships among satellites, UAVs, UGVs, and ground agents. Numerical results demonstrate that the proposed SCD framework outperforms conventional systems, highlighting the significant potential of world models for supporting unmanned systems.
2.3MMMar 27, 2025
WVSC: Wireless Video Semantic Communication with Multi-frame CompensationBingyan Xie, Yongpeng Wu, Yuxuan Shi et al.
Existing wireless video transmission schemes directly conduct video coding in pixel level, while neglecting the inner semantics contained in videos. In this paper, we propose a wireless video semantic communication framework, abbreviated as WVSC, which integrates the idea of semantic communication into wireless video transmission scenarios. WVSC first encodes original video frames as semantic frames and then conducts video coding based on such compact representations, enabling the video coding in semantic level rather than pixel level. Moreover, to further reduce the communication overhead, a reference semantic frame is introduced to substitute motion vectors of each frame in common video coding methods. At the receiver, multi-frame compensation (MFC) is proposed to produce compensated current semantic frame with a multi-frame fusion attention module. With both the reference frame transmission and MFC, the bandwidth efficiency improves with satisfying video transmission performance. Experimental results verify the performance gain of WVSC over other DL-based methods e.g. DVSC about 1 dB and traditional schemes about 2 dB in terms of PSNR.