CORTEXA
← Browse
arxiveess.SPcs.DC2026-07-18

Task-Oriented Communication with Hybrid-Precision Models

Songjie Xie, Wei Guo, Shenghui Song, Jun Zhang, Ying-Jun Angela Zhang, Khaled B. Letaief

Edge inference has emerged as a promising solution for the proliferation of artificial intelligence (AI) services by deploying models at the network edge to circumvent cloud-routing latency. Existing edge inference approaches mainly focused on either cooperative inference to reduce latency or lightweight model design to fit resource-constrained devices. These solutions often address the communication and computation challenges separately, and thus struggle to achieve a balanced trade-off among transmission efficiency, on-device processing cost, and inference accuracy. To bridge this gap, this paper proposes a hybrid-precision task-oriented communication framework for edge inference to holistically balance communication, on-device computation, and utility. In this framework, a binarized front-end is deployed on the edge device to extract and transmit binary features via orthogonal frequency-division multiplexing (OFDM) signals, while a full-precision back-end on the edge server performs the final inference. To ensure model consistency, we introduce an on-device binarization method tailored for split inference and develop an integrated channel-aware transmission scheme featuring subcarrier-based feature calibration. Furthermore, a knowledge distillation (KD)-based training strategy, supported by specialized gradient estimators, is developed to optimize the end-to-end system and inherit semantic knowledge from a full-precision teacher model. Extensive experiments on the large-scale ImageNet dataset demonstrate the superiority of the proposed hybrid system. Our analysis confirms that this design achieves an optimal trade-off among communication efficiency, on-device computational cost, and inference accuracy, outperforming existing edge inference solutions.

View free PDFSource page

Related papers

arxivcs.ITcs.DCcs.LGeess.SP2026-07-14

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

Chung-Hsuan Hu, Zheng Chen, Erik G. Larsson

In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by the temporal correlation between consecutive global models, differential coding can be applied to global model dissemination to r…

View free PDFSource page
arxivcs.LGcs.DCeess.SP2026-07-31

GQ-FSL: Green Quantized Federated Split Learning

Idan Roth, Lutz Lampe

Deploying state-of-the-art deep neural networks (DNNs) at the wireless edge is severely bottlenecked by the strict energy and resource constraints of mobile devices. While federated split learning (FSL) mitigates on-device computation by offloading workloads to an edge server, th…

View free PDFSource page
arxivcs.ITeess.SP2026-07-20

Task-Oriented Precoding for Edge Inference over Large-Scale MIMO Systems

Hongru Li, Zeyan Zhuang, Zixin Wang, Hengtao He, Shenghui Song, Jun Zhang, et al.

Future wireless networks are expected to support networked artificial intelligence (AI) services, where multiple devices transmit learned features to an edge server for distributed inference. This setting calls for task-oriented physical-layer optimization, where wireless transmi…

View free PDFSource page
arxiveess.SP2026-07-22

A Covert Precision Satellite Communication Framework Assisted by Cooperative IRSs

Haoyang Wu, Yunfan Bai, Mei Shen, Yuwen Qian, Guangji Chen, Long Shi, et al.

Satellite communication (SatCom), as an effective complement to terrestrial networks, has attracted considerable attention from both academia and industry owing to its wide coverage and high flexibility. However, the inherent openness of satellite links renders them highly vulner…

View free PDFSource page