SemanticNN: Compressive and Error-Resilient Semantic Offloading for Extremely Weak Devices
Jiaming Huang, Yi Gao, Fuchang Pan, Renjie Li, Wei Dong
Abstract
With the rapid growth of the Internet of Things (IoT), integrating artificial intelligence (AI) on extremely weak embedded devices has garnered significant attention, enabling improved real-time performance and enhanced data privacy. However, the resource limitations of such devices and unreliable network conditions necessitate error-resilient device-edge collaboration systems. Traditional approaches focus on bit-level transmission correctness, which can be inefficient under dynamic channel conditions. In contrast, we propose SemanticNN, a semantic codec that tolerates bit-level errors in pursuit of semantic-level correctness, enabling compressive and resilient collaborative inference offloading under strict computational and communication constraints. It incorporates a Bit Error Rate (BER)-aware decoder that adapts to dynamic channel conditions and a Soft Quantization (SQ)-based encoder to learn compact representations. Building on this architecture, we introduce Feature-augmentation Learning, a novel training strategy that enhances offloading efficiency. To address encoder-decoder capability mismatches from asymmetric resources, we propose XAI-based Asymmetry Compensation to enhance decoding semantic fidelity. We conduct extensive experiments on STM32 using three models and six datasets across image classification and object detection tasks. Experimental results demonstrate that, under varying transmission error rates, SemanticNN significantly reduces feature transmission volume by 56.82–344.83× while maintaining superior inference accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 203493a9-072b-4a42-9db3-ecae1d20c5c8Builds on8
- SPINN: synergistic progressive inference of neural networks over device and cloudStefanos Laskaridis, Stylianos I. Venieris, Mário Almeida, Ilias Leontiadis et al.MobiCom 2020 · 312 citations
- Flexible high-resolution object detection on edge devices with tunable latencyShiqi Jiang, Zhiqi Lin, Yuanchun Li, Yuanchao Shu et al.MobiCom 2021 · 103 citations
- Real-time neural network inference on extremely weak devices: agile offloading with explainable AIKai Huang, Wei GaoMobiCom 2022 · 57 citations
- InFi: end-to-end learnable input filter for resource-efficient mobile-centric inferenceMu Yuan, Lan Zhang, Fengxiang He, Xueting Tong et al.MobiCom 2022 · 36 citations
- TASTI: Semantic Indexes for Machine Learning-based Queries over Unstructured DataDaniel Kang, John Guibas, Peter D. Bailis, Tatsunori Hashimoto et al.SIGMOD 2022 · 23 citations
Related papers
- Compressive sensing based asymmetric semantic image compression for resource-constrained IoT systemYujun Huang, Bin Chen, Jianghui Zhang, Han Qiu et al.DAC 2022 · 5 citations
- DNN-Driven Compressive Offloading for Edge-Assisted Semantic Video SegmentationXuedou Xiao, Juecheng Zhang, Wei Wang, Jianhua He et al.INFOCOM 2022 · 27 citations
- OMNIS: Semantic RAN Slicing via Dynamic Split Neural NetworksLangtian Qin, Ian Harshbarger, Leïla Nasraoui, Carla Fabiana Chiasserini et al.INFOCOM 2026 · 1 citation
- Progressive Neural Compression for Adaptive Image Offloading Under Timing ConstraintsRuiqi Wang, Hanyang Liu, Jiaming Qiu, Moran Xu et al.RTSS 2023 · 10 citations
- Hierarchical Channel-spatial Encoding for Communication-efficient Collaborative LearningQihua Zhou, Song Guo, Yi Liu, Jie Zhang et al.NeurIPS 2022 · 4 citations
