SemanticNN: Compressive and Error-Resilient Semantic Offloading for Extremely Weak Devices
Jiaming Huang, Yi Gao, Fuchang Pan, Renjie Li, Wei Dong
摘要
With the rapid growth of the Internet of Things (IoT), integrating artificial intelligence (AI) on extremely weak embedded devices has garnered significant attention, enabling improved real-time performance and enhanced data privacy. However, the resource limitations of such devices and unreliable network conditions necessitate error-resilient device-edge collaboration systems. Traditional approaches focus on bit-level transmission correctness, which can be inefficient under dynamic channel conditions. In contrast, we propose SemanticNN, a semantic codec that tolerates bit-level errors in pursuit of semantic-level correctness, enabling compressive and resilient collaborative inference offloading under strict computational and communication constraints. It incorporates a Bit Error Rate (BER)-aware decoder that adapts to dynamic channel conditions and a Soft Quantization (SQ)-based encoder to learn compact representations. Building on this architecture, we introduce Feature-augmentation Learning, a novel training strategy that enhances offloading efficiency. To address encoder-decoder capability mismatches from asymmetric resources, we propose XAI-based Asymmetry Compensation to enhance decoding semantic fidelity. We conduct extensive experiments on STM32 using three models and six datasets across image classification and object detection tasks. Experimental results demonstrate that, under varying transmission error rates, SemanticNN significantly reduces feature transmission volume by 56.82–344.83× while maintaining superior inference accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- SPINN: synergistic progressive inference of neural networks over device and cloudStefanos Laskaridis, Stylianos I. Venieris, Mário Almeida, Ilias Leontiadis 等MobiCom 2020 · 被引用 312 次
- Flexible high-resolution object detection on edge devices with tunable latencyShiqi Jiang, Zhiqi Lin, Yuanchun Li, Yuanchao Shu 等MobiCom 2021 · 被引用 103 次
- Real-time neural network inference on extremely weak devices: agile offloading with explainable AIKai Huang, Wei GaoMobiCom 2022 · 被引用 57 次
- InFi: end-to-end learnable input filter for resource-efficient mobile-centric inferenceMu Yuan, Lan Zhang, Fengxiang He, Xueting Tong 等MobiCom 2022 · 被引用 36 次
- TASTI: Semantic Indexes for Machine Learning-based Queries over Unstructured DataDaniel Kang, John Guibas, Peter D. Bailis, Tatsunori Hashimoto 等SIGMOD 2022 · 被引用 23 次
相关 Paper
- Compressive sensing based asymmetric semantic image compression for resource-constrained IoT systemYujun Huang, Bin Chen, Jianghui Zhang, Han Qiu 等DAC 2022 · 被引用 5 次
- DNN-Driven Compressive Offloading for Edge-Assisted Semantic Video SegmentationXuedou Xiao, Juecheng Zhang, Wei Wang, Jianhua He 等INFOCOM 2022 · 被引用 27 次
- OMNIS: Semantic RAN Slicing via Dynamic Split Neural NetworksLangtian Qin, Ian Harshbarger, Leïla Nasraoui, Carla Fabiana Chiasserini 等INFOCOM 2026 · 被引用 1 次
- Progressive Neural Compression for Adaptive Image Offloading Under Timing ConstraintsRuiqi Wang, Hanyang Liu, Jiaming Qiu, Moran Xu 等RTSS 2023 · 被引用 10 次
- Hierarchical Channel-spatial Encoding for Communication-efficient Collaborative LearningQihua Zhou, Song Guo, Yi Liu, Jie Zhang 等NeurIPS 2022 · 被引用 4 次
