Neural Rate Estimator and Unsupervised Learning for Efficient Distributed Image Analytics in Split-DNN models
Nilesh A. Ahuja, Parual Datta, Bhavya Kanzariya, V. Srinivasa Somayazulu, Omesh Tickoo
摘要
Thanks to advances in computer vision and AI, there has been a large growth in the demand for cloud-based visual analytics in which images captured by a low-powered edge device are transmitted to the cloud for analytics. Use of conventional codecs (JPEG, MPEG, HEVC, etc.) for compressing such data introduces artifacts that can seriously degrade the performance of the downstream analytic tasks. Split-DNN computing has emerged as a paradigm to address such usages, in which a DNN is partitioned into a client-side portion and a server side portion. Lowcomplexity neural networks called 'bottleneck units' are introduced at the split point to transform the intermediate layer features into a lower-dimensional representation better suited for compression and transmission. Optimizing the pipeline for both compression and task-performance requires high-quality estimates of the information-theoretic rate of the intermediate features. Most works on compression for image analytics use heuristic approaches to estimate the rate, leading to suboptimal performance. We propose a high-quality 'neural rate-estimator' to address this gap. We interpret the lower-dimensional bottleneck output as a latent representation of the intermediate feature and cast the rate-distortion optimization problem as one of training an equivalent variational auto-encoder with an appropriate loss function. We show that this leads to improved rate-distortion outcomes. We further show that replacing supervised loss terms (such as cross-entropy loss) by distillation-based losses in a teacher-student framework allows for unsupervised training of bottleneck units without the need for explicit training labels. This makes our method very attractive for real world deployments where access to labeled training data is difficult or expensive. We demonstrate that our method outperforms several state-of-the-art methods by obtaining improved task accuracy at lower bitrates on image classification and semantic segmentation tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Diff-ICMH: Harmonizing Machine and Human Vision in Image Compression with Generative PriorRuoyu Feng, Yunpeng Qi, Jinming Liu, Yixin Gao 等NeurIPS 2025 · 被引用 5 次
- What Your Features Reveal: Data-Efficient Black-Box Feature Inversion Attack for Split DNNsZhihan Ren, Lijun He, Jiaxi Liang, Xinzhu Fu 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper1
相关 Paper
- VidIQ: Inference-Aware Neural Codecs for Quality-Enhanced, Real-Time Video AnalyticsAndong Zhu, Sheng Zhang, Xiaohang Shi, Hesheng Sun 等ACM MM 2025
- Optimizing Information Theory Based Bitwise Bottlenecks for Efficient Mixed-Precision Activation QuantizationXichuan Zhou, Kui Liu, Cong Shi, Haijun Liu 等AAAI 2021 · 被引用 3 次
- AdaStreamer: Machine-Centric High-Accuracy Multi-Video Analytics with Adaptive Neural CodecsAndong Zhu, Sheng Zhang, Ke Cheng, Xiaohang Shi 等INFOCOM 2024 · 被引用 8 次
- Flexible Neural Image Compression via Code EditingChenjian Gao, Tongda Xu, Dailan He, Yan Wang 等NeurIPS 2022 · 被引用 34 次
- Information Bottleneck Analysis of Deep Neural Networks via Lossy CompressionIvan Butakov, Aleksander Tolmachev, Sofia Malanchuk, Anna Neopryatnaya 等ICLR 2024 · 被引用 20 次
