The Last Byte: Learning Just Enough for Machine-Oriented Image Compression
Wuyuan Xie, Zhenming Li, Ye Liu, Jian Jin, Yun Song, Miaohui Wang
摘要
Just recognizable distortion (JRD) has been introduced for image compression for machines, aiming to quantify the maximum coding distortion that can be tolerated by a specific perception model, thereby defining the upper bound of machine vision redundancy (MVR). However, existing JRDbased redundancy estimation methods face three key challenges: limited dataset annotation accuracy, low prediction efficiency, and insufficient perception accuracy, all of which hinder their practical deployment. To address these limitations, we propose a new MVRNet, a frame-wise efficient JRD prediction method that generates the optimal encoding quantization map in a single inference pass. Furthermore, we refine the annotation standard for JRD datasets based on experimental insights, enhancing the precision of recognizable redundancy measurement. Compared to state-of-the-art methods, MVRNet achieves a superior balance between bitrate reduction and perception accuracy in JRD-guided compression, while offering up to a 40,000× speed improvement, demonstrating its practicality and efficiency for real-world applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- FBRT-YOLO: Faster and Better for Real-Time Aerial Image DetectionYao Xiao, Tingfa Xu, Yu Xin, Jianan LiAAAI 2025 · 被引用 115 次
- DeepSVC: Deep Scalable Video Coding for Both Machine and Human VisionHongbin Lin, Bolin Chen, Zhichen Zhang, Jielian Lin 等ACM MM 2023 · 被引用 32 次
- ICMH-Net: Neural Image Compression Towards both Machine Vision and Human VisionLei Liu, Zhihao Hu, Zhenghao Chen, Dong XuACM MM 2023 · 被引用 21 次
- Visual Redundancy Removal of Composite Images via Multimodal LearningWuyuan Xie, Shukang Wang, Rong Zhang, Miaohui WangACM MM 2023 · 被引用 1 次
- TR-DETR: Task-Reciprocal Transformer for Joint Moment Retrieval and Highlight DetectionHao Sun, Mingyao Zhou, Wenjing Chen, Wei XieAAAI 2024
相关 Paper
- Perceive More with Less: LiDAR Point Cloud Compression at Just Recognizable Distortion for 3D Scene UnderstandingMiaohui Wang, Runnan Huang, Taojun Liu, Shuyuan Lin 等AAAI 2026
- Firing Bits Where It Matters: Spiking-Guided Just Recognizable Distortion Modeling for Machine-Centric Video CodingWuyuan Xie, Zhenming Li, Yuwu Lu, Di Lin 等AAAI 2026
- Discernible Image CompressionZhaohui Yang, Yunhe Wang, Chang Xu, Peng Du 等ACM MM 2020 · 被引用 25 次
- Differentiable Vector Quantization for Rate-Distortion Optimization of Generative Image CompressionShiyin Jiang, Wei Long, Minghao Han, Zhenghao Chen 等CVPR 2026 · 被引用 3 次
- Just Noticeable Difference Modeling for Deep Visual FeaturesRui Zhao, Wenrui Li, Lin Zhu, Yajing Zheng 等ICML 2026 · 被引用 1 次
