Deep Residual Learning in the JPEG Transform Domain
Max Ehrlich, Larry Davis
Abstract
We introduce a general method of performing Residual Network inference and learning in the JPEG transform domain that allows the network to consume compressed images as input. Our formulation leverages the linearity of the JPEG transform to redefine convolution and batch normalization with a tune-able numerical approximation for ReLu. The result is mathematically equivalent to the spatial domain network up to the ReLu approximation accuracy. A formulation for image classification and a model conversion algorithm for spatial domain networks are given as examples of the method. We show skipping the costly decompression step allows for faster processing of images with little to no penalty in the network accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8dd72b08-2658-47ad-87f1-298fc49f6e48Cited by top-tier papers18
- FcaNet: Frequency Channel Attention NetworksZequn Qin, Pengyi Zhang, Fei Wu, Xi LiICCV 2021 · 1,049 citations
- Detecting Camouflaged Object in Frequency DomainYijie Zhong, Bo Li, Lv Tang, Senyun Kuang et al.CVPR 2022 · 271 citations
- Fourmer: An Efficient Global Modeling Paradigm for Image RestorationMan Zhou, Jie Huang, Chun-Le Guo, Chongyi LiICML 2023 · 148 citations
- Frequency Perception Network for Camouflaged Object DetectionRunmin Cong, Mengyao Sun, Sanyi Zhang, Xiaofei Zhou et al.ACM MM 2023 · 130 citations
- Towards Discriminative Representation Learning for Unsupervised Person Re-identificationTakashi Isobe, Dong Li, Lu Tian, Weihua Chen et al.ICCV 2021 · 76 citations
Related papers
- Practical Learned Lossless JPEG Recompression with Multi-Level Cross-Channel Entropy Model in the DCT DomainLina Guo, Xinjie Shi, Dailan He, Yuanyuan Wang et al.CVPR 2022 · 8 citations
- JPEG-ACT: Accelerating Deep Learning via Transform-based Lossy CompressionR. David Evans, Lufei Liu, Tor M. AamodtISCA 2020 · 47 citations
- Network DeconvolutionChengxi Ye, Matthew Evanusa, Hua He, Anton Mitrokhin et al.ICLR 2020
- JPEG Artifacts Reduction via Deep Convolutional Sparse CodingXueyang Fu, Zheng-Jun Zha, Feng Wu, Xinghao Ding et al.ICCV 2019 · 117 citations
- RGB No More: Minimally-Decoded JPEG Vision TransformersJeongsoo Park, Justin JohnsonCVPR 2023
