Deep Residual Learning in the JPEG Transform Domain
Max Ehrlich, Larry Davis
摘要
We introduce a general method of performing Residual Network inference and learning in the JPEG transform domain that allows the network to consume compressed images as input. Our formulation leverages the linearity of the JPEG transform to redefine convolution and batch normalization with a tune-able numerical approximation for ReLu. The result is mathematically equivalent to the spatial domain network up to the ReLu approximation accuracy. A formulation for image classification and a model conversion algorithm for spatial domain networks are given as examples of the method. We show skipping the costly decompression step allows for faster processing of images with little to no penalty in the network accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- FcaNet: Frequency Channel Attention NetworksZequn Qin, Pengyi Zhang, Fei Wu, Xi LiICCV 2021 · 被引用 1,049 次
- Detecting Camouflaged Object in Frequency DomainYijie Zhong, Bo Li, Lv Tang, Senyun Kuang 等CVPR 2022 · 被引用 271 次
- Fourmer: An Efficient Global Modeling Paradigm for Image RestorationMan Zhou, Jie Huang, Chun-Le Guo, Chongyi LiICML 2023 · 被引用 148 次
- Frequency Perception Network for Camouflaged Object DetectionRunmin Cong, Mengyao Sun, Sanyi Zhang, Xiaofei Zhou 等ACM MM 2023 · 被引用 130 次
- Towards Discriminative Representation Learning for Unsupervised Person Re-identificationTakashi Isobe, Dong Li, Lu Tian, Weihua Chen 等ICCV 2021 · 被引用 76 次
相关 Paper
- Practical Learned Lossless JPEG Recompression with Multi-Level Cross-Channel Entropy Model in the DCT DomainLina Guo, Xinjie Shi, Dailan He, Yuanyuan Wang 等CVPR 2022 · 被引用 8 次
- JPEG-ACT: Accelerating Deep Learning via Transform-based Lossy CompressionR. David Evans, Lufei Liu, Tor M. AamodtISCA 2020 · 被引用 47 次
- Network DeconvolutionChengxi Ye, Matthew Evanusa, Hua He, Anton Mitrokhin 等ICLR 2020
- JPEG Artifacts Reduction via Deep Convolutional Sparse CodingXueyang Fu, Zheng-Jun Zha, Feng Wu, Xinghao Ding 等ICCV 2019 · 被引用 117 次
- RGB No More: Minimally-Decoded JPEG Vision TransformersJeongsoo Park, Justin JohnsonCVPR 2023
