CosAE: Learnable Fourier Series for Image Restoration
Sifei Liu, Shalini De Mello, Jan Kautz
摘要
In this paper, we introduce Cosine Autoencoder (CosAE), a novel, generic Au-toencoder that seamlessly leverages the classic Fourier series with a feed-forward neural network. CosAE represents an input image as a series of 2D Cosine time series, each defined by a tuple of learnable frequency and Fourier coefficients. This method stands in contrast to a conventional Autoencoder that often sacrifices detail in their reduced-resolution bottleneck latent spaces. CosAE, however, encodes frequency coefficients, i.e., the amplitudes and phases, in its bottleneck. This encoding enables extreme spatial compression, e.g., 64 × downsampled feature maps in the bottleneck, without losing detail upon decoding. We showcase the advantage of CosAE via extensive experiments on flexible-resolution super-resolution and blind image restoration, two highly challenging tasks that demand the restoration network to effectively generalize to complex and even unknown image degradations. Our method surpasses state-of-the-art approaches, highlighting its capability to learn a generalizable representation for image restoration. The project page is maintained at https://sifeiliu.net/CosAE-page/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- MVAR: Visual Autoregressive Modeling with Scale and Spatial Markovian ConditioningJinhua Zhang, Wei Long, Minghao Han, Weiyi You 等ICLR 2026 · 被引用 9 次
- Latent Harmony: Synergistic Unified UHD Image Restoration via Latent Space Regularization and Controllable RefinementYidi Liu, Xueyang Fu, Jie Huang, Jie Xiao 等NeurIPS 2025 · 被引用 3 次
- FIPER: Factorized Features for Robust Image Super-Resolution and CompressionYang-Che Sun, Cheng Yu Yeo, Ernie Chu, Jun-Cheng Chen 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper23
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu 等ICML 2020 · 被引用 1,773 次
相关 Paper
- DreamUHD: Frequency Enhanced Variational Autoencoder for Ultra-High-Definition Image RestorationYidi Liu, Dong Li, Jie Xiao, Yuanfei Bao 等AAAI 2025 · 被引用 11 次
- JDEC: JPEG Decoding via Enhanced Continuous Cosine CoefficientsWoo Kyoung Han, Sunghoon Im, Jaedeok Kim, Kyong Hwan JinCVPR 2024
- Continuous Space-Time Video Super-Resolution with 3D Fourier FieldsAlexander Becker, Julius Erbach, Dominik Narnhofer, Konrad SchindlerICLR 2026 · 被引用 3 次
- Spectrum-to-Kernel Translation for Accurate Blind Image Super-ResolutionGuangpin Tao, Xiaozhong Ji, Wenzhuo Wang, Shuo Chen 等NeurIPS 2021 · 被引用 27 次
- Kernel Aware ResamplerMichael Bernasconi, Abdelaziz Djelouah, Farnood Salehi, Markus Gross 等CVPR 2023
