Model-Based Image Signal Processors via Learnable Dictionaries
Marcos V. Conde, Steven McDonagh, Matteo Maggioni, Ales Leonardis, Eduardo Pérez-Pellitero
摘要
Digital cameras transform sensor RAW readings into RGB images by means of their Image Signal Processor (ISP). Computational photography tasks such as image denoising and colour constancy are commonly performed in the RAW domain, in part due to the inherent hardware design, but also due to the appealing simplicity of noise statistics that result from the direct sensor readings. Despite this, the availability of RAW images is limited in comparison with the abundance and diversity of available RGB data. Recent approaches have attempted to bridge this gap by estimating the RGB to RAW mapping: handcrafted model-based methods that are interpretable and controllable usually require manual parameter fine-tuning, while end-to-end learnable neural networks require large amounts of training data, at times with complex training procedures, and generally lack interpretability and parametric control. Towards addressing these existing limitations, we present a novel hybrid model-based and data-driven ISP that builds on canonical ISP operations and is both learnable and interpretable. Our proposed invertible model, capable of bidirectional mapping between RAW and RGB domains, employs end-to-end learning of rich parameter representations, i.e. dictionaries, that are free from direct parametric supervision and additionally enable simple and plausible data augmentation. We evidence the value of our data generation process by extensive experiments under both RAW image reconstruction and RAW image denoising tasks, obtaining state-of-the-art performance in both. Additionally, we show that our ISP can learn meaningful mappings from few data samples, and that denoising models trained with our dictionary-based data augmentation are competitive despite having only few or zero ground-truth labels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- RAW-Flow: Advancing RGB-to-RAW Image Reconstruction with Deterministic Latent Flow MatchingZhen Liu, Diedong Feng, Hai Jiang, Liaoyuan Zeng 等AAAI 2026 · 被引用 3 次
- Bridging RGB and RAW: Single-step Deterministic Flow with Homogeneous Representation AlignmentDiedong Feng, Peiyi Zeng, Zhen Liu, Zhongyang Li 等ICML 2026
它引用的顶会 Paper4
- CycleISP: Real Image Restoration via Improved Data SynthesisSyed Waqas Zamir, Aditya Arora, Salman H. Khan, Munawar Hayat 等CVPR 2020
- Single-Image HDR Reconstruction by Learning to Reverse the Camera PipelineYu-Lun Liu, Wei-Sheng Lai, Yu-Sheng Chen, Yi-Lung Kao 等CVPR 2020
- Invertible Image Signal ProcessingYazhou Xing, Zian Qian, Qifeng ChenCVPR 2021
- A Multi-Hypothesis Approach to Color ConstancyDaniel Hernández Juárez, Sarah Parisot, Benjamin Busam, Ales Leonardis 等CVPR 2020
相关 Paper
- Modeling sRGB Camera Noise with Normalizing FlowsShayan Kousha, Ali Maleky, Michael S. Brown, Marcus A. BrubakerCVPR 2022 · 被引用 21 次
- ParamISP: Learned Forward and Inverse ISPs Using Camera ParametersWoohyeok Kim, Geonu Kim, Junyong Lee, Seungyong Lee 等CVPR 2024
- Learning Degradation-Independent Representations for Camera ISP PipelinesYanhui Guo, Fangzhou Luo, Xiaolin WuCVPR 2024
- Edit-aware RAW reconstructionAbhijith Punnappurath, Luxi Zhao, Ke Zhao, Hue Nguyen 等CVPR 2026
- ISPDiffuser: Learning RAW-to-sRGB Mappings with Texture-Aware Diffusion Models and Histogram-Guided Color ConsistencyYang Ren, Hai Jiang, Menglong Yang, Wei Li 等AAAI 2025 · 被引用 7 次
