MERIT: Multi-domain Efficient RAW Image Translation
Wenjun Huang, Shenghao Fu, Yian Jin, Yang Ni, Ziteng Cui, Hanning Chen, Yirui He, Yezi Liu, Sanggeon Yun, SungHeon Jeong, Ryozo Masukawa, William Youngwoo Chung, Mohsen Imani
Abstract
RAW images captured by different camera sensors exhibit substantial domain shifts due to varying spectral responses, noise characteristics, and tone behaviors, complicating their direct use in downstream computer vision tasks. Prior methods address this problem by training domain-specific RAW-to-RAW translators for each source-target pair, but such approaches do not scale to real-world scenarios involving multiple types of commercial cameras. In this work, we introduce MERIT, the first unified framework for multidomain RAW image translation, which leverages a single model to perform translations across arbitrary camera domains. To address domain-specific noise discrepancies, we propose a sensor-aware noise modeling loss that explicitly aligns the signal-dependent noise statistics of the generated images with those of the target domain. We further enhance the generator with a conditional multi-scale large kernel attention module for improved context and sensor-aware feature modeling. To facilitate standardized evaluation, we introduce MDRAW, the first dataset tailored for multidomain RAW image translation, comprising both paired and unpaired RAW captures from five diverse camera sensors across a wide range of scenes. Extensive experiments demonstrate that MERIT outperforms prior models in both quality (+5.56 dB) and scalability (80% reduction in training iterations). Our code is available here.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2ac8dc4-9556-4966-9705-55ff4f4e8822Builds on14
- NeRF in the Dark: High Dynamic Range View Synthesis from Noisy Raw ImagesBen Mildenhall, Peter Hedman, Ricardo Martin-Brualla, Pratul P. Srinivasan et al.CVPR 2022 · 307 citations
- AdaptiveISP: Learning an Adaptive Image Signal Processor for Object DetectionYujin Wang, Tianyi Xu, Zhang Fan, Tianfan Xue et al.NeurIPS 2024 · 38 citations
- Lighting Every Darkness with 3DGS: Fast Training and Real-Time Rendering for HDR View SynthesisXin Jin, Pengyi Jiao, Zheng-Peng Duan, Xingchao Yang et al.NeurIPS 2024 · 36 citations
- Swin-UNIT: Transformer-based GAN for High-resolution Unpaired Image TranslationYifan Li, Yaochen Li, Wenneng Tang, Zhifeng Zhu et al.ACM MM 2023 · 13 citations
- Learning to See in the Extremely DarkHai Jiang, Binhao Guan, Zhen Liu, Xiaohong Liu et al.ICCV 2025 · 10 citations
Related papers
- Towards General Low-Light Raw Noise Synthesis and ModelingFeng Zhang, Bin Xu, Zhiqiang Li, Xinran Liu et al.ICCV 2023 · 30 citations
- Adaptive Domain Learning for Cross-domain Image DenoisingZian Qian, Chenyang Qi, Ka Lung Law, Hao Fu et al.NeurIPS 2024 · 1 citation
- RawMetaDiff: Unlocking Extreme Darkness from Dual-Exposure RAW with Meta-Guided DiffusionPanjun Liu, Jiyuan Xia, YUANSHEN GUAN, Yong Li et al.CVPR 2026
- Dr. RAW: Towards General High-Level Vision from RAW with Efficient Task ConditioningWenjun Huang, Ziteng Cui, Yinqiang Zheng, Yirui He et al.NeurIPS 2025 · 5 citations
- SpiralDiff: Spiral Diffusion with LoRA for RGB-to-RAW Conversion Across CamerasHuanjing Yue, Shangbin Xie, Cong Cao, Qian Wu et al.CVPR 2026
