Degradation-Modeled Multipath Diffusion for Tunable Metalens Photography
Jianing Zhang, Jiayi Zhu, Feiyu Ji, Xiaokang Yang, Xiaoyun Yuan
摘要
Metalenses offer significant potential for ultra-compact computational imaging but face challenges from complex optical degradation and computational restoration difficulties. Existing methods typically rely on precise optical calibration or massive paired datasets, which are non-trivial for real-world imaging systems. Furthermore, a lack of control over the inference process often results in undesirable hallucinated artifacts. We introduce Degradation-Modeled Multipath Diffusion for tunable metalens photography, leveraging powerful natural image priors from pretrained models instead of large datasets. Our framework uses positive, neutral, and negative-prompt paths to balance high-frequency detail generation, structural fidelity, and suppression of metalens-specific degradation, alongside pseudo data augmentation. A tunable decoder enables controlled trade-offs between fidelity and perceptual quality. Additionally, a spatially varying degradation-aware attention (SVDA) module adaptively models complex optical and sensor-induced degradation. Finally, we design and build a millimeter-scale MetaCamera for real-world validation. Extensive results show that our approach outperforms state-of-the-art methods, achieving high-fidelity and sharp image reconstruction. More materials: https://dmdiff.github.io/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Learning Latent Transmission and Glare Maps for Lens Veiling Glare RemovalXiaolong Qian, Qi Jiang, Lei Sun, Zongxi Yu 等CVPR 2026 · 被引用 4 次
- 3M-TI: High-Quality Mobile Thermal Imaging via Calibration-free Multi-Camera Cross-Modal DiffusionMinchong Chen, Xiaoyun Yuan, Junzhe Wan, Jianing Zhang 等CVPR 2026 · 被引用 2 次
- LUCID: Learning Unified Control for Image Deflaring and Exposure Mastery in Nighttime PhotographyTingyu Yang, Yuan Cheng, Xiaoyun YuanSIGGRAPH 2026
它引用的顶会 Paper14
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Uformer: A General U-Shaped Transformer for Image RestorationZhendong Wang, Xiaodong Cun, Jianmin Bao, Wengang Zhou 等CVPR 2022 · 被引用 1,970 次
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar 等ICCV 2021 · 被引用 1,325 次
相关 Paper
- RawMetaDiff: Unlocking Extreme Darkness from Dual-Exposure RAW with Meta-Guided DiffusionPanjun Liu, Jiyuan Xia, YUANSHEN GUAN, Yong Li 等CVPR 2026
- Towards Illumination-Aware Restoration of Metalens-Captured Images: A New Dataset and a Strong BaselineFen Fang, Xinan Liang, Muli Yang, Jinghong Zheng 等AAAI 2026 · 被引用 1 次
- Extreme-Quality Computational Imaging via Degradation FrameworkShiqi Chen, Huajun Feng, Keming Gao, Zhihai Xu 等ICCV 2021 · 被引用 33 次
- Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationYuang Ai, Huaibo Huang, Xiaoqiang Zhou, Jiexiang Wang 等CVPR 2024
- UniRes: Universal Image Restoration for Complex DegradationsMo Zhou, Keren Ye, Mauricio Delbracio, Peyman Milanfar 等ICCV 2025
