Image Neural Field Diffusion Models
Yinbo Chen, Oliver Wang, Richard Zhang, Eli Shechtman, Xiaolong Wang, Michaël Gharbi
摘要
Diffusion models have shown an impressive ability to model complex data distributions, with several key advantages over GANs, such as stable training, better coverage of the training distribution's modes, and the ability to solve inverse problems without extra training. However, most diffusion models learn the distribution of fixed-resolution images. We propose to learn the distribution of continuous images by training diffusion models on image neural fields, which can be rendered at any resolution, and show its advantages over fixed-resolution models. To achieve this, a key challenge is to obtain a latent space that represents photorealistic image neural fields. We propose a simple and effective method, inspired by several recent techniques but with key changes to make the image neural fields photo-realistic. Our method can be used to convert existing latent diffusion autoencoders into image neural field autoen-coders. We show that image neural field diffusion models can be trained using mixed-resolution image datasets, outperform fixed-resolution diffusion models followed by super-resolution models, and can solve inverse problems with conditions applied at different scales efficiently.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- PixNerd: Pixel Neural Field DiffusionShuai Wang, Ziteng Gao, Chenhui Zhu, Weilin Huang 等ICLR 2026 · 被引用 78 次
- Optimize Any Topology: A Foundation Model for Shape- and Resolution-Free Structural Topology OptimizationAmin Heyrani Nobari, Lyle Regenwetter, Cyril Picard, Ligong Han 等NeurIPS 2025 · 被引用 6 次
- Progressive Artwork Outpainting Via Latent Diffusion ModelsDae-Young Song, Jung-Jae Yu, Donghyeon ChoICCV 2025 · 被引用 2 次
- NeRV-Diffusion: Diffuse Implicit Neural Representation for Video SynthesisYixuan Ren, Hanyu Wang, Bo He, Hao Chen 等ICLR 2026 · 被引用 1 次
- I-INR: Iterative Implicit Neural RepresentationsAli Haider, Muhammad Salman Ali, Maryam Qamar, Tahir Khalil 等AAAI 2026 · 被引用 1 次
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Neural Diffusion ModelsGrigory Bartosh, Dmitry P. Vetrov, Christian A. NaessethICML 2024 · 被引用 69 次
- Implicit Diffusion Models for Continuous Super-ResolutionSicheng Gao, Xuhui Liu, Bohan Zeng, Sheng Xu 等CVPR 2023
- Latent Diffusion for Language GenerationJustin Lovelace, Varsha Kishore, Chao Wan, Eliot Shekhtman 等NeurIPS 2023 · 被引用 177 次
- SILO: Solving Inverse Problems with Latent OperatorsRon Raphaeli, Sean Man, Michael EladICCV 2025
- DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete LatentsYilun Xu, Gabriele Corso, Tommi S. Jaakkola, Arash Vahdat 等ICML 2024 · 被引用 22 次
