Patronus: Interpretable Diffusion Models with Prototypes
Nina Weng, Aasa Feragen, Siavash Siavash Bigdeli
摘要
Uncovering the opacity of diffusion-based generative models is urgently needed, as their applications continue to expand while their underlying procedures largely remain a black box. With a critical question -- how can the diffusion generation process be interpreted and understood? -- we proposed Patronus, an interpretable diffusion model that incorporates a prototypical network to encode semantics in visual patches, revealing what visual patterns are learned and where and when they emerge throughout denoising. This interpretability of Patronus provides deeper insights into the generative mechanism, enabling the detection of shortcut learning via unwanted correlations and the tracing of semantic emergence across timesteps. We evaluate Patronus on four natural image datasets and one medical imaging dataset, demonstrating both faithful interpretability and strong generative performance. With this work, we open new avenues for understanding and steering diffusion models through prototype-based interpretability. Our code is available at nina-weng.github.io/patronus.github.io.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Autoencoders: Toward a Meaningful and Decodable RepresentationKonpat Preechakul, Nattanat Chatthee, Suttisak Wizadwongsa, Supasorn SuwajanakornCVPR 2022 · 被引用 276 次
- Understanding the Latent Space of Diffusion Models through the Lens of Riemannian GeometryYong-Hyun Park, Mingi Kwon, Jaewoong Choi, Junghyo Jo 等NeurIPS 2023 · 被引用 163 次
- On Provable Copyright Protection for Generative ModelsNikhil Vyas, Sham M. Kakade, Boaz BarakICML 2023 · 被引用 120 次
- FreeU: Free Lunch in Diffusion U-NetChenyang Si, Ziqi Huang, Yuming Jiang, Ziwei LiuCVPR 2024 · 被引用 111 次
相关 Paper
- NoiseCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable Directions in Diffusion ModelsYusuf Dalva, Pinar YanardagCVPR 2024
- PIP-Net: Patch-Based Intuitive Prototypes for Interpretable Image ClassificationMeike Nauta, Jörg Schlötterer, Maurice van Keulen, Christin SeifertCVPR 2023
- Interpretable Image Classification via Non-parametric Part Prototype LearningZhijie Zhu, Lei Fan, Maurice Pagnucco, Yang SongCVPR 2025
- ProgDiffusion: Progressively Self-encoding Diffusion ModelsZhangkai Wu, Xuhui Fan, Longbing CaoKDD 2025 · 被引用 3 次
- Revelio: Interpreting and Leveraging Semantic Information in Diffusion ModelsDahye Kim, Xavier Thomas, Deepti GhadiyaramICCV 2025 · 被引用 1 次
