Understanding the Latent Space of Diffusion Models through the Lens of Riemannian Geometry
Yong-Hyun Park, Mingi Kwon, Jaewoong Choi, Junghyo Jo, Youngjung Uh
Abstract
Despite the success of diffusion models (DMs), we still lack a thorough understanding of their latent space. To understand the latent space , we analyze them from a geometrical perspective. Our approach involves deriving the local latent basis within by leveraging the pullback metric associated with their encoding feature maps. Remarkably, our discovered local latent basis enables image editing capabilities by moving , the latent space of DMs, along the basis vector at specific timesteps. We further analyze how the geometric structure of DMs evolves over diffusion timesteps and differs across different text conditions. This confirms the known phenomenon of coarse-to-fine generation, as well as reveals novel insights such as the discrepancy between across timesteps, the effect of dataset complexity, and the time-varying influence of text prompts. To the best of our knowledge, this paper is the first to present image editing through -space traversal, editing only once at specific timestep without any additional training, and providing thorough analyses of the latent structure of DMs. The code to reproduce our experiments can be found at https://github.com/enkeejunior1/Diffusion-Pullback.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0b8aba3-bf28-426f-bd5a-92028e5d4543Cited by top-tier papers58
- Smoothed Energy Guidance: Guiding Diffusion Models with Reduced Energy Curvature of AttentionSusung HongNeurIPS 2024 · 67 citations
- Interpreting the Weight Space of Customized Diffusion ModelsAmil Dravid, Yossi Gandelsman, Kuan-Chieh Wang, Rameen Abdal et al.NeurIPS 2024 · 40 citations
- One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion ModelsViacheslav Surkov, Chris Wendler, Antonio Mari, Mikhail Terekhov et al.NeurIPS 2025 · 33 citations
- Exploring Low-Dimensional Subspace in Diffusion Models for Controllable Image EditingSiyi Chen, Huijie Zhang, Minzhe Guo, Yifu Lu et al.NeurIPS 2024 · 29 citations
- Emergence and Evolution of Interpretable Concepts in Diffusion ModelsBerk Tinaz, Zalan Fabian, Mahdi SoltanolkotabiNeurIPS 2025 · 23 citations
Builds on35
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- NoiseCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable Directions in Diffusion ModelsYusuf Dalva, Pinar YanardagCVPR 2024
- Isometric Representation Learning for Disentangled Latent Space of Diffusion ModelsJaehoon Hahm, Junho Lee, Sunghyun Kim, Joonseok LeeICML 2024 · 21 citations
- Smooth Diffusion: Crafting Smooth Latent Spaces in Diffusion ModelsJiayi Guo, Xingqian Xu, Yifan Pu, Zanlin Ni et al.CVPR 2024
- Blended Latent DiffusionOmri Avrahami, Ohad Fried, Dani LischinskiSIGGRAPH 2023 · 339 citations
- PIXELS: Progressive Image Xemplar-based Editing with Latent SurgeryShristi Das Biswas, Matthew Shreve, Xuelu Li, Prateek Singhal et al.AAAI 2025 · 2 citations
