Isometric Representation Learning for Disentangled Latent Space of Diffusion Models
Jaehoon Hahm, Junho Lee, Sunghyun Kim, Joonseok Lee
摘要
The latent space of diffusion model mostly still remains unexplored, despite its great success and potential in the field of generative modeling. In fact, the latent space of existing diffusion models are entangled, with a distorted mapping from its latent space to image space. To tackle this problem, we present Isometric Diffusion, equipping a diffusion model with a geometric regularizer to guide the model to learn a geometrically sound latent space of the training data manifold. This approach allows diffusion models to learn a more disentangled latent space, which enables smoother interpolation, more accurate inversion, and more precise control over attributes directly in the latent space. Our extensive experiments consisting of image interpolations, image inversions, and linear editing show the effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Latent Wavelet Diffusion For Ultra High-Resolution Image SynthesisLuigi Sigillo, Shengfeng He, Danilo ComminielloICLR 2026 · 被引用 8 次
- SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse AutoencodersEnrico Cassano, Riccardo Renzulli, Marco Nurisso, Mirko Zaffaroni 等ICML 2026 · 被引用 7 次
- Can Diffusion Models Disentangle? A Theoretical PerspectiveLiming Wang, Muhammad Jehanzeb Mirza, Yishu Gong, Yuan Gong 等NeurIPS 2025 · 被引用 4 次
- Latent Diffusion Models With Masked AutoencodersJunho Lee, Jeongwoo Shin, Hyungwook Choi, Joonseok LeeICCV 2025 · 被引用 3 次
- Contrastive Diffusion Alignment: Learning Structured Latents for Controllable GenerationRuchi Sandilya, Sumaira Perez, Charles Lynch, Lindsay Victoria 等ICML 2026 · 被引用 1 次
它引用的顶会 Paper23
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Smooth Diffusion: Crafting Smooth Latent Spaces in Diffusion ModelsJiayi Guo, Xingqian Xu, Yifan Pu, Zanlin Ni 等CVPR 2024
- InfoDiffusion: Representation Learning Using Information Maximizing Diffusion ModelsYingheng Wang, Yair Schiff, Aaron Gokaslan, Weishen Pan 等ICML 2023 · 被引用 64 次
- Hyperbolic Geometric Latent Diffusion Model for Graph GenerationXingcheng Fu, Yisen Gao, Yuecen Wei, Qingyun Sun 等ICML 2024 · 被引用 31 次
- NoiseCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable Directions in Diffusion ModelsYusuf Dalva, Pinar YanardagCVPR 2024
- Understanding the Latent Space of Diffusion Models through the Lens of Riemannian GeometryYong-Hyun Park, Mingi Kwon, Jaewoong Choi, Junghyo Jo 等NeurIPS 2023 · 被引用 163 次
