Diffusion Based Representation Learning
Sarthak Mittal, Korbinian Abstreiter, Stefan Bauer, Bernhard Schölkopf, Arash Mehrjou
Abstract
Diffusion-based methods, represented as stochastic differential equations on a continuous-time domain, have recently proven successful as nonadversarial generative models. Training such models relies on denoising score matching, which can be seen as multi-scale denoising autoencoders. Here, we augment the denoising score matching framework to enable representation learning without any supervised signal. GANs and VAEs learn representations by directly transforming latent codes to data samples. In contrast, the introduced diffusion-based representation learning relies on a new formulation of the denoising score matching objective and thus encodes the information needed for denoising. We illustrate how this difference allows for manual control of the level of details encoded in the representation. Using the same approach, we propose to learn an infinite-dimensional latent code that achieves improvements on state-of-the-art models on semisupervised image classification. We also compare the quality of learned representations of diffusion score matching with other methods like autoencoder and contrastively trained systems through their performances on downstream tasks. Finally, we also ablate with a different SDE formulation for diffusion models and show that the benefits on downstream tasks are still present on changing the underlying differential equation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a4536ace-f1dc-4937-91fd-ba179b51332fCited by top-tier papers27
- Denoising Diffusion Autoencoders are Unified Self-supervised LearnersWeilai Xiang, Hongyu Yang, Di Huang, Yunhong WangICCV 2023 · 145 citations
- Diffusion Model as Representation LearnerXingyi Yang, Xinchao WangICCV 2023 · 100 citations
- What matters for Representation Alignment: Global Information or Spatial Structure?Jaskirat Singh, Xingjian Leng, Zongze Wu, Liang Zheng et al.ICLR 2026 · 84 citations
- Directional diffusion models for graph representation learningRun Yang, Yuling Yang, Fan Zhou, Qiang SunNeurIPS 2023 · 39 citations
- Pre-Training Protein Encoder via Siamese Sequence-Structure Diffusion Trajectory PredictionZuobai Zhang, Minghao Xu, Aurélie C. Lozano, Vijil Chenthamarakshan et al.NeurIPS 2023 · 35 citations
Builds on15
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
Related papers
- A Variational Perspective on Diffusion-Based Generative Models and Score MatchingChin-Wei Huang, Jae Hyun Lim, Aaron C. CourvilleNeurIPS 2021 · 246 citations
- Deconstructing Denoising Diffusion Models for Self-Supervised LearningXinlei Chen, Zhuang Liu, Saining Xie, Kaiming HeICLR 2025 · 17 citations
- SODA: Bottleneck Diffusion Models for Representation LearningDrew A. Hudson, Daniel Zoran, Mateusz Malinowski, Andrew K. Lampinen et al.CVPR 2024
- ProgDiffusion: Progressively Self-encoding Diffusion ModelsZhangkai Wu, Xuhui Fan, Longbing CaoKDD 2025 · 3 citations
- Boosting Generative Image Modeling via Joint Image-Feature SynthesisTheodoros Kouzelis, Efstathios Karypidis, Ioannis Kakogeorgiou, Spyridon Gidaris et al.NeurIPS 2025 · 47 citations
