CoInD: Enabling Logical Compositions in Diffusion Models
Sachit Gaudi, Gautam Sreekumar, Vishnu Boddeti
摘要
How can we learn generative models to sample data with arbitrary logical compositions of statistically independent attributes? The prevailing solution is to sample from distributions expressed as a composition of attributes' conditional marginal distributions under the assumption that they are statistically independent. This paper shows that standard conditional diffusion models violate this assumption, even when all attribute compositions are observed during training. And, this violation is significantly more severe when only a subset of the compositions is observed. We propose COIND to address this problem. It explicitly enforces statistical independence between the conditional marginal distributions by minimizing Fisher's divergence between the joint and marginal distributions. The theoretical advantages of COIND are reflected in both qualitative and quantitative experiments, demonstrating a significantly more faithful and controlled generation of samples for arbitrary logical compositions of attributes. The benefit is more pronounced for scenarios that current solutions relying on the assumption of conditionally independent marginals struggle with, namely, logical compositions involving the NOT operation and when only a subset of compositions are observed during training. Our code is available at https://github.com/sachit3022/compositional-generation/ 1 INTRODUCTION 0 1 2 3 4 5 6 7 8 9 Digit (C 1 ) Color (C 2 ) (a) Uniform 0 1 2 3 4 5 6 7 8 9 Digit (C 1 ) Color (C 2 ) (b) Non-uniform 0 1 2 3 4 5 6 7 8 9 Digit (C 1 ) Color (C 2 ) (c) Partial support Composed GLIDE COIND Composed GLIDE COIND Composed GLIDE COIND Uniform Non-uniform Partial 4 ∧ Pink 4 ∧ Cyan 4 ∧ ¬(Pink ∨ Green) ¬(3 ∨ 4) ∧ Pink Support Method (d) Generated samples
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
- Composition of Pretrained Diffusion Models: A Logic-Based CalculusPeter Blohm, Vikas K GargICLR 2026
它引用的顶会 Paper17
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
相关 Paper
- Compositional Abilities Emerge Multiplicatively: Exploring Diffusion Models on a Synthetic TaskMaya Okawa, Ekdeep Singh Lubana, Robert P. Dick, Hidenori TanakaNeurIPS 2023 · 被引用 113 次
- Any-to-Any Generation via Composable DiffusionZineng Tang, Ziyi Yang, Chenguang Zhu, Michael Zeng 等NeurIPS 2023 · 被引用 294 次
- Logical Guidance for the Exact Composition of Diffusion ModelsFrancesco Alesiani, Jonathan Warrell, Tanja Bien, Henrik Christiansen 等ICML 2026
- Locally Coherent Parallel Decoding in Diffusion Language ModelsMichael Hersche, Nicolas Menet, Ronan Tanios, Abbas RahimiICML 2026 · 被引用 1 次
- CoDi: Co-evolving Contrastive Diffusion Models for Mixed-type Tabular SynthesisChaejeong Lee, Jayoung Kim, Noseong ParkICML 2023 · 被引用 97 次
