DISCO: DISCrete nOise for Conditional Control in Text-to-Image Diffusion Models
Longquan Dai, Ming Wu, Dejiao Xue, He Wang, Jinhui Tang
Abstract
A major challenge in using diffusion models is aligning outputs with user-defined conditions. Existing conditional generation methods fall into two major categories: classifier-based guidance, which requires differentiable target models and gradientbased correction; and classifier-free guidance, which embeds conditions directly into the diffusion model but demands expensive joint training and architectural coupling. In this work, we introduce a third paradigm: DISCrete nOise (DISCO) guidance, which replaces the continuous conditional correction term with a finite codebook of discrete noise vectors sampled from a Gaussian prior. Conditional generation is reformulated as a code selection task, and we train prediction network to choose the optimal code given the intermediate diffusion state and the conditioning input. Our approach is differentiability-free, and training-efficient, avoiding the gradient computation and architectural redundancy of prior methods. Empirical results demonstrate that DISCO achieves competitive controllability while substantially reducing resource demands, positioning it as a scalable and effective alternative for conditional diffusion generation. Code is available at https://github.com/dailongquan/disco .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4fdec9a7-855d-4ef6-8faa-81c93a30b88fBuilds on30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- Simple Guidance Mechanisms for Discrete Diffusion ModelsYair Schiff, Subham Sekhar Sahoo, Hao Phung, Guanghan Wang et al.ICLR 2025
- Domain Guidance: A Simple Transfer Approach for a Pre-trained Diffusion ModelJincheng Zhong, Xiangcheng Zhang, Jianmin Wang, Mingsheng LongICLR 2025
- Mitigating Noise Shift in Denoising Generative Models with Noise Awareness GuidanceJincheng Zhong, Boyuan Jiang, Xin Tao, Pengfei Wan et al.ICLR 2026 · 1 citation
- Elucidating the design space of classifier-guided diffusion generationJiajun Ma, Tianyang Hu, Wenjia Wang, Jiacheng SunICLR 2024 · 24 citations
- Inner Classifier-Free Guidance and Its Taylor Expansion for Diffusion ModelsShikun Sun, Longhui Wei, Zhicai Wang, Zixuan Wang et al.ICLR 2024 · 2 citations
