Composition and Alignment of Diffusion Models using Constrained Learning
Shervin Khalafi, Ignacio Hounie, Dongsheng Ding, Alejandro Ribeiro
摘要
Diffusion models have become prevalent in generative modeling due to their ability to sample from complex distributions. To improve the quality of generated samples and their compliance with user requirements, two commonly used methods are: (i) Alignment, which involves finetuning a diffusion model to align it with a reward; and (ii) Composition, which combines several pretrained diffusion models together, each emphasizing a desirable attribute in the generated outputs. However, trade-offs often arise when optimizing for multiple rewards or combining multiple models, as they can often represent competing properties. Existing methods cannot guarantee that the resulting model faithfully generates samples with all the desired properties. To address this gap, we propose a constrained optimization framework that unifies alignment and composition of diffusion models by enforcing that the aligned model satisfies reward constraints and/or remains close to each pretrained model. We provide a theoretical characterization of the solutions to the constrained alignment and composition problems and develop a Lagrangian-based primal-dual training algorithm to approximate these solutions. Empirically, we demonstrate our proposed approach in image generation, applying it to alignment and composition, and show that our aligned or composed model satisfies constraints effectively. Our implementation can be found at: https://github.com/shervinkhalafi/constrained_comp_align
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Calibrating Generative Models to Distributional ConstraintsHenry Smith, Nathaniel Diamant, Brian TrippeICML 2026 · 被引用 4 次
- Constrained Flow Optimization via Sequential Fine-Tuning for Molecular DesignSven Gutjahr, Riccardo De Santi, Luca Schaufelberger, Kjell Jorner 等ICML 2026 · 被引用 3 次
- Unlearning in Diffusion Models: A Unified Framework with KL Divergence and Likelihood ConstraintsShervin Khalafi, Alejandro Ribeiro, Dongsheng DingICML 2026
- Composition of Pretrained Diffusion Models: A Logic-Based CalculusPeter Blohm, Vikas K GargICLR 2026
它引用的顶会 Paper36
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 被引用 6,549 次
相关 Paper
- Constrained Diffusion Models via Dual TrainingShervin Khalafi, Dongsheng Ding, Alejandro RibeiroNeurIPS 2024 · 被引用 24 次
- Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion ModelsMin Cheng, Fatemeh Doudi, Dileep Kalathil, Mohammad Ghavamzadeh 等ICLR 2026 · 被引用 6 次
- CRAFT: Aligning Diffusion Models with Fine-Tuning Is Easier Than You ThinkZening Sun, Zhengpeng Xie, Lichen Bai, Shitong Shao 等CVPR 2026 · 被引用 3 次
- Training-Free Constrained Generation With Stable Diffusion ModelsStefano Zampini, Jacob K. Christopher, Luca Oneto, Davide Anguita 等NeurIPS 2025 · 被引用 21 次
- Direct Consistency Optimization for Robust Customization of Text-to-Image Diffusion modelsKyungmin Lee, Sangkyung Kwak, Kihyuk Sohn, Jinwoo ShinNeurIPS 2024 · 被引用 13 次
