Beyond Textual Constraints: Learning Novel Diffusion Conditions with Fewer Examples
Yuyang Yu, Bangzhen Liu, Chenxi Zheng, Xuemiao Xu, Shengfeng He, Huaidong Zhang
Abstract
In this paper, we delve into a novel aspect of learning novel diffusion conditions with datasets an order of magnitude smaller. The rationale behind our approach is the elimination of textual constraints during the few-shot learning process. To that end, we implement two optimization strategies. The first, prompt-free conditional learning, utilizes a prompt-free encoder derived from a pre-trained Stable Diffusion model. This strategy is designed to adapt new conditions to the diffusion process by minimizing the textual-visual cor-relation, thereby ensuring a more precise alignment between the generated content and the specified conditions. The second strategy entails condition-specific negative rectification, which addresses the inconsistencies typically brought about by Classifier-free guidance in few-shot training con-texts. Our extensive experiments across a variety of condition modalities demonstrate the effectiveness and efficiency of our framework, yielding results comparable to those obtained with datasets a thousand times larger. Our codes are available at https://github.com/Yuyan9Yu/BeyondTextConstraint.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic PropagationHaofeng Liu, Chenshu Xu, Yifei Yang, Lihua Zeng et al.CVPR 2024 · 12 citations
- Neptune-X: Active X-to-Maritime Generation for Universal Maritime Object DetectionYu Guo, Shengfeng He, Yuxu Lu, Haonan An et al.NeurIPS 2025 · 7 citations
- StableGuard: Towards Unified Copyright Protection and Tamper Localization in Latent Diffusion ModelsHaoxin Yang, Bangzhen Liu, Xuemiao Xu, Cheng Xu et al.NeurIPS 2025 · 5 citations
- Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch GenerationRui Yang, Huining Li, Yiyi Long, Xiaojun Wu et al.ICCV 2025 · 2 citations
- FR²Seg: Continual Segmentation Across Multiple Sites via Fourier Style Replay and Adaptive Consistency RegularizationCheng Xu, Weiwen Zhang, Hongrui Zhang, Xuemiao Xu et al.AAAI 2025 · 1 citation
Builds on26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Universal Few-shot Spatial Control for Diffusion ModelsKiet T. Nguyen, Chanhyuk Lee, Donggyun Kim, Dong Hoon Lee et al.NeurIPS 2025 · 2 citations
- SD-FSMIS: Adapting Stable Diffusion for Few-Shot Medical Image SegmentationMeihua Li, Yang Zhang, Weizhao He, Hu Qu et al.CVPR 2026 · 1 citation
- DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric FinetuningYuxuan Duan, Yan Hong, Bo Zhang, Jun Lan et al.NeurIPS 2024 · 2 citations
- D2C: Diffusion-Decoding Models for Few-Shot Conditional GenerationAbhishek Sinha, Jiaming Song, Chenlin Meng, Stefano ErmonNeurIPS 2021 · 149 citations
- Understanding and Improving Training-free Loss-based Diffusion GuidanceYifei Shen, Xinyang Jiang, Yifan Yang, Yezhen Wang et al.NeurIPS 2024 · 36 citations
