Few-Shot Diffusion Models Escape the Curse of Dimensionality
Ruofeng Yang, Bo Jiang, Cheng Chen, Ruinan Jin, Baoxiang Wang, Shuai Li
摘要
While diffusion models have demonstrated impressive performance, there is a growing need for generating samples tailored to specific user-defined concepts. The customized requirements promote the development of few-shot diffusion models, which use limited n ta target samples to fine-tune a pre-trained diffusion model trained on n s source samples. Despite the empirical success, no theoretical work specifically analyzes few-shot diffusion models. Moreover, the existing results for diffusion models without a fine-tuning phase can not explain why few-shot models generate great samples due to the curse of dimensionality. In this work, we analyze few-shot diffusion models under a linear structure distribution with a latent dimension d . From the approximation perspective, we prove that few-shot models have a (cid:101) O ( n − 2 /d s + n − 1 / 2 ta ) bound to approximate the target score function, which is better than n − 2 /d ta results. From the optimization perspective, we consider a latent Gaussian special case and prove that the optimization problem has a closed-form minimizer. This means few-shot models can directly obtain an approximated minimizer without a complex optimization process. Furthermore, we also provide the accuracy bound (cid:101) O (1 /n ta + 1 / √ n s ) for the empirical solution, which still has better dependence on n ta compared to n s . The results of the real-world experiments also show that the models obtained by only fine-tuning the encoder and decoder specific to the target distribution can produce novel images with the target feature, which supports our theoretical results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and MethodHaochen Wang, Xiangtai Li, Zilong Huang, Anran Wang 等ICLR 2026 · 被引用 45 次
- Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMsHaochen Wang, Yuhao Wang, Tao Zhang, Yikang Zhou 等ICLR 2026 · 被引用 18 次
- Generalization of Diffusion Models Arises with a Balanced Representation SpaceZekai Zhang, Xiao Li, Xiang Li, Lianghe Shi 等ICLR 2026 · 被引用 14 次
- Provable Sample-Efficient Transfer Learning Conditional Diffusion Models via Representation LearningZiheng Cheng, Tianyu Xie, Shiyue Zhang, Cheng ZhangNeurIPS 2025 · 被引用 5 次
- Multi-Subspace Multi-Modal Modeling for Diffusion Models: Estimation, Convergence and Mixture of ExpertsRuofeng Yang, Yongcan Li, Bo Jiang, Cheng Chen 等ICLR 2026 · 被引用 4 次
它引用的顶会 Paper23
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual InversionRinon Gal, Yuval Alaluf, Yuval Atzmon, Or Patashnik 等ICLR 2023 · 被引用 464 次
相关 Paper
- HuTuMotion: Human-Tuned Navigation of Latent Motion Diffusion Models with Minimal FeedbackGaoge Han, Shaoli Huang, Mingming Gong, Jinglei TangAAAI 2024 · 被引用 4 次
- DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric FinetuningYuxuan Duan, Yan Hong, Bo Zhang, Jun Lan 等NeurIPS 2024 · 被引用 2 次
- Specialist Diffusion: Plug-and-Play Sample-Efficient Fine-Tuning of Text-to-Image Diffusion Models to Learn Any Unseen StyleHaoming Lu, Hazarapet Tunanyan, Kai Wang, Shant Navasardyan 等CVPR 2023
- Score Approximation, Estimation and Distribution Recovery of Diffusion Models on Low-Dimensional DataMinshuo Chen, Kaixuan Huang, Tuo Zhao, Mengdi WangICML 2023 · 被引用 168 次
- On the Generalization Properties of Diffusion ModelsPuheng Li, Zhong Li, Huishuai Zhang, Jiang BianNeurIPS 2023 · 被引用 86 次
