PureCC: Pure Learning for Text-to-Image Concept Customization
Zhichao Liao, Xiaole Xian, Qingyu Li, Wenyu Qin, Meng Wang, Weicheng Xie, Siyang Song, Pingfa Feng, Long ZENG, Liang Pan
摘要
Existing concept customization methods have achieved remarkable outcomes in high-fidelity and multi-concept customization.However, they often neglect the influence on the original model's behavior and capabilities when learning new personalized concepts.To address this issue, we propose PureCC. PureCC novelly introduces a decoupled learning objective for concept customization, which combines the implicit guidance of the target concept with the original conditional prediction. This separated form enables PureCC to substantially focus on the original model during training. Moreover, based on this objective, PureCC designs a dual-branch training pipeline that includes a frozen extractor providing purified target concept representation as implicit guidance and a trainable flow model producing the original conditional prediction, jointly achieving pure learning for personalized concept. Furthermore, PureCC introduces an novel adaptive guidance scale to dynamically adjust the guidance strength of the target concept, balancing between customization fidelity and model preservation. Extensive experiments show that PureCC achieves state-of-the-art performance in preserving the original behavior and capabilities while enabling high-fidelity concept customization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper36
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- DCoAR: Deep Concept Injection into Unified Autoregressive Models for Personalized Text-to-Image GenerationFangtai Wu, Mushui Liu, Weijie He, Zhao Wang 等CVPR 2026 · 被引用 1 次
- Concept Conductor: Orchestrating Multiple Personalized Concepts in Text-to-Image SynthesisZebin Yao, Fangxiang Feng, Ruifan Li, Xiaojie WangAAAI 2025 · 被引用 3 次
- ClassDiffusion: More Aligned Personalization Tuning with Explicit Class GuidanceJiannan Huang, Jun Hao Liew, Hanshu Yan, Yuyang Yin 等ICLR 2025 · 被引用 1 次
- Orthogonal Adaptation for Modular Customization of Diffusion ModelsRyan Po, Guandao Yang, Kfir Aberman, Gordon WetzsteinCVPR 2024 · 被引用 18 次
- Personalized Residuals for Concept-Driven Text-to-Image GenerationCusuh Ham, Matthew Fisher, James Hays, Nicholas I. Kolkin 等CVPR 2024
