Magic Clothing: Controllable Garment-Driven Image Synthesis
Weifeng Chen, Tao Gu, Yuhao Xu, Arlene Chen
摘要
We propose Magic Clothing, a latent diffusion model (LDM)-based network architecture for an unexplored garment-driven image synthesis task. Aiming at generating customized characters wearing the target garments with diverse text prompts, the image controllability is the most critical issue, i.e., to preserve the garment details and maintain faithfulness to the text prompts. To this end, we introduce a garment extractor to capture the detailed garment features, and employ self-attention fusion to incorporate them into the pretrained LDMs, ensuring that the garment details remain unchanged on the target character. Then, we leverage the joint classifier-free guidance to balance the control of garment features and text prompts over the generated results. Meanwhile, the proposed garment extractor is a plug-in module applicable to various finetuned LDMs, and it can be combined with other extensions like ControlNet and IP-Adapter to enhance the diversity and controllability of the generated characters. Furthermore, we design Matched-Points-LPIPS (MP-LPIPS), a robust metric for evaluating the consistency of the target image to the source garment. Extensive experiments demonstrate that our Magic Clothing achieves state-of-the-art results under various conditional controls for garment-driven image synthesis. Our source code is available at https://github.com/ShineChen1024/MagicClothing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- IMAGDressing-v1: Customizable Virtual DressingFei Shen, Xin Jiang, Xin He, Hu Ye 等AAAI 2025 · 被引用 128 次
- OmniTry: Virtual Try-On Anything without MasksYutong Feng, Linlin Zhang, Hengyuan Cao, Yiming Chen 等NeurIPS 2025 · 被引用 16 次
- Any2anytryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing TasksHailong Guo, Bohan Zeng, Yiren Song, Wentao Zhang 等ICCV 2025 · 被引用 13 次
- FashionComposer: Compositional Fashion Image GenerationSihui Ji, Yiyang Wang, Xi Chen, Xiaogang Xu 等SIGGRAPH 2025 · 被引用 3 次
- FashionTailor: Controllable Clothing Editing for Human Images with Appearance PreservingJie Hou, Jianghong Ma, Xiangyu Mu, Haijun Zhang 等AAAI 2025 · 被引用 1 次
它引用的顶会 Paper34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 被引用 6,549 次
相关 Paper
- AnyDressing: Customizable Multi-Garment Virtual Dressing via Latent Diffusion ModelsXinghui Li, Qichao Sun, Pengze Zhang, Fulong Ye 等CVPR 2025
- OOTDiffusion: Outfitting Fusion Based Latent Diffusion for Controllable Virtual Try-OnYuhao Xu, Tao Gu, Weifeng Chen, Arlene ChenAAAI 2025 · 被引用 177 次
- IMAGGarment+: Efficient Attribute-Wise Diffusion for Garment GenerationJian Yu, Fei Shen, Cong Wang, Yanpeng Sun 等AAAI 2026
- Multi-focal Conditioned Latent Diffusion for Person Image SynthesisJiaqi Liu, Jichao Zhang, Paolo Rota, Nicu SebeCVPR 2025
- MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware DiffusionDi Chang, Yichun Shi, Quankai Gao, Hongyi Xu 等ICML 2024 · 被引用 125 次
