Text-Driven Fashion Image Editing with Compositional Concept Learning and Counterfactual Abduction
Shanshan Huang, Haoxuan Li, Chunyuan Zheng, Mingyuan Ge, Wei Gao, Lei Wang, Li Liu
Abstract
Fashion image editing is a valuable tool for designers to convey their creative ideas by visualizing design concepts. With the recent advances in text editing methods, significant progress has been made in fashion image editing. However, they face two key challenges: spurious correlations in training data often induce changes in other areas when editing an area representing the intended editing concept, and these models typically lack the ability to edit multiple concepts simultaneously. To address the above challenges, we propose a novel Text-driven Fashion Image ediTing framework called T-FIT to mitigate the impact of spurious correlation by integrating counterfactual reasoning with compositional concept learning to precisely ensure compositional multiconcept fashion image editing relying solely on text descriptions. Specifically, T-FIT includes three key components. (i) Counterfactual abduction module, which learns an exogenous variable of the source image by a denoising U-Net model. (ii) Concept learning module, which identifies concepts in fashion image editing-such as clothing types and colors and projects a target concept into the space spanned from a series of textual prompts. (iii) Concept composition module, which enables simultaneous adjustments of multiple concepts by aggregating each concept's direction vector obtained from the concept learning module. Extensive experiments show that our method can achieve state-of-theart performance on various fashion image editing tasks, including single-concept editing (e.g., sleeve length, clothing type) and multi-concept editing (e.g., color & sleeve length).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7331e907-02b3-41c4-a14e-d9f0c463f002Cited by top-tier papers3
- Unveiling Extraneous Sampling Bias with Data Missing-Not-At-RandomChunyuan Zheng, Haocheng Yang, Haoxuan Li, Mengyue YangNeurIPS 2025 · 15 citations
- Addressing Correlated Latent Exogenous Variables in Debiased Recommender SystemsShuqiang Zhang, Yuchao Zhang, Jinkun Chen, Haochen SuiKDD 2025 · 4 citations
- Mitigating Data Imbalance in Time Series Classification Based on Counterfactual Minority Samples AugmentationLei Wang, Shanshan Huang, Chunyuan Zheng, Jun Liao et al.KDD 2025 · 1 citation
Builds on22
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- MasaCtrl: Tuning-Free Mutual Self-Attention Control for Consistent Image Synthesis and EditingMingdeng Cao, Xintao Wang, Zhongang Qi, Ying Shan et al.ICCV 2023 · 770 citations
- Diffusion Autoencoders: Toward a Meaningful and Decodable RepresentationKonpat Preechakul, Nattanat Chatthee, Suttisak Wizadwongsa, Supasorn SuwajanakornCVPR 2022 · 276 citations
Related papers
- TexFit: Text-Driven Fashion Image Editing with Diffusion ModelsTongxin Wang, Mang YeAAAI 2024 · 20 citations
- Edit Like A Designer: Modeling Design Workflows for Unaligned Fashion EditingQiyu Dai, Shuai Yang, Wenjing Wang, Wei Xiang et al.ACM MM 2021 · 5 citations
- FashionTailor: Controllable Clothing Editing for Human Images with Appearance PreservingJie Hou, Jianghong Ma, Xiangyu Mu, Haijun Zhang et al.AAAI 2025 · 1 citation
- Doubly Abductive Counterfactual Inference for Text-Based Image EditingXue Song, Jiequan Cui, Hanwang Zhang, Jingjing Chen et al.CVPR 2024 · 7 citations
- FEAT: Fashion Editing and Try-On from Any DesignSoye Kwon, Keonyoung Lee, Dahuin Jung, Jaekoo LeeCVPR 2026
