Full-Range Virtual Try-On with Recurrent Tri-Level Transform
Han Yang, Xinrui Yu, Ziwei Liu
Abstract
Virtual try-on aims to transfer a target clothing image onto a reference person. Though great progress has been achieved, the functioning zone of existing works is still limited to standard clothes (e.g., plain shirt without complex laces or ripped effect), while the vast complexity and variety of non-standard clothes (e.g., off-shoulder shirt, word-shoulder dress) are largely ignored. In this work, we propose a principled framework, Re-current Tri-Level Transform (RT-VTON), that performs full-range virtual try-on on both standard and non-standard clothes. We have two key insights towards the framework design: 1) Semantics transfer requires a gradual feature transform on three different levels of clothing representations, namely clothes code, pose code and parsing code. 2) Geometry transfer requires a regularized image deformation between rigidity and flexibility. Firstly, we predict the semantics of the “after-try-on” person by recurrently refining the tri-level feature codes using local gated attention and non-local correspondence learning. Next, we design a semi-rigid deformation to align the clothing image and the predicted semantics, which preserves local warping similarity. Finally, a canonical try-on synthesizer fuses all the processed information to generate the clothed person image. Extensive experiments on conventional benchmarks along with user studies demonstrate that our framework achieves state-of-the-art performance both quantitatively and qualitatively. Notably, RT-VTON shows compelling results on a wide range of non-standard clothes. Project page: https://lzqhardworker.github.io/RT-VTON/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d658489-d7e4-4570-8f07-e52cc65b22d5Cited by top-tier papers16
- Size Does Matter: Size-aware Virtual Try-on via Clothing-oriented Transformation Try-on NetworkChieh-Yun Chen, Yi-Chung Chen, Hong-Han Shuai, Wen-Huang ChengICCV 2023 · 38 citations
- Texture-Preserving Diffusion Models for High-Fidelity Virtual Try-OnXu Yang, Changxing Ding, Zhibin Hong, Junhao Huang et al.CVPR 2024 · 25 citations
- M&M VTO: Multi-Garment Virtual Try-On and EditingLuyang Zhu, Yingwei Li, Nan Liu, Hao Peng et al.CVPR 2024 · 13 citations
- Greatness in Simplicity: Unified Self-Cycle Consistency for Parser-Free Virtual Try-OnChenghu Du, Junyin Wang, Shuqing Liu, Shengwu XiongNeurIPS 2023 · 11 citations
- FaD-VLP: Fashion Vision-and-Language Pre-training towards Unified Retrieval and CaptioningSuvir Mirchandani, Licheng Yu, Mengjiao Wang, Animesh Sinha et al.EMNLP 2022 · 9 citations
Builds on11
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- ClothFlow: A Flow-Based Model for Clothed Person GenerationXintong Han, Weilin Huang, Xiaojun Hu, Matthew R. ScottICCV 2019 · 297 citations
- FiNet: Compatible and Diverse Fashion Image InpaintingXintong Han, Zuxuan Wu, Weilin Huang, Matthew R. Scott et al.ICCV 2019 · 85 citations
- WAS-VTON: Warping Architecture Search for Virtual Try-on NetworkZhenyu Xie, Xujie Zhang, Fuwei Zhao, Haoye Dong et al.ACM MM 2021 · 24 citations
- Toward Realistic Virtual Try-on Through Landmark Guided Shape MatchingGuoqiang Liu, Dan Song, Ruofeng Tong, Min TangAAAI 2021 · 19 citations
Related papers
- Towards Multi-Pose Guided Virtual Try-On NetworkHaoye Dong, Xiaodan Liang, Xiaohui Shen, Bochao Wang et al.ICCV 2019 · 226 citations
- GP-VTON: Towards General Purpose Virtual Try-On via Collaborative Local-Flow Global-Parsing LearningZhenyu Xie, Zaiyu Huang, Xin Dong, Fuwei Zhao et al.CVPR 2023
- Disentangled Cycle Consistency for Highly-Realistic Virtual Try-OnChongjian Ge, Yibing Song, Yuying Ge, Han Yang et al.CVPR 2021
- M3D-VTON: A Monocular-to-3D Virtual Try-On NetworkFuwei Zhao, Zhenyu Xie, Michael Kampffmeyer, Haoye Dong et al.ICCV 2021 · 81 citations
- MV-TON: Memory-based Video Virtual Try-on networkXiaojing Zhong, Zhonghua Wu, Taizhe Tan, Guosheng Lin et al.ACM MM 2021 · 27 citations
