Down to the Last Detail: Virtual Try-on with Fine-grained Details
Jiahang Wang, Tong Sha, Wei Zhang, Zhoujun Li, Tao Mei
Abstract
Virtual try-on has attracted lots of research attention due to its potential applications in e-commerce, virtual reality and fashion design. However, existing methods can hardly preserve the fine-grained details (e.g., clothing texture, facial identity, hair style, skin tone) during generation, due to the non-rigid body deformation and multi-scale details. In this work, we propose a multi-stage framework to synthesize person images, where fine-grained details can be well preserved. To address the long-range translation and rich-details generation, we propose a Tree-Block (tree dilated fusion block) to replace standard ResNet-block where applicable. Notably, multi-scale feature maps can be smoothly fused for fine-grained detail generation, by incorporating larger spatial context at multiple scales. With a delicate end-to-end training scheme, our whole framework can be jointly optimized for results with significantly better visual fidelity and richer details. Moreover, we also explore the potential application in video-based virtual try-on. By harnessing the well-trained image generator and an extra video-level adaptor, a model photo can be well animated with a driving pose sequence. Extensive evaluations on standard datasets and user study demonstrate that our proposed framework achieves the state-of-the-art results, especially in preserving visual details in clothing texture and facial identity. Our implementation is publicly available via https://github.com/JDAI-CV/Down-to-the-Last-Detail-Virtual-Try-on-with-Detail-Carving.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 5d974c95-e191-4ad1-887c-d5bcc6f938c0Cited by top-tier papers6
- Style-Based Global Appearance Flow for Virtual Try-OnSen He, Yi-Zhe Song, Tao XiangCVPR 2022 · 112 citations
- MV-VTON: Multi-View Virtual Try-On with Diffusion ModelsHaoyu Wang, Zhilu Zhang, Donglin Di, Shiliang Zhang et al.AAAI 2025 · 32 citations
- Dressing in the Wild by Watching Dance VideosXin Dong, Fuwei Zhao, Zhenyu Xie, Xijin Zhang et al.CVPR 2022 · 30 citations
- Structure-transformed Texture-enhanced Network for Person Image SynthesisMunan Xu, Yuanqi Chen, Shan Liu, Thomas H. Li et al.ICCV 2021 · 3 citations
- PaRUS: A Virtual Reality Shopping Method Focusing on Contextual Information between Products and Real Usage ScenesYinyu Lu, Weitao You, Ziqing Zheng, Yizhan Shao et al.IEEE VR 2025 · 2 citations
Related papers
- ZFlow: Gated Appearance Flow-based Virtual Try-on with 3D PriorsAyush Chopra, Rishabh Jain, Mayur Hemani, Balaji KrishnamurthyICCV 2021 · 73 citations
- GPD-VVTO: Preserving Garment Details in Video Virtual Try-OnYuanbin Wang, Weilun Dai, Long Chan, Huanyu Zhou et al.ACM MM 2024 · 4 citations
- VTNFP: An Image-Based Virtual Try-On Network With Body and Clothing Feature PreservationRuiyun Yu, Xiaoqi Wang, Xiaohui XieICCV 2019 · 184 citations
- ClothFlow: A Flow-Based Model for Clothed Person GenerationXintong Han, Weilin Huang, Xiaojun Hu, Matthew R. ScottICCV 2019 · 297 citations
- Robust-MVTON: Learning Cross-Pose Feature Alignment and Fusion for Robust Multi-View Virtual Try-OnNannan Zhang, Yijiang Li, Dong Du, Zheng Chong et al.CVPR 2025
