Garments2Look: A Multi-Reference Dataset for High-Fidelity Outfit-Level Virtual Try-On with Clothing and Accessories
Junyao Hu, Zhongwei Cheng, Waikeung Wong, Xingxing Zou
摘要
Virtual try-on (VTON) has advanced single-garment visualization, yet real-world fashion centers on full outfits with multiple garments, accessories, fine-grained categories, layering, and diverse styling, remaining beyond current VTON systems. Existing datasets are category-limited and lack outfit diversity. We introduce Garments2Look, the first large-scale multimodal dataset for outfit-level VTON, comprising 80K many-garments-toone-look pairs across 40 major categories and 300+ finegrained subcategories. Each pair includes an outfit with 3-12 reference garment images (Average 4.48), a model image wearing the outfit, and detailed item and try-on textual annotations. To balance authenticity and diversity, we propose a synthesis pipeline. It involves heuristically constructing outfit lists before generating try-on results, with the entire process subjected to strict automated filtering and human validation to ensure data quality. To probe task difficulty, we adapt SOTA VTON methods and general-purpose image editing models to establish baselines. Results show current methods struggle to try on complete outfits seamlessly and to infer correct layering and styling, leading to misalignment and artifacts. Our code and data are open-sourced on https://github.com/ ArtmeScienceLab/Garments2Look.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- On Aliased Resizing and Surprising Subtleties in GAN EvaluationGaurav Parmar, Richard Zhang, Jun-Yan ZhuCVPR 2022 · 被引用 250 次
- Dressing in Order: Recurrent Person Image Generation for Pose Transfer, Virtual Try-on and Outfit EditingAiyu Cui, Daniel McKee, Svetlana LazebnikICCV 2021 · 被引用 110 次
- Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image EditingYusu Qian, Eli Bocek-Rivele, Liangchen Song, Jialing Tong 等CVPR 2026 · 被引用 63 次
- XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT ModulationBowen Chen, Brynn zhao, Haomiao Sun, Li Chen 等NeurIPS 2025 · 被引用 60 次
- Size Does Matter: Size-aware Virtual Try-on via Clothing-oriented Transformation Try-on NetworkChieh-Yun Chen, Yi-Chung Chen, Hong-Han Shuai, Wen-Huang ChengICCV 2023 · 被引用 38 次
相关 Paper
- IMAGDressing-v1: Customizable Virtual DressingFei Shen, Xin Jiang, Xin He, Hu Ye 等AAAI 2025 · 被引用 128 次
- MV-Fashion: Towards Enabling Virtual Try-On and Size Estimation with Multi-View Paired DataHunor Laczkó, Libang Jia, Loc-Phat Truong, Diego Hernández 等CVPR 2026
- OmniTry: Virtual Try-On Anything without MasksYutong Feng, Linlin Zhang, Hengyuan Cao, Yiming Chen 等NeurIPS 2025 · 被引用 16 次
- Any2anytryon: Leveraging Adaptive Position Embeddings for Versatile Virtual Clothing TasksHailong Guo, Bohan Zeng, Yiren Song, Wentao Zhang 等ICCV 2025 · 被引用 13 次
- MV-VTON: Multi-View Virtual Try-On with Diffusion ModelsHaoyu Wang, Zhilu Zhang, Donglin Di, Shiliang Zhang 等AAAI 2025 · 被引用 32 次
