SwiftTailor: Efficient 3D Garment Generation with Geometry Image Representation
Phuc Pham, Uy Dieu Tran, Binh-Son Hua, Phong Nguyen
Abstract
Realistic and efficient 3D garment generation remains a longstanding challenge in computer vision and digital fashion. Existing methods typically rely on large vision- language models to produce serialized representations of 2D sewing patterns, which are then transformed into simulation-ready 3D meshes using garment modeling framework such as GarmentCode. Although these approaches yield high-quality results, they often suffer from slow inference times, ranging from 30 seconds to a minute. In this work, we introduce SwiftTailor, a novel two-stage framework that unifies sewing-pattern reasoning and geometry-based mesh synthesis through a compact geometry image representation. SwiftTailor comprises two lightweight modules: PatternMaker, an efficient vision-language model that predicts sewing patterns from diverse input modalities, and GarmentSewer, an efficient dense prediction transformer that converts these patterns into a novel Garment Geometry Image, encoding the 3D surface of all garment panels in a unified UV space. The final 3D mesh is reconstructed through an efficient inverse mapping process that incorporates remeshing and dynamic stitching algorithms to directly assemble the garment, thereby amortizing the cost of physical simulation. Extensive experiments on the Multimodal GarmentCodeData demonstrate that SwiftTailor achieves state-of-the-art accuracy and visual fidelity while significantly reducing inference time. This work offers a scalable, interpretable, and high-performance solution for next-generation 3D garment generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0adea0e7-862c-4af2-afcf-1a3f63476dfcBuilds on19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Palette: Image-to-Image Diffusion ModelsChitwan Saharia, William Chan, Huiwen Chang, Chris A. Lee et al.SIGGRAPH 2022 · 1,638 citations
- Codimensional incremental potential contactMinchen Li, Danny M. Kaufman, Chenfanfu JiangSIGGRAPH 2021 · 117 citations
- NeuralTailor: reconstructing sewing pattern structures from 3D point clouds of garmentsMaria Korosteleva, Sung-Hee LeeSIGGRAPH 2022 · 45 citations
Related papers
- PatternGSL: A Structured Specification Language for Template-Free and Simulation-Ready 3D GarmentsZhenyang Li, Lutao Jiang, Yizhou Zhao, Ying-Cong Chen et al.SIGGRAPH 2026
- GarmentGPT: Compositional Garment Pattern Generation via Discrete Latent TokenizationFangsheng Weng, Junhao Chen, Xiang Li, Jie Qin et al.ICLR 2026
- Learning Sewing Patterns via Latent Flow Matching of Implicit FieldsCong Cao, Ren Li, Corentin Dumery, Hao LiSIGGRAPH 2026
- What You See Is What You Wear: Crafting Garments for Diverse Avatars with Consistent Wearing EffectsZan Wang, Anqi Li, Yixuan Li, Wei Liang et al.IEEE VR 2026
- ReWeaver: Towards Simulation-Ready and Topology-Accurate Garment ReconstructionMing Li, Hui Shan, Kai Zheng, Chentao Shen et al.CVPR 2026 · 2 citations
