High-Fidelity 4D Cloth Capture Pipeline with a Two-Level Pattern
Ziheng Liu, Anka He Chen, Shu Chen, Yin Yang, Cem Yuksel, Jenny Han Lin
Abstract
Capturing cloth motion with high fidelity is challenging due to fine-scale wrinkles, large deformation, and frequent self-occlusion. We present a 4D (spatio-temporal) cloth capture system that achieves 1 mm spatial resolution using only 16 RGB cameras. Our approach uses a two-level marker pattern: sparse, colored L-shaped markers provide robust detection and orientation, while dense noise patterns within each marker enable both marker identification and precise keypoint localization. By unwarping detected markers to a canonical frame, we factor out perspective distortion and most of cloth deformation, allowing the localizer to achieve sub-pixel accuracy. The localized keypoints are triangulated across views to form an incomplete point cloud. A physics-based optimization then deforms a template mesh to match the captured geometry while maintaining penetration-free constraints and physical plausibility for occluded regions. Our method produces temporally coherent sequences that faithfully capture fine wrinkles and folds even during complex motions with self-contact.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get a661de9a-8e95-4e71-8e66-d948a0702510Related papers
- Capturing detailed deformations of moving human bodiesHe Chen, Hyojoon Park, Kutay Macit, Ladislav KavanSIGGRAPH 2021 · 30 citations
- GPU-based simulation of cloth wrinkles at submillimeter levelsHuamin WangSIGGRAPH 2021 · 115 citations
- Holoported Characters: Real-Time Free-Viewpoint Rendering of Humans from Sparse RGB CamerasAshwath Shetty, Marc Habermann, Guoxing Sun, Diogo C. Luvizon et al.CVPR 2024 · 9 citations
- SAFT: Shape and Appearance of Fabrics from Template via Differentiable Physical Simulations from Monocular VideoDavid Stotko, Reinhard KleinICCV 2025 · 1 citation
- ChallenCap: Monocular 3D Capture of Challenging Human Performances Using Multi-Modal ReferencesYannan He, Anqi Pang, Xin Chen, Han Liang et al.CVPR 2021
