Sparkle: A Robust and Versatile Representation for Point Cloud-based Human Motion Capture
Yiming Ren, Yujing Sun, Aoru Xue, Kwok-Yan Lam, Yuexin Ma
摘要
Point cloud-based motion capture leverages rich spatial geometry and privacy-preserving sensing, but learning robust representations from noisy, unstructured point clouds remains challenging. Existing approaches face a struggle trade-off between point-based methods (geometrically detailed but noisy) and skeleton-based ones (robust but oversimplified). We address the fundamental challenge: how to construct an effective representation for human motion capture that can balance expressiveness and robustness. In this paper, we propose Sparkle, a structured representation unifying skeletal joints and surface anchors with explicit kinematic-geometric factorization. Our framework, SparkleMotion, learns this representation through hierarchical modules embedding geometric continuity and kinematic constraints. By explicitly disentangling internal kinematic structure from external surface geometry, SparkleMotion achieves state-of-the-art performance not only in accuracy but crucially in robustness and generalization under severe domain shifts, noise, and occlusion. Extensive experiments demonstrate our superiority across diverse sensor types and challenging real-world scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 被引用 1,105 次
- SoftGroup for 3D Instance Segmentation on Point CloudsThang Vu, Kookhoi Kim, Tung Minh Luu, Thanh Xuan Nguyen 等CVPR 2022 · 被引用 251 次
- TransPose: real-time 3D human translation and pose estimation with six inertial sensorsXinyu Yi, Yuxiao Zhou, Feng XuSIGGRAPH 2021 · 被引用 200 次
- Physical Inertial Poser (PIP): Physics-aware Real-time Human Motion Tracking from Sparse Inertial SensorsXinyu Yi, Yuxiao Zhou, Marc Habermann, Soshi Shimada 等CVPR 2022 · 被引用 198 次
- WHAM: Reconstructing World-Grounded Humans with Accurate 3D MotionSoyong Shin, Juyong Kim, Eni Halilaj, Michael J. BlackCVPR 2024 · 被引用 66 次
相关 Paper
- SPEAL: Skeletal Prior Embedded Attention Learning for Cross-Source Point Cloud RegistrationKezheng Xiong, Maoji Zheng, Qingshan Xu, Chenglu Wen 等AAAI 2024 · 被引用 24 次
- Glimpse: Geometry Learning of Multi-scale Structural Priors for 3D Pose EstimationZhenhua TANG, Jihua Peng, Yanbin Hao, Qiguang Miao 等ICML 2026
- Geo-CF2Net: Geometry-Prior Cross-Frequency Interactive Fusion Network for 3D Human Action RecognitionZhaoyu Chen, Qian Huang, Xing Li, Yunfei Zhang 等ACM MM 2025
- SOMA: Solving Optical Marker-Based MoCap AutomaticallyNima Ghorbani, Michael J. BlackICCV 2021 · 被引用 48 次
- VoteHMR: Occlusion-Aware Voting Network for Robust 3D Human Mesh Recovery from Partial Point CloudsGuanze Liu, Yu Rong, Lu ShengACM MM 2021 · 被引用 27 次
