EventHPE: Event-based 3D Human Pose and Shape Estimation
Shihao Zou, Chuan Guo, Xinxin Zuo, Sen Wang, Pengyu Wang, Xiaoqin Hu, Shoushun Chen, Minglun Gong, Li Cheng
Abstract
Event camera is an emerging imaging sensor for capturing dynamics of moving objects as events, which motivates our work in estimating 3D human pose and shape from the event signals. Events, on the other hand, have their unique challenges: rather than capturing static body postures, the event signals are best at capturing local motions. This leads us to propose a two-stage deep learning approach, called EventHPE. The first-stage, FlowNet, is trained by unsupervised learning to infer optical flow from events. Both events and optical flow are closely related to human body dynamics, which are fed as input to the ShapeNet in the second stage, to estimate 3D human shapes. To mitigate the discrepancy between image-based flow (optical flow) and shape-based flow (vertices movement of human body shape), a novel flow coherence loss is introduced by exploiting the fact that both flows are originated from the identical human motion. An in-house event-based 3D human dataset is curated that comes with 3D pose and shape annotations, which is by far the largest one to our knowledge. Empirical evaluations on DHP19 dataset and our in-house dataset demonstrate the effectiveness of our approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e86cdce6-4848-4114-9a41-7a71b8e36855Cited by top-tier papers14
- VIRD: Immersive Match Video Analysis for High-Performance Badminton CoachingTica Lin, Alexandre Aouididi, Chen Zhu-Tian, Johanna Beyer et al.IEEE VIS 2023 · 31 citations
- Lightweight Super-Resolution Head for Human Pose EstimationHaonan Wang, Jie Liu, Jie Tang, Gangshan WuACM MM 2023 · 23 citations
- EventEgo3D: 3D Human Motion Capture from Egocentric Event StreamsChristen Millerdurai, Hiroyasu Akada, Jian Wang, Diogo C. Luvizon et al.CVPR 2024 · 11 citations
- RELI11D: A Comprehensive Multimodal Human Motion Dataset and MethodMing Yan, Yan Zhang, Shuqiang Cai, Shuqi Fan et al.CVPR 2024 · 5 citations
- E-4DGS: High-Fidelity Dynamic Reconstruction from the Multi-view Event CamerasChaoran Feng, Zhenyu Tang, Wangbo Yu, Yatian Pang et al.ACM MM 2025 · 3 citations
Builds on5
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 427 citations
- DenseRaC: Joint 3D Pose and Shape Estimation by Dense Render-and-CompareYuanlu Xu, Song-Chun Zhu, Tony TungICCV 2019 · 204 citations
- VIBE: Video Inference for Human Body Pose and Shape EstimationMuhammed Kocabas, Nikos Athanasiou, Michael J. BlackCVPR 2020
- EventCap: Monocular 3D Capture of High-Speed Human Motions Using an Event CameraLan Xu, Weipeng Xu, Vladislav Golyanik, Marc Habermann et al.CVPR 2020
- Learning Event-Based Motion DeblurringZhe Jiang, Yu Zhang, Dongqing Zou, Jimmy S. J. Ren et al.CVPR 2020
Related papers
- Unsupervised 3d Motion Estimation Using Event CameraHan Han, Wei Zhai, Tiesong Zhao, Bin Li et al.CVPR 2026
- E2PNet: Event to Point Cloud Registration with Spatio-Temporal Representation LearningXiuhong Lin, Changjie Qiu, Zhipeng Cai, Siqi Shen et al.NeurIPS 2023 · 18 citations
- Efficient Meshflow and Optical Flow Estimation from Event CamerasXinglong Luo, Ao Luo, Zhengning Wang, Chunyu Lin et al.CVPR 2024 · 10 citations
- Adaptive Vision Transformer for Event-Based Human Pose EstimationNannan Yu, Tao Ma, Jiqing Zhang, Yuji Zhang et al.ACM MM 2024 · 9 citations
- Time Lens: Event-Based Video Frame InterpolationStepan Tulyakov, Daniel Gehrig, Stamatios Georgoulis, Julius Erbach et al.CVPR 2021
