HybridTrak: Adding Full-Body Tracking to VR Using an Off-the-Shelf Webcam
Jackie (Junrui) Yang, Tuochao Chen, Fang Qin, Monica S. Lam, James A. Landay
Abstract
Full-body tracking in virtual reality improves presence, allows interaction via body postures, and facilitates better social expression among users. However, full-body tracking systems today require a complex setup fixed to the environment (e.g., multiple lighthouses/cameras) and a laborious calibration process, which goes against the desire to make VR systems more portable and integrated. We present HybridTrak, which provides accurate, real-time full-body tracking by augmenting inside-out1 upper-body VR tracking systems with a single external off-the-shelf RGB web camera. HybridTrak uses a full-neural solution to convert and transform users’ 2D full-body poses from the webcam to 3D poses leveraging the inside-out upper-body tracking data. We showed HybridTrak is more accurate than RGB or depth-based tracking methods on the MPI-INF-3DHP dataset. We also tested HybridTrak in the popular VRChat app and showed that body postures presented by HybridTrak are more distinguishable and more natural than a solution using an RGBD camera.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b9bb5439-4932-4e8a-b8f1-25c3068eceecCited by top-tier papers5
- ReactGenie: A Development Framework for Complex Multimodal Interactions Using Large Language ModelsJackie (Junrui) Yang, Yingtian Shi, Yuhan Zhang, Karina Li et al.CHI 2024 · 17 citations
- WAVE: Anticipatory Movement Visualization for VR DancingMarkus Laattala, Roosa Piitulainen, Nadia M. Ady, Monica Tamariz et al.CHI 2024 · 15 citations
- AMMA: Adaptive Multimodal Assistants Through Automated State Tracking and User Model-Directed Guidance PlanningJackie (Junrui) Yang, Leping Qiu, Emmanuel Angel Corona-Moreno, Louisa Shi et al.IEEE VR 2024 · 11 citations
- Systematic Literature Review of Using Virtual Reality as a Social Platform in HCI CommunityXiaoying Wei, Xiaofu Jin, Ge Lin Kan, Yukang Yan et al.CSCW 2025 · 10 citations
- GFPose: Learning 3D Human Pose Prior with Gradient FieldsHai Ci, Mingdong Wu, Wentao Zhu, Xiaoxuan Ma et al.CVPR 2023
Builds on8
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Learnable Triangulation of Human PoseKarim Iskakov, Egor Burkov, Victor S. Lempitsky, Yury MalkovICCV 2019 · 419 citations
- XNect: real-time multi-person 3D motion capture with a single RGB cameraDushyant Mehta, Oleksandr Sotnychenko, Franziska Mueller, Weipeng Xu et al.SIGGRAPH 2020 · 267 citations
- Cross View Fusion for 3D Human Pose EstimationHaibo Qiu, Chunyu Wang, Jingdong Wang, Naiyan Wang et al.ICCV 2019 · 242 citations
- Occlusion-Aware Networks for 3D Human Pose Estimation in VideoYu Cheng, Bo Yang, Bo Wang, Wending Yan et al.ICCV 2019 · 223 citations
Related papers
- EgoPoseVR: Spatiotemporal Multi-Modal Reasoning for Egocentric Full-Body Pose in Virtual RealityHaojie Cheng, Shaun Jing Heng Ong, Shaoyu Cai, Aiden Tat Yang Koh et al.IEEE VR 2026 · 1 citation
- BodyTrak: Inferring Full-body Poses from Body Silhouettes Using a Miniature Camera on a WristbandHyunchul Lim, Yaxuan Li, Matthew Dressa, Fang Hu et al.UbiComp 2022 · 18 citations
- FRAME: Floor-aligned Representation for Avatar Motion from Egocentric VideoAndrea Boscolo Camiletto, Jian Wang, Eduardo Alvarado, Rishabh Dabral et al.CVPR 2025
- MonoEye: Multimodal Human Motion Capture System Using A Single Ultra-Wide Fisheye CameraDong-Hyun Hwang, Kohei Aso, Ye Yuan, Kris M. Kitani et al.UIST 2020 · 37 citations
- EnvPoser: Environment-aware Realistic Human Motion Estimation from Sparse Observations with Uncertainty ModelingSongpengcheng Xia, Yu Zhang, Zhuo Su, Xiaozheng Zheng et al.CVPR 2025
