JRDB-Pose: A Large-Scale Dataset for Multi-Person Pose Estimation and Tracking
Edward Vendrow, Duy-Tho Le, Jianfei Cai, Hamid Rezatofighi
Abstract
Autonomous robotic systems operating in human environments must understand their surroundings to make accurate and safe decisions. In crowded human scenes with close-up human-robot interaction and robot navigation, a deep understanding of surrounding people requires reasoning about human motion and body dynamics over time with human body pose estimation and tracking. However, existing datasets captured from robot platforms either do not provide pose annotations or do not reflect the scene distribution of social robots. In this paper, we introduce JRDB-Pose, a large-scale dataset and benchmark for multi-person pose estimation and tracking. JRDB-Pose extends the existing JRDB which includes videos captured from a social navigation robot in a university campus environment, containing challenging scenes with crowded indoor and outdoor locations and a diverse range of scales and occlusion types. JRDB-Pose provides human pose annotations with per-keypoint occlusion labels and track IDs consistent across the scene and with existing annotations in JRDB. We conduct a thorough experimental study of state-of-theart multi-person pose estimation and tracking methods on JRDB-Pose, showing that our dataset imposes new challenges for the existing methods. JRDB-Pose is available at https://jrdb.erc.monash.edu/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a0ba3bf9-3550-4d84-989c-1a8a1a4bf09fCited by top-tier papers12
- Neural Localizer Fields for Continuous 3D Human Pose and Shape EstimationIstván Sárándi, Gerard Pons-MollNeurIPS 2024 · 76 citations
- PolySLGen: Online Multimodal Speaking-Listening Reaction Generation in Polyadic InteractionZhi-Yi Lin, Thomas Markhorst, Jouh Yeong Chew, Xucong ZhangCVPR 2026 · 3 citations
- JRDB-Reasoning: A Difficulty-Graded Benchmark for Visual Reasoning in RoboticsSimindokht Jahangard, Mehrzad Mohammadi, Yi Shen, Zhixi Cai et al.AAAI 2026 · 2 citations
- JRDB-Social: A Multifaceted Robotic Dataset for Understanding of Context and Dynamics of Human Interactions Within Social GroupsSimindokht Jahangard, Zhixi Cai, Shiki Wen, Hamid RezatofighiCVPR 2024
- Bézier Degradation Modeling for LiDAR-based Human Motion CaptureXiaoqi An, Lin Zhao, Jun Li, Chen Gong et al.CVPR 2026
Builds on16
- Tracking Without Bells and WhistlesPhilipp Bergmann, Tim Meinhardt, Laura Leal-TaixéICCV 2019 · 1,030 citations
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- Occlusion-Aware Networks for 3D Human Pose Estimation in VideoYu Cheng, Bo Yang, Bo Wang, Wending Yan et al.ICCV 2019 · 223 citations
- 3D Human Pose Estimation Using Spatio-Temporal Networks with Explicit Occlusion TrainingYu Cheng, Bo Yang, Bo Wang, Robby T. TanAAAI 2020 · 145 citations
- MOTSynth: How Can Synthetic Data Help Pedestrian Detection and Tracking?Matteo Fabbri, Guillem Brasó, Gianluca Maugeri, Orcun Cetintas et al.ICCV 2021 · 128 citations
Related papers
- JRDB-PanoTrack: An Open-World Panoptic Segmentation and Tracking Robotic Dataset in Crowded Human EnvironmentsDuy-Tho Le, Chenhui Gou, Stavya Datta, Hengcan Shi et al.CVPR 2024
- JRDB-Act: A Large-scale Dataset for Spatio-temporal Action, Social Group and Activity DetectionMahsa Ehsanpour, Fatemeh Sadat Saleh, Silvio Savarese, Ian D. Reid et al.CVPR 2022 · 49 citations
- PoseTrack21: A Dataset for Person Search, Multi-Object Tracking and Multi-Person Pose TrackingAndreas Doering, Di Chen, Shanshan Zhang, Bernt Schiele et al.CVPR 2022 · 47 citations
- CDTB: A Color and Depth Visual Object Tracking Dataset and BenchmarkAlan Lukezic, Ugur Kart, Jani Käpylä, Ahmed Durmush et al.ICCV 2019 · 79 citations
- Reconstructing Close Human Interaction with Appearance and Proxemics ReasoningBuzhen Huang, Chen Li, Chongyang Xu, Dongyue Lu et al.CVPR 2025
