RacketVision: A Multiple Racket Sports Benchmark for Unified Ball and Racket Analysis
Linfeng Dong, Yuchen Yang, Hao Wu, Wei Wang, Yuenan Hou, Zhihang Zhong, Xiao Sun
摘要
We introduce RacketVision, a novel dataset and benchmark for advancing computer vision in sports analytics, covering table tennis, tennis, and badminton. The dataset is the first to provide large-scale, fine-grained annotations for racket pose alongside traditional ball positions, enabling research into complex human-object interactions. It is designed to tackle three interconnected tasks: fine-grained ball tracking, articulated racket pose estimation, and predictive ball trajectory forecasting. Our evaluation of established baselines reveals a critical insight for multi-modal fusion: while naively concatenating racket pose features degrades performance, a Cross-Attention mechanism is essential to unlock their value, leading to trajectory prediction results that surpass strong unimodal baselines. RacketVision provides a versatile resource and a strong starting point for future research in dynamic object tracking, conditional motion forecasting, and multi-modal analysis in sports.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- ViTPose: Simple Vision Transformer Baselines for Human Pose EstimationYufei Xu, Jing Zhang, Qiming Zhang, Dacheng TaoNeurIPS 2022 · 被引用 1,105 次
- Shape from Blur: Recovering Textured 3D Shape and Motion of Fast Moving ObjectsDenys Rozumnyi, Martin R. Oswald, Vittorio Ferrari, Marc PollefeysNeurIPS 2021 · 被引用 15 次
- Towards Explicit Exoskeleton for the Reconstruction of Complicated 3D Human AvatarsYifan Zhan, Qingtian Zhu, Muyao Niu, Mingze Ma 等ICCV 2025 · 被引用 3 次
- Sequential Gaussian Avatars with Hierarchical Motion ContextWangze Xu, Yifan Zhan, Zhihang Zhong, Xiao SunICCV 2025 · 被引用 1 次
- SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language ModelsHaotian Xia, Zhengbang Yang, Junbo Zou, Rhys Tracy 等ICLR 2025
相关 Paper
- EventAnchor: Reducing Human Interactions in Event Annotation of Racket Sports VideosDazhen Deng, Jiang Wu, Jiachen Wang, Yihong Wu 等CHI 2021 · 被引用 29 次
- ShuttleNet: Position-Aware Fusion of Rally Progress and Player Styles for Stroke Forecasting in BadmintonWei-Yao Wang, Hong-Han Shuai, Kai-Shiang Chang, Wen-Chih PengAAAI 2022 · 被引用 56 次
- ViSTec: Video Modeling for Sports Technique Recognition and Tactical AnalysisYuchen He, Zeqing Yuan, Yihong Wu, Liqi Cheng 等AAAI 2024 · 被引用 13 次
- SportsMOT: A Large Multi-Object Tracking Dataset in Multiple Sports ScenesYutao Cui, Chenkai Zeng, Xiaoyu Zhao, Yichun Yang 等ICCV 2023 · 被引用 187 次
- Augmenting Sports Videos with VisCommentatorChen Zhu-Tian, Shuainan Ye, Xiangtong Chu, Haijun Xia 等IEEE VIS 2021 · 被引用 60 次
