FineParser: A Fine-Grained Spatio-Temporal Action Parser for Human-Centric Action Quality Assessment
Jinglin Xu, Sibo Yin, Guohao Zhao, Zishuo Wang, Yuxin Peng
摘要
Existing action quality assessment (AQA) methods mainly learn deep representations at the video level for scoring diverse actions. Due to the lack of a fine-grained understanding of actions in videos, they harshly suffer from low credibility and interpretability, thus insufficient for stringent applications, such as Olympic diving events. We argue that a fine-grained understanding of actions requires the model to perceive and parse actions in both time and space, which is also the key to the credibility and inter-pretability of the AQA technique. Based on this insight, we propose a new fine-grained spatial-temporal action parser named FineParser. It learns human-centric foreground action representations by focusing on target action regions within each frame and exploiting their fine-grained alignments in time and space to minimize the impact of in-valid backgrounds during the assessment. In addition, we construct fine-grained annotations of human-centric fore-ground action masks for the FineDiving dataset, called FineDiving-HM. With refined annotations on diverse target action procedures, FineDiving-HM can promote the development of real-world AQA systems. Through extensive experiments, we demonstrate the effectiveness of FineParser, which outperforms state-of-the-art methods while supporting more tasks of fine-grained action understanding. Data and code are available at https://github.com/PKU-ICST-MIPL/FineParser_CVPR2024.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- HumanSAM: Classifying Human-Centric Forgery Videos in Human Spatial, Appearance, and Motion AnomalyChang Liu, Yunfan Ye, Fan Zhang, Qingyang Zhou 等ICCV 2025 · 被引用 6 次
- SkillSight: Efficient First-Person Skill Assessment with GazeChi Hsuan Wu, Kumar Ashutosh, Kristen GraumanCVPR 2026 · 被引用 4 次
- BriMA: Bridged Modality Adaptation for Multi-Modal Continual Action Quality AssessmentKanglei Zhou, Chang Li, Qingyi Pan, Liyuan WangCVPR 2026 · 被引用 3 次
- MCMoE: Completing Missing Modalities with Mixture of Experts for Incomplete Multimodal Action Quality AssessmentHuangbiao Xu, Huanqi Wu, Xiao Ke, Junyi Wu 等AAAI 2026 · 被引用 2 次
- Learning Long-Range Action Representation by Two-Stream Mamba Pyramid Network for Figure Skating AssessmentFengshun Wang, Qiurui Wang, Peilin ZhaoACM MM 2025 · 被引用 1 次
它引用的顶会 Paper21
- SportsMOT: A Large Multi-Object Tracking Dataset in Multiple Sports ScenesYutao Cui, Chenkai Zeng, Xiaoyu Zhao, Yichun Yang 等ICCV 2023 · 被引用 187 次
- Group-aware Contrastive Regression for Action Quality AssessmentXumin Yu, Yongming Rao, Wenliang Zhao, Jiwen Lu 等ICCV 2021 · 被引用 147 次
- Action Assessment by Joint Relation GraphsJiahui Pan, Jibin Gao, Wei-Shi ZhengICCV 2019 · 被引用 141 次
- MultiSports: A Multi-Person Video Dataset of Spatio-Temporally Localized Sports ActionsYixuan Li, Lei Chen, Runyu He, Zhenzhi Wang 等ICCV 2021 · 被引用 131 次
- FineDiving: A Fine-grained Dataset for Procedure-aware Action Quality AssessmentJinglin Xu, Yongming Rao, Xumin Yu, Guangyi Chen 等CVPR 2022 · 被引用 118 次
相关 Paper
- TSA-Net: Tube Self-Attention Network for Action Quality AssessmentShunli Wang, Dingkang Yang, Peng Zhai, Chixiao Chen 等ACM MM 2021 · 被引用 93 次
- Localization-assisted Uncertainty Score Disentanglement Network for Action Quality AssessmentYanli Ji, Lingfeng Ye, Huili Huang, Lijing Mao 等ACM MM 2023 · 被引用 25 次
- Uncertainty-Aware Score Distribution Learning for Action Quality AssessmentYansong Tang, Zanlin Ni, Jiahuan Zhou, Danyang Zhang 等CVPR 2020
- Intra- and Inter-Action Understanding via Temporal Action ParsingDian Shao, Yue Zhao, Bo Dai, Dahua LinCVPR 2020
- HAA500: Human-Centric Atomic Action Dataset with Curated VideosJihoon Chung, Cheng-hsin Wuu, Hsuan-ru Yang, Yu-Wing Tai 等ICCV 2021 · 被引用 62 次
