Hybrid Dynamic-static Context-aware Attention Network for Action Assessment in Long Videos
Ling-An Zeng, Fa-Ting Hong, Wei-Shi Zheng, Qi-Zhi Yu, Wei Zeng, Yaowei Wang, Jian-Huang Lai
Abstract
The objective of action quality assessment is to score sports videos. However, most existing works focus only on video dynamic information (i.e., motion information) but ignore the specific postures that an athlete is performing in a video, which is important for action assessment in long videos. In this work, we present a novel hybrid dynAmic-static Context-aware attenTION NETwork (ACTION-NET) for action assessment in long videos. To learn more discriminative representations for videos, we not only learn the video dynamic information but also focus on the static postures of the detected athletes in specific frames, which represent the action quality at certain moments, along with the help of the proposed hybrid dynamic-static architecture. Moreover, we leverage a context-aware attention module consisting of a temporal instancewise graph convolutional network unit and an attention unit for both streams to extract more robust stream features, where the former is for exploring the relations between instances and the latter for assigning a proper weight to each instance. Finally, we combine the features of the two streams to regress the final video score, supervised by ground-truth scores given by experts. Additionally, we have collected and annotated the new Rhythmic Gymnastics dataset, which contains videos of four different types of gymnastics routines, for evaluation of action quality assessment in long videos. Extensive experimental results validate the efficacy of our proposed method, which outperforms related approaches. The codes and dataset are available at https://github.com/lingan1996/ACTION-NET.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 652bc197-5e72-4be1-ab44-e913330721dbCited by top-tier papers12
- Cross-modal Consensus Network for Weakly Supervised Temporal Action LocalizationFa-Ting Hong, Jia-Chang Feng, Dan Xu, Ying Shan et al.ACM MM 2021 · 104 citations
- Likert Scoring with Grade Decoupling for Long-term Action AssessmentAngchi Xu, Ling-An Zeng, Wei-Shi ZhengCVPR 2022 · 41 citations
- Skating-Mixer: Long-Term Sport Audio-Visual Modeling with MLPsJingfei Xia, Mingchen Zhuge, Tiantian Geng, Shun Fan et al.AAAI 2023 · 38 citations
- Human-Activity AGV Quality Assessment: A Benchmark Dataset and an Objective Evaluation MetricZhichao Zhang, Wei Sun, Xinyue Li, Yunhao Li et al.ACM MM 2025 · 7 citations
- BriMA: Bridged Modality Adaptation for Multi-Modal Continual Action Quality AssessmentKanglei Zhou, Chang Li, Qingyi Pan, Liyuan WangCVPR 2026 · 3 citations
Builds on1
Related papers
- FineDiving: A Fine-grained Dataset for Procedure-aware Action Quality AssessmentJinglin Xu, Yongming Rao, Xumin Yu, Guangyi Chen et al.CVPR 2022 · 118 citations
- TSA-Net: Tube Self-Attention Network for Action Quality AssessmentShunli Wang, Dingkang Yang, Peng Zhai, Chixiao Chen et al.ACM MM 2021 · 93 citations
- Uncertainty-Aware Score Distribution Learning for Action Quality AssessmentYansong Tang, Zanlin Ni, Jiahuan Zhou, Danyang Zhang et al.CVPR 2020
- Accurate Temporal Action Proposal Generation with Relation-Aware Pyramid NetworkJialin Gao, Zhixiang Shi, Guanshuo Wang, Jiani Li et al.AAAI 2020 · 78 citations
- Localization-assisted Uncertainty Score Disentanglement Network for Action Quality AssessmentYanli Ji, Lingfeng Ye, Huili Huang, Lijing Mao et al.ACM MM 2023 · 25 citations
