Uncertainty-Guided Probabilistic Transformer for Complex Action Recognition
Hongji Guo, Hanjing Wang, Qiang Ji
Abstract
A complex action consists of a sequence of atomic actions that interact with each other over a relatively long period of time. This paper introduces a probabilistic model named Uncertainty-Guided Probabilistic Transformer (UGPT) for complex action recognition. The self-attention mechanism of a Transformer is used to capture the complex and long-term dynamics of the complex actions. By explicitly modeling the distribution of the attention scores, we extend the deterministic Transformer to a probabilistic Transformer in order to quantify the uncertainty of the pre-diction. The model prediction uncertainty is used to improve both training and inference. Specifically, we propose a novel training strategy by introducing a majority model and a minority model based on the epistemic uncertainty. During the inference, the prediction is jointly made by both models through a dynamic fusion strategy. Our method is validated on the benchmark datasets, including Breakfast Actions, MultiTHUMOS, and Charades. The experiment re-sults show that our model achieves the state-of-the-art per-formance under both sufficient and insufficient data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 915085fc-8f8c-4d71-ae0d-01d49a3feaaaCited by top-tier papers13
- ProbVLM: Probabilistic Adapter for Frozen Vison-Language ModelsUddeshya Upadhyay, Shyamgopal Karthik, Massimiliano Mancini, Zeynep AkataICCV 2023 · 41 citations
- Uncertainty-guided Learning for Improving Image Manipulation DetectionKaixiang Ji, Feng Chen, Xin Guo, Yadong Xu et al.ICCV 2023 · 25 citations
- Physics-Augmented Autoencoder for 3D Skeleton-Based Gait RecognitionHongji Guo, Qiang JiICCV 2023 · 24 citations
- Towards Understanding Future: Consistency Guided Probabilistic Modeling for Action AnticipationZhao Xie, Yadong Shi, Kewei Wu, Yaru Cheng et al.AAAI 2024 · 9 citations
- Dynamic Aggregated Network for Gait RecognitionKang Ma, Ying Fu, Dezhi Zheng, Chunshui Cao et al.CVPR 2023
Builds on11
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- An End-to-End Transformer Model for 3D Object DetectionIshan Misra, Rohit Girdhar, Armand JoulinICCV 2021 · 602 citations
- Uncertainty-Guided Transformer Reasoning for Camouflaged Object DetectionFan Yang, Qiang Zhai, Xin Li, Rui Huang et al.ICCV 2021 · 293 citations
- Bayesian Graph Convolution LSTM for Skeleton Based Action RecognitionRui Zhao, Kang Wang, Hui Su, Qiang JiICCV 2019 · 104 citations
- Uncertainty-Aware Audiovisual Activity Recognition Using Deep Bayesian Variational InferenceMahesh Subedar, Ranganath Krishnan, Paulo Lopez-Meyer, Omesh Tickoo et al.ICCV 2019 · 81 citations
Related papers
- Uncertainty-aware Action Decoupling Transformer for Action AnticipationHongji Guo, Nakul Agarwal, Shao-Yuan Lo, Kwonjoon Lee et al.CVPR 2024
- Future Transformer for Long-term Action AnticipationDayoung Gong, Joonseok Lee, Manjin Kim, Seong Jong Ha et al.CVPR 2022 · 56 citations
- Probabilistic Distillation Transformer: Modelling Uncertainties for Visual Abductive ReasoningWanru Xu, Zhenjiang Miao, Yi Tian, Yigang Cen et al.ACM MM 2024
- Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose EstimationHan Li, Bowen Shi, Wenrui Dai, Hongwei Zheng et al.AAAI 2023 · 76 citations
- Breaking Spurious Correlations: Uncertainty-Driven Causal Transformers for AU DetectionYuru Wang, Yue ZhouCVPR 2026
