Uncertainty-Guided Probabilistic Transformer for Complex Action Recognition
Hongji Guo, Hanjing Wang, Qiang Ji
摘要
A complex action consists of a sequence of atomic actions that interact with each other over a relatively long period of time. This paper introduces a probabilistic model named Uncertainty-Guided Probabilistic Transformer (UGPT) for complex action recognition. The self-attention mechanism of a Transformer is used to capture the complex and long-term dynamics of the complex actions. By explicitly modeling the distribution of the attention scores, we extend the deterministic Transformer to a probabilistic Transformer in order to quantify the uncertainty of the pre-diction. The model prediction uncertainty is used to improve both training and inference. Specifically, we propose a novel training strategy by introducing a majority model and a minority model based on the epistemic uncertainty. During the inference, the prediction is jointly made by both models through a dynamic fusion strategy. Our method is validated on the benchmark datasets, including Breakfast Actions, MultiTHUMOS, and Charades. The experiment re-sults show that our model achieves the state-of-the-art per-formance under both sufficient and insufficient data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- ProbVLM: Probabilistic Adapter for Frozen Vison-Language ModelsUddeshya Upadhyay, Shyamgopal Karthik, Massimiliano Mancini, Zeynep AkataICCV 2023 · 被引用 41 次
- Uncertainty-guided Learning for Improving Image Manipulation DetectionKaixiang Ji, Feng Chen, Xin Guo, Yadong Xu 等ICCV 2023 · 被引用 25 次
- Physics-Augmented Autoencoder for 3D Skeleton-Based Gait RecognitionHongji Guo, Qiang JiICCV 2023 · 被引用 24 次
- Towards Understanding Future: Consistency Guided Probabilistic Modeling for Action AnticipationZhao Xie, Yadong Shi, Kewei Wu, Yaru Cheng 等AAAI 2024 · 被引用 9 次
- Dynamic Aggregated Network for Gait RecognitionKang Ma, Ying Fu, Dezhi Zheng, Chunshui Cao 等CVPR 2023
它引用的顶会 Paper11
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 被引用 4,104 次
- An End-to-End Transformer Model for 3D Object DetectionIshan Misra, Rohit Girdhar, Armand JoulinICCV 2021 · 被引用 602 次
- Uncertainty-Guided Transformer Reasoning for Camouflaged Object DetectionFan Yang, Qiang Zhai, Xin Li, Rui Huang 等ICCV 2021 · 被引用 293 次
- Bayesian Graph Convolution LSTM for Skeleton Based Action RecognitionRui Zhao, Kang Wang, Hui Su, Qiang JiICCV 2019 · 被引用 104 次
- Uncertainty-Aware Audiovisual Activity Recognition Using Deep Bayesian Variational InferenceMahesh Subedar, Ranganath Krishnan, Paulo Lopez-Meyer, Omesh Tickoo 等ICCV 2019 · 被引用 81 次
相关 Paper
- Uncertainty-aware Action Decoupling Transformer for Action AnticipationHongji Guo, Nakul Agarwal, Shao-Yuan Lo, Kwonjoon Lee 等CVPR 2024
- Future Transformer for Long-term Action AnticipationDayoung Gong, Joonseok Lee, Manjin Kim, Seong Jong Ha 等CVPR 2022 · 被引用 56 次
- Probabilistic Distillation Transformer: Modelling Uncertainties for Visual Abductive ReasoningWanru Xu, Zhenjiang Miao, Yi Tian, Yigang Cen 等ACM MM 2024
- Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose EstimationHan Li, Bowen Shi, Wenrui Dai, Hongwei Zheng 等AAAI 2023 · 被引用 76 次
- Breaking Spurious Correlations: Uncertainty-Driven Causal Transformers for AU DetectionYuru Wang, Yue ZhouCVPR 2026
