Evidential Deep Learning for Open Set Action Recognition
Wentao Bao, Qi Yu, Yu Kong
Abstract
In a real-world scenario, human actions are typically out of the distribution from training data, which requires a model to both recognize the known actions and reject the unknown. Different from image data, video actions are more challenging to be recognized in an open-set setting due to the uncertain temporal dynamics and static bias of human actions. In this paper, we propose a Deep Evidential Action Recognition (DEAR) method to recognize actions in an open testing set. Specifically, we formulate the action recognition problem from the evidential deep learning (EDL) perspective and propose a novel model calibration method to regularize the EDL training. Besides, to mitigate the static bias of video representation, we propose a plug-and-play module to debias the learned representation through contrastive learning. Experimental results show that our DEAR method achieves consistent performance gain on multiple mainstream action recognition models and benchmarks. Code and pre-trained models are available at https://www.rit.edu/actionlab/dear .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1e1ed7ed-f56a-48c0-8569-4f0141d4e8f3Cited by top-tier papers60
- Is Out-of-Distribution Detection Learnable?Zhen Fang, Yixuan Li, Jie Lu, Jiahua Dong et al.NeurIPS 2022 · 188 citations
- UBnormal: New Benchmark for Supervised Open-Set Video Anomaly DetectionAndra Acsintoae, Andrei Florescu, Mariana-Iuliana Georgescu, Tudor Mare et al.CVPR 2022 · 153 citations
- TrEP: Transformer-Based Evidential Prediction for Pedestrian Intention with UncertaintyZhengming Zhang, Renran Tian, Zhengming DingAAAI 2023 · 84 citations
- Uncertainty Estimation by Fisher Information-based Evidential Deep LearningDanruo Deng, Guangyong Chen, Yang Yu, Furui Liu et al.ICML 2023 · 82 citations
- Nearest Neighbor Guidance for Out-of-Distribution DetectionJaewoo Park, Yoon Gyo Jung, Andrew Beng Jin TeohICCV 2023 · 74 citations
Builds on14
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- TSM: Temporal Shift Module for Efficient Video UnderstandingJi Lin, Chuang Gan, Song HanICCV 2019 · 2,049 citations
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 777 citations
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz et al.NeurIPS 2020 · 674 citations
- Learning De-biased Representations with Biased RepresentationsHyojin Bahng, Sanghyuk Chun, Sangdoo Yun, Jaegul Choo et al.ICML 2020 · 332 citations
Related papers
- OpenTAL: Towards Open Set Temporal Action LocalizationWentao Bao, Qi Yu, Yu KongCVPR 2022 · 30 citations
- SOAR: Scene-debiasing Open-set Action RecognitionYuanhao Zhai, Ziyi Liu, Zhenyu Wu, Yi Wu et al.ICCV 2023 · 15 citations
- Towards Evidential and Class Separable Open Set Object DetectionRuofan Wang, Rui-Wei Zhao, Xiaobo Zhang, Rui FengAAAI 2024 · 12 citations
- Open Set Action Recognition via Multi-Label Evidential LearningChen Zhao, Dawei Du, Anthony Hoogs, Christopher FunkCVPR 2023
- OpenAVE: Moving towards Open Set Audio-Visual Event LocalizationJiale Yu, Baopeng Zhang, Zhu Teng, Jianping FanACM MM 2024 · 2 citations
