OpenTAL: Towards Open Set Temporal Action Localization
Wentao Bao, Qi Yu, Yu Kong
摘要
Temporal Action Localization (TAL) has experienced remarkable success under the supervised learning paradigm. However, existing TAL methods are rooted in the closed set assumption, which cannot handle the inevitable unknown actions in open-world scenarios. In this paper, we, for the first time, step toward the Open Set TAL (OSTAL) problem and propose a general framework Open TAL based on Evidential Deep Learning (EDL). Specifically, the OpenTAL consists of uncertainty-aware action classification, actionness prediction, and temporal location regression. With the proposed importance-balanced EDL method, classification uncertainty is learned by collecting categorical evidence majorly from important samples. To distinguish the unknown actions from background video frames, the actionness is learned by the positive-unlabeled learning. The classification uncertainty is further calibrated by leveraging the guidance from the temporal localization quality. The OpenTAL is general to enable existing TAL models for open set scenarios, and experimental results on THUMOS14 and ActivityNet1.3 benchmarks show the effectiveness of our method. The code and pre-trained models are released at https://www.rit.edu/actionlab/opental.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Is Out-of-Distribution Detection Learnable?Zhen Fang, Yixuan Li, Jie Lu, Jiahua Dong 等NeurIPS 2022 · 被引用 188 次
- Uncertainty Estimation by Fisher Information-based Evidential Deep LearningDanruo Deng, Guangyong Chen, Yang Yu, Furui Liu 等ICML 2023 · 被引用 82 次
- Does Video-Text Pretraining Help Open-Vocabulary Online Action Detection?Qingsong Zhao, Yi Wang, Jilan Xu, Yinan He 等NeurIPS 2024 · 被引用 16 次
- Weakly-Supervised Residual Evidential Learning for Multi-Instance Uncertainty EstimationPei Liu, Luping JiICML 2024 · 被引用 9 次
- HSIC-based Moving Weight Averaging for Few-Shot Open-Set Object DetectionBinyi Su, Hua Zhang, Zhong ZhouACM MM 2023 · 被引用 8 次
它引用的顶会 Paper23
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 被引用 4,104 次
- Deep Evidential RegressionAlexander Amini, Wilko Schwarting, Ava Soleimany, Daniela RusNeurIPS 2020 · 被引用 777 次
- BMN: Boundary-Matching Network for Temporal Action Proposal GenerationTianwei Lin, Xiao Liu, Xin Li, Errui Ding 等ICCV 2019 · 被引用 709 次
- Graph Convolutional Networks for Temporal Action LocalizationRunhao Zeng, Wenbing Huang, Chuang Gan, Mingkui Tan 等ICCV 2019 · 被引用 536 次
- Evidential Deep Learning for Open Set Action RecognitionWentao Bao, Qi Yu, Yu KongICCV 2021 · 被引用 204 次
相关 Paper
- Cascade Evidential Learning for Open-world Weakly-supervised Temporal Action LocalizationMengyuan Chen, Junyu Gao, Changsheng XuCVPR 2023
- OpenAVE: Moving towards Open Set Audio-Visual Event LocalizationJiale Yu, Baopeng Zhang, Zhu Teng, Jianping FanACM MM 2024 · 被引用 2 次
- Learning Generalized Representations for Open-Set Temporal Action LocalizationJunshan Hu, Liansheng Zhuang, Weisong Dong, Shiming Ge 等ACM MM 2023 · 被引用 4 次
- CAG-QIL: Context-Aware Actionness Grouping via Q Imitation Learning for Online Temporal Action LocalizationHyolim Kang, Kyungmin Kim, Yumin Ko, Seon Joo KimICCV 2021 · 被引用 18 次
- Towards Evidential and Class Separable Open Set Object DetectionRuofan Wang, Rui-Wei Zhao, Xiaobo Zhang, Rui FengAAAI 2024 · 被引用 12 次
