Prompt-Enhanced Multiple Instance Learning for Weakly Supervised Video Anomaly Detection
Junxi Chen, Liang Li, Li Su, Zheng-Jun Zha, Qingming Huang
摘要
Weakly-supervised Video Anomaly Detection (wVAD) aims to detect frame-level anomalies using only videolevel labels in training. Due to the limitation of coarsegrained labels, Multi-Instance Learning (MIL) is prevailing in wVAD. However, MIL suffers from insufficiency of binary supervision to model diverse abnormal patterns. Besides, the coupling between abnormality and its context hinders the learning of clear abnormal event boundary. In this paper, we propose prompt-enhanced MIL to detect various abnormal events while ensuring clear event boundaries. Concretely, we design the abnormal-aware prompts by using abnormal class annotations together with learnable prompt, which can incorporate semantic priors into video features dynamically. The detector can utilize the semantic-rich features to capture diverse abnormal patterns. In addition, normal context prompt is introduced to amplify the distinction between abnormality and its context, facilitating the generation of clear boundary. With the mutual enhancement of abnormal-aware and normal context prompt, the model can construct discriminative representations to detect divergent anomalies without ambiguous event boundaries. Extensive experiments demonstrate our method achieves SOTA performance on three public benchmarks. The code is available at https://github. com/Junxi-Chen/PE-MIL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Vad-R1: Towards Video Anomaly Reasoning via Perception-to-Cognition Chain-of-ThoughtChao Huang, Benfeng Wang, Wei Wang, Jie Wen 等NeurIPS 2025 · 被引用 30 次
- Generalizing Single-Frame Supervision to Event-Level Understanding for Video Anomaly DetectionJunxi Chen, Liang Li, Yunbin Tu, Li Su 等NeurIPS 2025 · 被引用 4 次
- Learning to Watch: Active Video Anomaly Understanding via Interleaved Policy OptimizationMengjingcheng Mo, Jiaxu Leng, Xinbo GaoICML 2026 · 被引用 1 次
- Designing Multi-Robot Ground Video Sensemaking with Public Safety ProfessionalsPuqi Zhou, Ali Asgarov, Aafiya Hussain, Wonjoon Park 等CHI 2026 · 被引用 1 次
- Linguistic Relative Policy Optimization for Video Anomaly ReasoningJiaxu Leng, Jiankang Zheng, Mengjingcheng Mo, Zhanjie Wu 等ICML 2026
它引用的顶会 Paper13
- Weakly-supervised Video Anomaly Detection with Robust Temporal Feature Magnitude LearningYu Tian, Guansong Pang, Yuanhong Chen, Rajvinder Singh 等ICCV 2021 · 被引用 495 次
- A Hybrid Video Anomaly Detection Framework via Memory-Augmented Flow Reconstruction and Flow-Guided Frame PredictionZhian Liu, Yongwei Nie, Chengjiang Long, Qing Zhang 等ICCV 2021 · 被引用 341 次
- DetCLIP: Dictionary-Enriched Visual-Concept Paralleled Pre-training for Open-world DetectionLewei Yao, Jianhua Han, Youpeng Wen, Xiaodan Liang 等NeurIPS 2022 · 被引用 285 次
- Self-Training Multi-Sequence Learning with Transformer for Weakly Supervised Video Anomaly DetectionShuo Li, Fang Liu, Licheng JiaoAAAI 2022 · 被引用 282 次
- VadCLIP: Adapting Vision-Language Models for Weakly Supervised Video Anomaly DetectionPeng Wu, Xuerong Zhou, Guansong Pang, Lingru Zhou 等AAAI 2024 · 被引用 220 次
相关 Paper
- Weakly Supervised Video Anomaly Detection and Localization with Spatio-Temporal PromptsPeng Wu, Xuerong Zhou, Guansong Pang, Zhiwei Yang 等ACM MM 2024 · 被引用 50 次
- TLMA: Mitigating the Impact of Weakly Labeled Information for Video Anomaly DetectionRong Xu, Runqi Wang, Yingjun Zhang, Tao Tao 等CVPR 2026
- Unbiased Multiple Instance Learning for Weakly Supervised Video Anomaly DetectionHui Lv, Zhongqi Yue, Qianru Sun, Bin Luo 等CVPR 2023
- Text Prompt with Normality Guidance for Weakly Supervised Video Anomaly DetectionZhiwei Yang, Jing Liu, Peng WuCVPR 2024 · 被引用 55 次
- MIST: Multiple Instance Self-Training Framework for Video Anomaly DetectionJia-Chang Feng, Fa-Ting Hong, Wei-Shi ZhengCVPR 2021
