Hierarchical Semantic Contrast for Scene-aware Video Anomaly Detection
Shengyang Sun, Xiaojin Gong
Abstract
Increasing scene-awareness is a key challenge in video anomaly detection (VAD). In this work, we propose a hierarchical semantic contrast (HSC) method to learn a sceneaware VAD model from normal videos. We first incorporate foreground object and background scene features with highlevel semantics by taking advantage of pre-trained video parsing models. Then, building upon the autoencoderbased reconstruction framework, we introduce both scenelevel and object-level contrastive learning to enforce the encoded latent features to be compact within the same semantic classes while being separable across different classes. This hierarchical semantic contrast strategy helps to deal with the diversity of normal patterns and also increases their discrimination ability. Moreover, for the sake of tackling rare normal activities, we design a skeleton-based motion augmentation to increase samples and refine the model further. Extensive experiments on three public datasets and scene-dependent mixture datasets validate the effectiveness of our proposed method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30a0e0d5-673e-43bb-80f2-8254f0738158Cited by top-tier papers11
- Harnessing Large Language Models for Training-Free Video Anomaly DetectionLuca Zanella, Willi Menapace, Massimiliano Mancini, Yiming Wang et al.CVPR 2024 · 57 citations
- Open-Vocabulary Video Anomaly DetectionPeng Wu, Xuerong Zhou, Guansong Pang, Yujia Sun et al.CVPR 2024 · 56 citations
- Multi-Scale Video Anomaly Detection by Multi-Grained Spatio-Temporal Representation LearningMenghao Zhang, Jingyu Wang, Qi Qi, Haifeng Sun et al.CVPR 2024 · 29 citations
- Towards Multi-Domain Learning for Generalizable Video Anomaly DetectionMyeongAh Cho, Taeoh Kim, Minho Shim, Dongyoon Wee et al.NeurIPS 2024 · 14 citations
- Open-World Semantic Segmentation Including Class SimilarityMatteo Sodano, Federico Magistri, Lucas Nunes, Jens Behley et al.CVPR 2024 · 8 citations
Builds on29
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Revisiting Skeleton-based Action RecognitionHaodong Duan, Yue Zhao, Kai Chen, Dahua Lin et al.CVPR 2022 · 752 citations
- Self-paced Contrastive Learning with Hybrid Memory for Domain Adaptive Object Re-IDYixiao Ge, Feng Zhu, Dapeng Chen, Rui Zhao et al.NeurIPS 2020 · 688 citations
- Weakly-supervised Video Anomaly Detection with Robust Temporal Feature Magnitude LearningYu Tian, Guansong Pang, Yuanhong Chen, Rajvinder Singh et al.ICCV 2021 · 495 citations
Related papers
- Learning to Tell Apart: Weakly Supervised Video Anomaly Detection via Disentangled Semantic AlignmentWenti Yin, Huaxin Zhang, Xiang Wang, Yuqing Lu et al.AAAI 2026
- Cluster Attention Contrast for Video Anomaly DetectionZiming Wang, Yuexian Zou, Zeming ZhangACM MM 2020 · 90 citations
- Hierarchical Scene Normality-Binding Modeling for Anomaly Detection in Surveillance VideosQianyue Bao, Fang Liu, Yang Liu, Licheng Jiao et al.ACM MM 2022 · 50 citations
- Effective Video Abnormal Event Detection by Learning A Consistency-Aware High-Level Feature ExtractorGuang Yu, Siqi Wang, Zhiping Cai, Xinwang Liu et al.ACM MM 2022 · 7 citations
- Prompt-Enhanced Multiple Instance Learning for Weakly Supervised Video Anomaly DetectionJunxi Chen, Liang Li, Li Su, Zheng-Jun Zha et al.CVPR 2024
