Dual Conditioned Motion Diffusion for Pose-Based Video Anomaly Detection
Hongsong Wang, Andi Xu, Pinle Ding, Jie Gui
Abstract
Video Anomaly Detection (VAD) is essential for computer vision and multimedia research. Existing VAD methods utilize either reconstruction-based or prediction-based frameworks. The former excels at detecting irregular patterns or structures, whereas the latter is capable of spotting abnormal deviations or trends. We address pose-based video anomaly detection and introduce a novel framework called Dual Conditioned Motion Diffusion (DCMD), which enjoys the advantages of both approaches. The DCMD integrates conditioned motion and conditioned embedding to comprehensively utilize the pose characteristics and latent semantics of observed movements, respectively. In the reverse diffusion process, a motion transformer is proposed to capture potential correlations from multi-layered characteristics within the spectrum space of human motion. To enhance the discriminability between normal and abnormal instances, we design a novel United Association Discrepancy (UAD) regularization that primarily relies on a Gaussian kernel-based time association and a self-attention-based global association. Finally, a mask completion strategy is introduced during the inference stage of the reverse diffusion process to enhance the utilization of conditioned motion for the prediction branch of anomaly detection. Extensive experiments conducted on four datasets demonstrate that our method dramatically outperforms state-of-the-art methods and exhibits superior generalization performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 42ea2e0d-f0eb-4129-bb09-ffeff1b5926bCited by top-tier papers1
Ask how each one uses itBuilds on12
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Memorizing Normality to Detect Anomaly: Memory-Augmented Deep Autoencoder for Unsupervised Anomaly DetectionDong Gong, Lingqiao Liu, Vuong Le, Budhaditya Saha et al.ICCV 2019 · 1,646 citations
- Anomaly Transformer: Time Series Anomaly Detection with Association DiscrepancyJiehui Xu, Haixu Wu, Jianmin Wang, Mingsheng LongICLR 2022 · 960 citations
- Learning Trajectory Dependencies for Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu Salzmann, Hongdong LiICCV 2019 · 534 citations
- Appearance-Motion Memory Consistency Network for Video Anomaly DetectionRuichu Cai, Hao Zhang, Wen Liu, Shenghua Gao et al.AAAI 2021 · 223 citations
Related papers
- Multimodal Motion Conditioned Diffusion Model for Skeleton-based Video Anomaly DetectionAlessandro Flaborea, Luca Collorone, Guido Maria D'Amely di Melendugno, Stefano D'Arrigo et al.ICCV 2023 · 83 citations
- Video Event Restoration Based on Keyframes for Video Anomaly DetectionZhiwei Yang, Jing Liu, Zhaoyang Wu, Peng Wu et al.CVPR 2023
- Feature Prediction Diffusion Model for Video Anomaly DetectionCheng Yan, Shiyu Zhang, Yang Liu, Guansong Pang et al.ICCV 2023 · 76 citations
- Convolutional Transformer based Dual Discriminator Generative Adversarial Networks for Video Anomaly DetectionXinyang Feng, Dongjin Song, Yuncong Chen, Zhengzhang Chen et al.ACM MM 2021 · 101 citations
- Autoregressive Denoising Score Matching Is a Good Video Anomaly DetectorHanwen Zhang, Congqi Cao, Qinyi Lv, Lingtong Min et al.ICCV 2025 · 3 citations
