Multimodal Motion Conditioned Diffusion Model for Skeleton-based Video Anomaly Detection
Alessandro Flaborea, Luca Collorone, Guido Maria D'Amely di Melendugno, Stefano D'Arrigo, Bardh Prenkaj, Fabio Galasso
摘要
Anomalies are rare and anomaly detection is often therefore framed as One-Class Classification (OCC), i.e. trained solely on normalcy. Leading OCC techniques constrain the latent representations of normal 1 motions to limited volumes and detect as abnormal anything outside, which accounts satisfactorily for the openset'ness of anomalies. But normalcy shares the same openset'ness property since humans can perform the same action in several ways, which the leading techniques neglect. We propose a novel generative model for video anomaly detection (VAD), which assumes that both normality and abnormality are multimodal. We consider skeletal representations and leverage state-of-the-art diffusion probabilistic models to generate multimodal future human poses. We contribute a novel conditioning on the past motion of people and exploit the improved mode coverage capabilities of diffusion processes to generate different-but-plausible future motions. Upon the statistical aggregation of future modes, an anomaly is detected when the generated set of motions is not pertinent to the actual future. We validate our model on 4 established benchmarks: UBnormal, HR-UBnormal, HR-STC, and HR-Avenue, with extensive experiments surpassing state-of-the-art results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- On Diffusion Modeling for Anomaly DetectionVictor Livernoche, Vineet Jain, Yashar Hezaveh, Siamak RavanbakhshICLR 2024 · 被引用 74 次
- Multi-Scale Video Anomaly Detection by Multi-Grained Spatio-Temporal Representation LearningMenghao Zhang, Jingyu Wang, Qi Qi, Haifeng Sun 等CVPR 2024 · 被引用 29 次
- MULDE: Multiscale Log-Density Estimation via Denoising Score Matching for Video Anomaly DetectionJakub Micorek, Horst Possegger, Dominik Narnhofer, Horst Bischof 等CVPR 2024 · 被引用 21 次
- Dual Conditioned Motion Diffusion for Pose-Based Video Anomaly DetectionHongsong Wang, Andi Xu, Pinle Ding, Jie GuiAAAI 2025 · 被引用 8 次
- EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion GenerationXiaofeng Tan, Wanjiang Weng, Haodong Lei, Hongsong WangICLR 2026 · 被引用 6 次
它引用的顶会 Paper15
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
相关 Paper
- UBnormal: New Benchmark for Supervised Open-Set Video Anomaly DetectionAndra Acsintoae, Andrei Florescu, Mariana-Iuliana Georgescu, Tudor Mare 等CVPR 2022 · 被引用 153 次
- Normalizing Flows for Human Pose Anomaly DetectionOr Hirschorn, Shai AvidanICCV 2023 · 被引用 97 次
- Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion ModelHang Zhou, Jiale Cai, Yuteng Ye, Yonghui Feng 等AAAI 2025 · 被引用 23 次
- Graph Embedded Pose Clustering for Anomaly DetectionAmir Markovitz, Gilad Sharir, Itamar Friedman, Lihi Zelnik-Manor 等CVPR 2020
- A Multilevel Guidance-Exploration Network and Behavior-Scene Matching Method for Human Behavior Anomaly DetectionGuoqing Yang, Zhiming Luo, Jianzhe Gao, Yingxin Lai 等ACM MM 2024 · 被引用 1 次
