Towards Visual Discrimination and Reasoning of Real-World Physical Dynamics: Physics-Grounded Anomaly Detection
Wenqiao Li, Yao Gu, Xintao Chen, Xiaohao Xu, Ming Hu, Xiaonan Huang, Yingna Wu
摘要
Band (b) Interaction (a) Object (c) Video with Physical Dynamics Figure 1. Towards visual discrimination of physical dynamics in real-world industrial object anomaly detection. We illustrate objects, interactions, and time-sequenced videos from the Physics-Grounded Anomaly Detection dataset: (a) Object; (b) Interaction: Applied actions shown with directional arrows; (c) Video with Physical Dynamics: Temporal sequences showing normal and abnormal states, highlighting anomalies like leaks, misalignments, and cracks. By focusing on the dynamic behaviors of complex objects, we enhance understanding of interactions and failure modes in real-world settings, where both structure and motion contribute to anomaly detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- A Hybrid Video Anomaly Detection Framework via Memory-Augmented Flow Reconstruction and Flow-Guided Frame PredictionZhian Liu, Yongwei Nie, Chengjiang Long, Qing Zhang 等ICCV 2021 · 被引用 341 次
- Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language ModelsMuhammad Maaz, Hanoona Abdul Rasheed, Salman Khan, Fahad KhanACL 2024 · 被引用 279 次
- Video-LLaVA: Learning United Visual Representation by Alignment Before ProjectionBin Lin, Yang Ye, Bin Zhu, Jiaxi Cui 等EMNLP 2024 · 被引用 231 次
- VadCLIP: Adapting Vision-Language Models for Weakly Supervised Video Anomaly DetectionPeng Wu, Xuerong Zhou, Guansong Pang, Lingru Zhou 等AAAI 2024 · 被引用 220 次
相关 Paper
- Interactive Anomaly Detection for Articulated Objects via Motion AnticipationAnkan Bhunia, Changjian Li, Hakan BilenNeurIPS 2025 · 被引用 1 次
- Human-Object-Object Interaction: Towards Human-Centric Complex Interaction DetectionMingxuan Zhang, Xiao Wu, Zhaoquan Yuan, Qi He 等ACM MM 2023 · 被引用 6 次
- PhysInOne: Visual Physics Learning and Reasoning in One SuiteSiyuan Zhou, Hejun Wang, Hu Cheng, Jinxi Li 等CVPR 2026 · 被引用 9 次
- CRIPP-VQA: Counterfactual Reasoning about Implicit Physical Properties via Video Question AnsweringMaitreya Patel, Tejas Gokhale, Chitta Baral, Yezhou YangEMNLP 2022 · 被引用 5 次
- Kaputt: A Large-Scale Dataset for Visual Defect DetectionSebastian Höfer, Dorian Fritz Henning, Artemij Amiranashvili, Douglas Morrison 等ICCV 2025 · 被引用 2 次
