Unsupervised Multi-View Visual Anomaly Detection via Progressive Homography-Guided Alignment
Xintao Chen, Xiaohao Xu, Bozhong Zheng, Yun Liu, Yingna Wu
摘要
Unsupervised visual anomaly detection from multi-view images presents a significant challenge: distinguishing genuine defects from benign appearance variations caused by viewpoint changes. Existing methods, often designed for single-view inputs, treat multiple views as a disconnected set of images, leading to inconsistent feature representations and a high false-positive rate. To address this, we introduce ViewSense-AD (VSAD), a novel framework that learns viewpoint-invariant representations by explicitly modeling geometric consistency across views. At its core is our Multi-View Alignment Module (MVAM), which leverages homography to project and align corresponding feature regions between neighboring views. We integrate MVAM into a View-Align Latent Diffusion Model (VALDM), enabling progressive and multi-stage alignment during the denoising process. This allows the model to build a coherent and holistic understanding of the object's surface from coarse to fine scales. Furthermore, a lightweight Fusion Refiner Module (FRM) enhances the global consistency of the aligned features, suppressing noise and improving discriminative power. Anomaly detection is performed by comparing multi-level features from the diffusion model against a learned memory bank of normal prototypes. Extensive experiments on the challenging RealIAD and MANTA datasets demonstrate that VSAD sets a new state-of-the-art, significantly outperforming existing methods in pixel, view, and sample-level visual anomaly detection, proving its robustness to large viewpoint shifts and complex textures. Our code will be released to drive further research.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Towards Total Recall in Industrial Anomaly DetectionKarsten Roth, Latha Pemula, Joaquin Zepeda, Bernhard Schölkopf 等CVPR 2022 · 被引用 1,301 次
- A Diffusion-Based Framework for Multi-Class Anomaly DetectionHaoyang He, Jiangning Zhang, Hongxu Chen, Xuhai Chen 等AAAI 2024 · 被引用 231 次
- Appearance-Motion Memory Consistency Network for Video Anomaly DetectionRuichu Cai, Hao Zhang, Wen Liu, Shenghua Gao 等AAAI 2021 · 被引用 223 次
相关 Paper
- Learning Invariant Discriminative Patterns for Unified Anomaly DetectionChengcheng Xing, Yanyu Xu, Yonghui Xu, Lizhen CuiACM MM 2025
- Unsupervised Surface Anomaly Detection with Diffusion Probabilistic ModelXinyi Zhang, Naiqi Li, Jiawei Li, Tao Dai 等ICCV 2023 · 被引用 112 次
- DZAD: Diffusion-based Zero-shot Anomaly DetectionTianrui Zhang, Liang Gao, Xinyu Li, Yiping GaoAAAI 2025 · 被引用 5 次
- Is Task-Specific Training Necessary for Anomaly Detection?Xingwu Zhang, Guanxuan Li, Paul Henderson, Gerardo Aragon-Camarasa 等ICML 2026 · 被引用 1 次
- UniAD: Integrating Geometric and Semantic Cues for Unified Anomaly DetectionXiaodong Wang, Hongmin Hu, Fei Yan, Junwen Lu 等ACM MM 2025 · 被引用 2 次
