Unsupervised Multi-View Visual Anomaly Detection via Progressive Homography-Guided Alignment
Xintao Chen, Xiaohao Xu, Bozhong Zheng, Yun Liu, Yingna Wu
Abstract
Unsupervised visual anomaly detection from multi-view images presents a significant challenge: distinguishing genuine defects from benign appearance variations caused by viewpoint changes. Existing methods, often designed for single-view inputs, treat multiple views as a disconnected set of images, leading to inconsistent feature representations and a high false-positive rate. To address this, we introduce ViewSense-AD (VSAD), a novel framework that learns viewpoint-invariant representations by explicitly modeling geometric consistency across views. At its core is our Multi-View Alignment Module (MVAM), which leverages homography to project and align corresponding feature regions between neighboring views. We integrate MVAM into a View-Align Latent Diffusion Model (VALDM), enabling progressive and multi-stage alignment during the denoising process. This allows the model to build a coherent and holistic understanding of the object's surface from coarse to fine scales. Furthermore, a lightweight Fusion Refiner Module (FRM) enhances the global consistency of the aligned features, suppressing noise and improving discriminative power. Anomaly detection is performed by comparing multi-level features from the diffusion model against a learned memory bank of normal prototypes. Extensive experiments on the challenging RealIAD and MANTA datasets demonstrate that VSAD sets a new state-of-the-art, significantly outperforming existing methods in pixel, view, and sample-level visual anomaly detection, proving its robustness to large viewpoint shifts and complex textures. Our code will be released to drive further research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2b1ef3e9-a727-4a8e-9ab9-b992ccc7b9ebBuilds on24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Towards Total Recall in Industrial Anomaly DetectionKarsten Roth, Latha Pemula, Joaquin Zepeda, Bernhard Schölkopf et al.CVPR 2022 · 1,301 citations
- A Diffusion-Based Framework for Multi-Class Anomaly DetectionHaoyang He, Jiangning Zhang, Hongxu Chen, Xuhai Chen et al.AAAI 2024 · 231 citations
- Appearance-Motion Memory Consistency Network for Video Anomaly DetectionRuichu Cai, Hao Zhang, Wen Liu, Shenghua Gao et al.AAAI 2021 · 223 citations
Related papers
- Learning Invariant Discriminative Patterns for Unified Anomaly DetectionChengcheng Xing, Yanyu Xu, Yonghui Xu, Lizhen CuiACM MM 2025
- Unsupervised Surface Anomaly Detection with Diffusion Probabilistic ModelXinyi Zhang, Naiqi Li, Jiawei Li, Tao Dai et al.ICCV 2023 · 112 citations
- DZAD: Diffusion-based Zero-shot Anomaly DetectionTianrui Zhang, Liang Gao, Xinyu Li, Yiping GaoAAAI 2025 · 5 citations
- Is Task-Specific Training Necessary for Anomaly Detection?Xingwu Zhang, Guanxuan Li, Paul Henderson, Gerardo Aragon-Camarasa et al.ICML 2026 · 1 citation
- UniAD: Integrating Geometric and Semantic Cues for Unified Anomaly DetectionXiaodong Wang, Hongmin Hu, Fei Yan, Junwen Lu et al.ACM MM 2025 · 2 citations
