Dinomaly: The Less Is More Philosophy in Multi-Class Unsupervised Anomaly Detection
Jia Guo, Shuai Lu, Weihang Zhang, Fang Chen, Huiqi Li, Hongen Liao
摘要
Recent studies highlighted a practical setting of unsupervised anomaly detection (UAD) that builds a unified model for multi-class images. Despite various advancements addressing this challenging task, the detection performance under the multi-class setting still lags far behind state-of-the-art class-separated models. Our research aims to bridge this substantial performance gap. In this paper, we present Dinomaly, a minimalist reconstructionbased anomaly detection framework that harnesses pure Transformer architectures without relying on complex designs, additional modules, or specialized tricks. Given this powerful framework consisting of only Attentions and MLPs, we found four simple components that are essential to multi-class anomaly detection: (1) Scalable foundation Transformers that extracts universal and discriminative features, (2) Noisy Bottleneck where pre-existing Dropouts do all the noise injection tricks, (3) Linear Attention that naturally cannot focus, and (4) Loose Reconstruction that does not force layer-to-layer and point-bypoint reconstruction. Extensive experiments are conducted across popular anomaly detection benchmarks including MVTec-AD, VisA, Real-IAD, etc. Our proposed Dinomaly achieves impressive image-level AUROC of 99.6%, 98.7%, and 89.3% on the three datasets respectively, which is not only superior to state-of-the-art multi-class UAD methods, but also achieves the most advanced class-separated UAD records. Code is available at: https://github.com/ guojiajeremy/Dinomaly
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Normal-Abnormal Guided Generalist Anomaly DetectionYuexin Wang, Xiaolei Wang, Yizheng Gong, Jimin XiaoNeurIPS 2025 · 被引用 16 次
- Towards an Incremental Unified Multimodal Anomaly Detection: Augmenting Multimodal Denoising From an Information Bottleneck PerspectiveKaifang Long, Lianbo Ma, Jiaqi Liu, liming liu 等CVPR 2026 · 被引用 5 次
- Reasoning-Driven Anomaly Detection and Localization with Image-Level SupervisionYizhou Jin, Yuezhu Feng, Jinjin Zhang, Peng Wang 等CVPR 2026 · 被引用 4 次
- A Semantically Disentangled Unified Model for Multi-category 3D Anomaly DetectionSuYeon Kim, Wongyu Lee, MyeongAh ChoCVPR 2026 · 被引用 3 次
- AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly DetectionZhaopeng Gu, Bingke Zhu, Guibo Zhu, Yingying Chen 等AAAI 2026 · 被引用 3 次
它引用的顶会 Paper30
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- An Empirical Study of Training Self-Supervised Vision TransformersXinlei Chen, Saining Xie, Kaiming HeICCV 2021 · 被引用 2,340 次
- Towards Total Recall in Industrial Anomaly DetectionKarsten Roth, Latha Pemula, Joaquin Zepeda, Bernhard Schölkopf 等CVPR 2022 · 被引用 1,301 次
相关 Paper
- A Unified Model for Multi-class Anomaly DetectionZhiyuan You, Lei Cui, Yujun Shen, Kai Yang 等NeurIPS 2022 · 被引用 585 次
- Hierarchical Vector Quantized Transformer for Multi-class Unsupervised Anomaly DetectionRuiying Lu, Yujie Wu, Long Tian, Dongsheng Wang 等NeurIPS 2023 · 被引用 121 次
- A Diffusion-Based Framework for Multi-Class Anomaly DetectionHaoyang He, Jiangning Zhang, Hongxu Chen, Xuhai Chen 等AAAI 2024 · 被引用 231 次
- Learning Unsupervised Metaformer for Anomaly DetectionJhih-Ciang Wu, Ding-Jie Chen, Chiou-Shann Fuh, Tyng-Luh LiuICCV 2021 · 被引用 101 次
- Is Task-Specific Training Necessary for Anomaly Detection?Xingwu Zhang, Guanxuan Li, Paul Henderson, Gerardo Aragon-Camarasa 等ICML 2026 · 被引用 1 次
