ODIM: Outlier Detection via Likelihood of Under-Fitted Generative Models
Dongha Kim, Jaesung Hwang, Jongjin Lee, Kunwoong Kim, Yongdai Kim
Abstract
The unsupervised outlier detection (UOD) problem refers to a task to identify inliers given training data which contain outliers as well as inliers, without any labeled information about inliers and outliers. It has been widely recognized that using fully-trained likelihood-based deep generative models (DGMs) often results in poor performance in distinguishing inliers from outliers. In this study, we claim that the likelihood itself could serve as powerful evidence for identifying inliers in UOD tasks, provided that DGMs are carefully under-fitted. Our approach begins with a novel observation called the inlier-memorization (IM) effect-when training a deep generative model with data including outliers, the model initially memorizes inliers before outliers. Based on this finding, we develop a new method called the outlier detection via the IM effect (ODIM). Remarkably, the ODIM requires only a few updates, making it computationally efficient-at least tens of times faster than other deep-learning-based algorithms. Also, the ODIM filters out outliers excellently, regardless of the data type, including tabular, image, and text data. To validate the superiority and efficiency of our method, we provide extensive empirical analyses on close to 60 datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 37d755b2-ac5e-4831-bbb4-0715b39df770Cited by top-tier papers2
- UniOD: A Universal Model for Outlier Detection across Diverse DomainsDazhi Fu, Jicong FanICLR 2026 · 1 citation
- Memorize Early, Then Query: Inlier-Memorization-Guided Active Outlier DetectionMinseo Kang, Seunghwan Park, Dongha KimAAAI 2026
Builds on15
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan et al.CCS 2016 · 7,620 citations
- Deep Semi-Supervised Anomaly DetectionLukas Ruff, Robert A. Vandermeulen, Nico Görnitz, Alexander Binder et al.ICLR 2020 · 678 citations
Related papers
- ALTBI: Constructing Improved Outlier Detection Models via Optimization of Inlier-Memorization EffectSeoyoung Cho, Jaesung Hwang, Kwan-Young Bak, Dongha KimAAAI 2025
- Automatic Unsupervised Outlier Model SelectionYue Zhao, Ryan A. Rossi, Leman AkogluNeurIPS 2021 · 104 citations
- Further Analysis of Outlier Detection with Deep Generative ModelsZiyu Wang, Bin Dai, David P. Wipf, Jun ZhuNeurIPS 2020 · 45 citations
- Understanding Failures in Out-of-Distribution Detection with Deep Generative ModelsLily H. Zhang, Mark Goldstein, Rajesh RanganathICML 2021 · 129 citations
- Hierarchical VAEs Know What They Don't KnowJakob Drachmann Havtorn, Jes Frellsen, Søren Hauberg, Lars MaaløeICML 2021 · 87 citations
