Rethinking Out-of-distribution (OOD) Detection: Masked Image Modeling is All You Need
Jingyao Li, Pengguang Chen, Zexin He, Shaozuo Yu, Shu Liu, Jiaya Jia
摘要
The core of out-of-distribution (OOD) detection is to learn the in-distribution (ID) representation, which is distinguishable from OOD samples. Previous work applied recognition-based methods to learn the ID features, which tend to learn shortcuts instead of comprehensive representations. In this work, we find surprisingly that simply using reconstruction-based methods could boost the performance of OOD detection significantly. We deeply explore the main contributors of OOD detection and find that reconstruction-based pretext tasks have the potential to provide a generally applicable and efficacious prior, which benefits the model in learning intrinsic data distributions of the ID dataset. Specifically, we take Masked Image Modeling as a pretext task for our OOD detection framework (MOOD). Without bells and whistles, MOOD outperforms previous SOTA of one-class OOD detection by 5.7%, multiclass OOD detection by 3.0%, and near-distribution OOD detection by 2.1%. It even defeats the 10-shot-per-class outlier exposure OOD detection, although we do not include any OOD samples for our detection. Codes are available at https://github.com/lijingyao20010602/MOOD .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- DAPE: Data-Adaptive Positional Encoding for Length ExtrapolationChuanyang Zheng, Yihang Gao, Han Shi, Minbin Huang 等NeurIPS 2024 · 被引用 42 次
- Out-of-Distribution Detection in Long-Tailed Recognition with Calibrated Outlier Class LearningWenjun Miao, Guansong Pang, Xiao Bai, Tianqi Li 等AAAI 2024 · 被引用 31 次
- Secure On-Device Video OOD Detection without BackpropagationShawn Li, Peilin Cai, Yuxiao Zhou, Zhiyu Ni 等ICCV 2025 · 被引用 28 次
- Adapting to Distribution Shift by Visual Domain Prompt GenerationZhixiang Chi, Li Gu, Tao Zhong, Huan Liu 等ICLR 2024 · 被引用 23 次
- Conjugated Semantic Pool Improves OOD Detection with Pre-trained Vision-Language ModelsMengyuan Chen, Junyu Gao, Changsheng XuNeurIPS 2024 · 被引用 21 次
它引用的顶会 Paper17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
相关 Paper
- Zero-Shot Out-of-Distribution Detection Based on the Pre-trained Model CLIPSepideh Esmaeilpour, Bing Liu, Eric Robertson, Lei ShuAAAI 2022 · 被引用 219 次
- Detecting Out-of-distribution Data through In-distribution Class PriorXue Jiang, Feng Liu, Zhen Fang, Hong Chen 等ICML 2023 · 被引用 33 次
- MOOD: Multi-Level Out-of-Distribution DetectionZiqian Lin, Sreya Dutta Roy, Yixuan LiCVPR 2021
- YolOOD: Utilizing Object Detection Concepts for Multi-Label Out-of-Distribution DetectionAlon Zolfi, Guy Amit, Amit Baras, Satoru Koda 等CVPR 2024 · 被引用 7 次
- Diffusion-based Layer-wise Semantic Reconstruction for Unsupervised Out-of-Distribution DetectionYing Yang, De Cheng, Chaowei Fang, Yubiao Wang 等NeurIPS 2024 · 被引用 9 次
