FodFoM: Fake Outlier Data by Foundation Models Creates Stronger Visual Out-of-Distribution Detector
Jiankang Chen, Ling Deng, Zhiyong Gan, Wei-Shi Zheng, Ruixuan Wang
摘要
Out-of-Distribution (OOD) detection is crucial when deploying machine learning models in open-world applications. The core challenge in OOD detection is mitigating the model's overconfidence on OOD data. While recent methods using auxiliary outlier datasets or synthesizing outlier features have shown promising OOD detection performance, they are limited due to costly data collection or simplified assumptions. In this paper, we propose a novel OOD detection framework FodFoM that innovatively combines multiple foundation models to generate two types of challenging fake outlier images for classifier training. The first type is based on BLIP-2's image captioning capability, CLIP's vision-language knowledge, and Stable Diffusion's image generation ability. Jointly utilizing these foundation models constructs fake outlier images which are semantically similar to but different from in-distribution (ID) images. For the second type, GroundingDINO's object detection ability is utilized to help construct pure background images by blurring foreground ID objects in ID images. The proposed framework can be flexibly combined with multiple existing OOD detection methods. Extensive empirical evaluations show that image classifiers with the help of constructed fake images can more accurately differentiate real OOD image from ID ones. New state-of-the-art OOD detection performance is achieved on multiple benchmarks. The code is available at https://github.com/Cverchen/ACMMM2024-FodFoM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution DetectionXin Gao, Jiyao Liu, Guanghao Li, Yueming Lyu 等NeurIPS 2025 · 被引用 9 次
- TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language ModelsJinlun Ye, Jiang Liao, Runhe Lai, Xinhua Lu 等CVPR 2026 · 被引用 2 次
- DCAC: Dynamic Class-Aware Cache Creates Stronger Out-of-Distribution DetectorsYanqi Wu, Qichao Chen, Runhe Lai, Xinhua Lu 等AAAI 2026 · 被引用 1 次
- Image-based Outlier Synthesis With Training DataSudarshan RegmiCVPR 2026
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
相关 Paper
- Synthesizing Near-Boundary OOD Samples for Out-of-Distribution DetectionJinglun Li, Kaixun Jiang, Zhaoyu Chen, Bo Li 等ICCV 2025 · 被引用 2 次
- BOOD: Boundary-based Out-Of-Distribution Data GenerationQilin Liao, Shuo Yang, Bo Zhao, Ping Luo 等ICML 2025
- OpenSDI: Spotting Diffusion-Generated Images in the Open WorldYabin Wang, Zhiwu Huang, Xiaopeng HongCVPR 2025
- Dream the Impossible: Outlier Imagination with Diffusion ModelsXuefeng Du, Yiyou Sun, Jerry Zhu, Yixuan LiNeurIPS 2023 · 被引用 114 次
- ID-like Prompt Learning for Few-Shot Out-of-Distribution DetectionYichen Bai, Zongbo Han, Bing Cao, Xiaoheng Jiang 等CVPR 2024
