Intriguing Properties of Generative Classifiers
Priyank Jaini, Kevin Clark, Robert Geirhos
摘要
What is the best paradigm to recognize objects -- discriminative inference (fast but potentially prone to shortcut learning) or using a generative model (slow but potentially more robust)? We build on recent advances in generative modeling that turn text-to-image models into classifiers. This allows us to study their behavior and to compare them against discriminative models and human psychophysical data. We report four intriguing emergent properties of generative classifiers: they show a record-breaking human-like shape bias (99% for Imagen), near human-level out-of-distribution accuracy, state-of-the-art alignment with human classification errors, and they understand certain perceptual illusions. Our results indicate that while the current dominant paradigm for modeling human object recognition is discriminative inference, zero-shot generative models approximate human object recognition data surprisingly well.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Adversarial Robustness Limits via Scaling-Law and Human-Alignment StudiesBrian R. Bartoldson, James Diffenderfer, Konstantinos Parasyris, Bhavya KailkhuraICML 2024 · 被引用 45 次
- Visual Anagrams Reveal Hidden Differences in Holistic Shape Processing Across Vision ModelsFenil R. Doshi, Thomas Fel, Talia Konkle, George A. AlvarezNeurIPS 2025 · 被引用 5 次
- Your VAR Model is Secretly an Efficient and Explainable Generative ClassifierYi-Chung Chen, David I. Inouye, Jing GaoICLR 2026 · 被引用 2 次
- Exploring Structured Semantic Priors Underlying Diffusion Score for Test-time AdaptationMingjia Li, Shuang Li, Tongrui Su, Longhui Yuan 等NeurIPS 2024 · 被引用 2 次
- Leveraging Prior Knowledge of Diffusion Model for Person SearchGiyeol Kim, Sooyoung Yang, Jihyong Oh, Myungjoo Kang 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper15
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Text-to-Image Diffusion Models are Zero Shot ClassifiersKevin Clark, Priyank JainiNeurIPS 2023 · 被引用 192 次
- A Universal Discriminator for Zero-Shot GeneralizationHaike Xu, Zongyu Lin, Jing Zhou, Yanan Zheng 等ACL 2023 · 被引用 6 次
- Compositional Scene Understanding through Inverse Generative ModelingYanbo Wang, Justin Dauwels, Yilun DuICML 2025
- Generative Multi-modal Models are Good Class-Incremental LearnersXusheng Cao, Haori Lu, Linlan Huang, Xialei Liu 等CVPR 2024
- Is Synthetic Data from Generative Models Ready for Image Recognition?Ruifei He, Shuyang Sun, Xin Yu, Chuhui Xue 等ICLR 2023 · 被引用 56 次
