Diversity vs. Recognizability: Human-like generalization in one-shot generative models
Victor Boutin, Lakshya Singhal, Xavier Thomas, Thomas Serre
Abstract
Robust generalization to new concepts has long remained a distinctive feature of human intelligence. However, recent progress in deep generative models has now led to neural architectures capable of synthesizing novel instances of unknown visual concepts from a single training example. Yet, a more precise comparison between these models and humans is not possible because existing performance metrics for generative models (i.e., FID, IS, likelihood) are not appropriate for the one-shot generation scenario. Here, we propose a new framework to evaluate one-shot generative models along two axes: sample recognizability vs. diversity (i.e., intra-class variability). Using this framework, we perform a systematic evaluation of representative one-shot generative models on the Omniglot handwritten dataset. We first show that GAN-like and VAE-like models fall on opposite ends of the diversity-recognizability space. Extensive analyses of the effect of key model parameters further revealed that spatial attention and context integration have a linear contribution to the diversity-recognizability trade-off. In contrast, disentanglement transports the model along a parabolic curve that could be used to maximize recognizability. Using the diversity-recognizability framework, we were able to identify models and parameters that closely approximate human data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 09a45060-255a-4a2c-b339-50d61c855be5Cited by top-tier papers3
- Diffusion Models as Artists: Are we Closing the Gap between Humans and Machines?Victor Boutin, Thomas Fel, Lakshya Singhal, Rishav Mukherji et al.ICML 2023 · 13 citations
- Follow the Energy, Find the Path: Riemannian Metrics from Energy-Based ModelsLouis Béthune, David Vigouroux, Yilun Du, Rufin VanRullen et al.NeurIPS 2025 · 8 citations
- Latent Representation Matters: Human-like Sketches in One-shot Drawing TasksVictor Boutin, Rishav Mukherji, Aditya Agrawal, Sabine Muzellec et al.NeurIPS 2024 · 3 citations
Builds on2
Related papers
- Rarity Score : A New Metric to Evaluate the Uncommonness of Synthesized ImagesJiyeon Han, Hwanil Choi, Yunjey Choi, Junho Kim et al.ICLR 2023 · 7 citations
- Feature Likelihood Score: Evaluating the Generalization of Generative Models Using SamplesMarco Jiralerspong, Avishek Joey Bose, Ian Gemp, Chongli Qin et al.NeurIPS 2023 · 39 citations
- Learning to Memorize Feature Hallucination for One-Shot Image GenerationYu Xie, Yanwei Fu, Ying Tai, Yun Cao et al.CVPR 2022 · 10 citations
- Enhancing Identity-Deformation Disentanglement in StyleGAN for One-Shot Face Video Re-EnactmentQing Chang, Yao-Xiang Ding, Kun ZhouAAAI 2025 · 3 citations
- Lost in Latent Space: Examining failures of disentangled models at combinatorial generalisationMilton Llera Montero, Jeffrey S. Bowers, Rui Ponte Costa, Casimir J. H. Ludwig et al.NeurIPS 2022 · 29 citations
