Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
Sol Park, Soobin Um
摘要
Minority sampling aims to generate low-density instances on a data manifold and is of central importance in applications such as medical diagnosis, anomaly detection, and creative AI. Existing approaches, however, define minority samples relative to generative priors learned from training data, confining rarity to model-specific notions that may poorly reflect real-world semantics. In this work, we propose a world-centric perspective on minority sampling, which defines rarity with respect to real-world priors rather than generator-induced densities. To this end, we introduce JEPA guidance, a diffusion sampling framework guided by a Joint-Embedding Predictive Architecture (JEPA)-a class of world models that encode broad, semantically rich representations. JEPA guidance steers diffusion trajectories toward low-density regions under the implicit density induced by the JEPA, thereby aligning generated minorities with real-world semantic rarity. To make JEPA guidance computationally practical, we develop principled approximation strategies accompanied by theoretical error bounds, significantly reducing the overhead of guidance computation. Extensive experiments across unconditional, class-conditional, and text-to-image generation demonstrate that JEPA guidance consistently improves the fidelity and semantic validity of minority samples, outperforming generator-centric baselines in capturing real-world notions of rarity. Code is available at https://github.com/ soobin-um/jepa-guidance . †Corresponding author.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- ImageReward: Learning and Evaluating Human Preferences for Text-to-Image GenerationJiazheng Xu, Xiao Liu, Yuchen Wu, Yuxuan Tong 等NeurIPS 2023 · 被引用 1,310 次
相关 Paper
- Don't Play Favorites: Minority Guidance for Diffusion ModelsSoobin Um, Suhyeon Lee, Jong Chul YeICLR 2024 · 被引用 37 次
- Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority GenerationSoobin Um, Beomsu Kim, Jong Chul YeICML 2025
- Minority-Focused Text-to-Image Generation via Prompt OptimizationSoobin Um, Jong Chul YeCVPR 2025
- RAIGen: Rare Attribute Identification in Text-to-Image Generative ModelsSilpa Vadakkeeveetil Sreelatha, Dan Wang, Serge Belongie, Muhammad Awais 等ICML 2026
- GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution DetectionXin Gao, Jiyao Liu, Guanghao Li, Yueming Lyu 等NeurIPS 2025 · 被引用 9 次
