DiffuLT: Diffusion for Long-tail Recognition Without External Knowledge
Jie Shao, Ke Zhu, Hanxiao Zhang, Jianxin Wu
摘要
This paper introduces a novel pipeline for long-tail (LT) recognition that diverges from conventional strategies. Instead, it leverages the long-tailed dataset itself to generate a balanced proxy dataset without utilizing external data or model. We deploy a diffusion model trained from scratch on only the long-tailed dataset to create this proxy and verify the effectiveness of the data produced. Our analysis identifies approximately-in-distribution (AID) samples, which slightly deviate from the real data distribution and incorporate a blend of class information, as the crucial samples for enhancing the generative model's performance in long-tail classification. We promote the generation of AID samples during the training of a generative model by utilizing a feature extractor to guide the process and filter out detrimental samples during generation. Our approach, termed Diffusion model for Long-Tail recognition (DiffuLT), represents a pioneer application of generative models in long-tail recognition. DiffuLT achieves state-of-the-art results on CIFAR10-LT, CIFAR100-LT, and ImageNet-LT, surpassing leading competitors by significant margins. Comprehensive ablations enhance the interpretability of our pipeline. Notably, the entire generative process is conducted without relying on external data or pre-trained model weights, which leads to its generalizability to real-world long-tailed scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Reframing Long-Tailed Learning via Loss Landscape GeometryShenghan Chen, Yiming Liu, Yanzhen Wang, Yujia Wang 等CVPR 2026 · 被引用 2 次
- Principled Long-Tailed Generative Modeling via Diffusion ModelsPranoy Das, Kexin Fu, Abolfazl Hashemi, Vijay GuptaNeurIPS 2025 · 被引用 2 次
- Diffusion Curriculum: Synthetic-to-Real Data Curriculum via Image-Guided DiffusionYijun Liang, Shweta Bhardwaj, Tianyi ZhouICCV 2025 · 被引用 1 次
- LT-Soups: Bridging Head and Tail Classes via Subsampled Model SoupsMasih Aminbeidokhti, Subhankar Roy, Eric Granger, Elisa Ricci 等NeurIPS 2025 · 被引用 1 次
- CORAL: Disentangling Latent Representations in Long-Tailed DiffusionEsther Rodriguez, Monica Welfert, Samuel McDowell, Nathan Stromberg 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper35
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- LTGC: Long-Tail Recognition via Leveraging LLMs-Driven Generated ContentQihao Zhao, Yalun Dai, Hao Li, Wei Hu 等CVPR 2024 · 被引用 22 次
- Class-Balancing Diffusion ModelsYiming Qin, Huangjie Zheng, Jiangchao Yao, Mingyuan Zhou 等CVPR 2023
- Generative Data Mining with Longtail-Guided DiffusionDavid S. Hayden, Mao Ye, Timur Garipov, Gregory P. Meyer 等ICML 2025
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- Self Supervision to Distillation for Long-Tailed Visual RecognitionTianhao Li, Limin Wang, Gangshan WuICCV 2021 · 被引用 122 次
