SDDGR: Stable Diffusion-Based Deep Generative Replay for Class Incremental Object Detection
Junsu Kim, Hoseong Cho, Jihyeon Kim, Yihalem Yimolal Tiruneh, Seungryul Baek
Abstract
In the field of class incremental learning (CIL), generative replay has become increasingly prominent as a method to mitigate the catastrophic forgetting, alongside the continuous improvements in generative models. However, its application in class incremental object detection (CIOD) has been significantly limited, primarily due to the complexities of scenes involving multiple labels. In this paper, we propose a novel approach called stable diffusion deep generative replay (SDDGR) for CIOD. Our method utilizes a diffusion-based generative model with pre-trained text-to-image diffusion networks to generate realistic and diverse synthetic images. SDDGR incorporates an iterative refinement strategy to produce high-quality images encompassing old classes. Additionally, we adopt an L2 knowledge distillation technique to improve the retention of prior knowledge in synthetic images. Furthermore, our approach includes pseudo-labeling for old objects within new task images, preventing misclassification as background elements. Extensive experiments on the COCO 2017 dataset demonstrate that SD-DGR significantly outperforms existing algorithms, achieving a new state-of-the-art in various CIOD scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 87dbebe6-b103-4c75-8f12-f41c96c782d6Cited by top-tier papers24
- When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data GeneratorsKrzysztof Adamkiewicz, Brian B. Moser, Stanislav Frolov, Tobias Christian Nauen et al.CVPR 2026 · 8 citations
- Task-Free Continual Generation and Representation Learning via Dynamic Expansionable Memory ClusterFei Ye, Adrian G. BorsAAAI 2024 · 8 citations
- BlurDM: A Blur Diffusion Model for Image DeblurringJin-Ting He, Fu-Jen Tsai, Yan-Tsung Peng, Min-Hung Chen et al.NeurIPS 2025 · 7 citations
- Enhanced Continual Learning of Vision-Language Models with Model FusionHaoyuan Gao, Zicong Zhang, Yuqi Wei, Linglan Zhao et al.ICLR 2026 · 5 citations
- GCD: Advancing Vision-Language Models for Incremental Object Detection via Global Alignment and Correspondence DistillationXu Wang, Zilei Wang, Zihan LinAAAI 2025 · 4 citations
Builds on28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
Related papers
- Revisiting Generative Replay for Class Incremental Object DetectionShizhou Zhang, Xueqiang Lv, Yinghui Xing, Qirui Wu et al.CVPR 2025
- DDGR: Continual Learning with Deep Diffusion-based Generative ReplayRui Gao, Weiwei LiuICML 2023 · 101 citations
- Pseudo Object Replay and Mining for Incremental Object DetectionDongbao Yang, Yu Zhou, Xiaopeng Hong, Aoting Zhang et al.ACM MM 2023 · 6 citations
- Gradient Decomposition and Alignment for Incremental Object DetectionWenlong Luo, Shizhou Zhang, De Cheng, Yinghui Xing et al.ICCV 2025 · 5 citations
- Learning Task-Aware Language-Image Representation for Class-Incremental Object DetectionHongquan Zhang, Bin-Bin Gao, Yi Zeng, Xudong Tian et al.AAAI 2024 · 12 citations
