DetDiffusion: Synergizing Generative and Perceptive Models for Enhanced Data Generation and Perception
Yibo Wang, Ruiyuan Gao, Kai Chen, Kaiqiang Zhou, Yingjie Cai, Lanqing Hong, Zhenguo Li, Lihui Jiang, Dit-Yan Yeung, Qiang Xu, Kai Zhang
摘要
Current perceptive models heavily depend on resource-intensive datasets, prompting the need for innovative solutions. Leveraging recent advances in diffusion models, synthetic data, by constructing image inputs from various annotations, proves beneficial for downstream tasks. While prior methods have separately addressed generative and perceptive models, DetDiffusion, for the first time, harmonizes both, tackling the challenges in generating effective data for perceptive models. To enhance image generation with perceptive models, we introduce perception-aware loss (P.A. loss) through segmentation, improving both quality and controllability. To boost the performance of specific perceptive models, our method customizes data augmentation by extracting and utilizing perception-aware attribute (P.A. Attr) during generation. Experimental results from the object detection task highlight DetDiffusion's superior performance, establishing a new state-of-the-art in layout-guided generation. Furthermore, image syntheses from DetDiffusion can effectively augment training data, significantly enhancing downstream detection performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- SubjectDrive: Scaling Generative Data in Autonomous Driving via Subject ControlBinyuan Huang, Yuqing Wen, Yucheng Zhao, Yaosi Hu 等AAAI 2025 · 被引用 28 次
- Advancing Fine-Grained Classification by Structure and Subject Preserving AugmentationEyal Michaeli, Ohad FriedNeurIPS 2024 · 被引用 19 次
- Mixture of insighTful Experts (MoTE): The Synergy of Reasoning Chains and Expert Mixtures in Self-AlignmentZhili Liu, Yunhao Gou, Kai Chen, Lanqing Hong 等ACL 2025 · 被引用 13 次
- MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive ControlRuiyuan Gao, Kai Chen, Bo Xiao, Lanqing Hong 等ICCV 2025 · 被引用 11 次
- Unbiased Object Detection Beyond Frequency with Visually Prompted Image SynthesisXinhao Cai, Liulei Li, Gensheng Pei, Tao Chen 等ICLR 2026 · 被引用 7 次
它引用的顶会 Paper33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- ReCon: Region-Controllable Data Augmentation with Rectification and Alignment for Object DetectionHaowei Zhu, Tianxiang Pan, Rui Qin, Jun-Hai Yong 等NeurIPS 2025 · 被引用 5 次
- Cycle-Consistent Learning for Joint Layout-to-Image Generation and Object DetectionXinhao Cai, Qiuxia Lai, Gensheng Pei, Xiangbo Shu 等ICCV 2025 · 被引用 1 次
- GeoDiffusion: Text-Prompted Geometric Control for Object Detection Data GenerationKai Chen, Enze Xie, Zhe Chen, Yibo Wang 等ICLR 2024 · 被引用 60 次
- ChangeDiff: A Multi-Temporal Change Detection Data Generator with Flexible Text Prompts via Diffusion ModelQi Zang, Jiayi Yang, Shuang Wang, Dong Zhao 等AAAI 2025 · 被引用 2 次
- DatasetDM: Synthesizing Data with Perception Annotations Using Diffusion ModelsWeijia Wu, Yuzhong Zhao, Hao Chen, Yuchao Gu 等NeurIPS 2023 · 被引用 191 次
