RDF-MIG: A Robust Diffusion Framework for Masked Image Generation to Augment Semantic Segmentation and Change Detection
Zian Cao, Wei Wei, Qingshan Gao, Yuanyuan Fu
Abstract
Change detection and semantic segmentation are key techniques for satellite image analysis in remote sensing. However, acquiring high-quality labeled data is costly and timeconsuming. Although recent studies have explored generative models to ease data scarcity, a unified framework supporting both tasks is still lacking, and most methods overlook noise accumulation and cannot generate multispectral images. To address this, we propose the robust diffusion framework for masked image generation (RDF-MIG). RDF-MIG generates bi-temporal change-labeled and single-temporal segmentation-labeled images to enhance downstream change detection and semantic segmentation tasks. Furthermore, to address noise accumulation and improve the quality of generated image-mask pairs, we reformulate the diffusion model training objective by proposing the Maximum Correntropy Robust Diffusion (MCRD) loss, and further design an MSE-consistency calibration that analytically aligns small-error gradients with the MSE objective while preserving robustness to outliers. Experiments indicate that the proposed RDF-MIG framework can generate multispectral image-mask pairs to improve downstream performance, while MCRD loss further enhances the quality of the synthesized data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 32adee0b-e7d8-4d0b-9ef8-1face9ebc978Builds on13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo et al.ICCV 2019 · 1,125 citations
- Effective Data Augmentation With Diffusion ModelsBrandon Trabucco, Kyle Doherty, Max Gurinas, Ruslan SalakhutdinovICLR 2024 · 380 citations
- Generating images with sparse representationsCharlie Nash, Jacob Menick, Sander Dieleman, Peter W. BattagliaICML 2021 · 291 citations
Related papers
- SatSynth: Augmenting Image-Mask Pairs Through Diffusion Models for Aerial Semantic SegmentationAysim Toker, Marvin Eisenberger, Daniel Cremers, Laura Leal-TaixéCVPR 2024 · 36 citations
- Task-Oriented Data Synthesis and Control-Rectify Sampling for Remote Sensing Semantic SegmentationYunkai Yang, Yudong Zhang, Kunquan Zhang, Jinxiao Zhang et al.CVPR 2026 · 2 citations
- Stochastic Conditional Diffusion Models for Robust Semantic Image SynthesisJuyeon Ko, Inho Kong, Dogyun Park, Hyunwoo J. KimICML 2024 · 14 citations
- ChangeDiff: A Multi-Temporal Change Detection Data Generator with Flexible Text Prompts via Diffusion ModelQi Zang, Jiayi Yang, Shuang Wang, Dong Zhao et al.AAAI 2025 · 2 citations
- JoDiffusion: Jointly Diffusing Image with Pixel-Level Annotations for Semantic Segmentation PromotionHaoyu Wang, Lei Zhang, Wenrui Liu, Dengyang Jiang et al.AAAI 2026
