Revealing the Invisible: Latent Structure Modeling for Semantically Consistent Cloud Removal
Jingwei Xin, Kai Guo, Jie Li, Nannan Wang
摘要
Cloud removal (CR) in remote sensing imagery is a critical yet challenging task due to complex cloud patterns and diverse underlying ground structures. Despite recent progress in generative models such as diffusion models, CR remains limited by their inadequate capability to perceive and reconstruct structured information beneath cloud-covered areas. In this work, we propose a Visibility-guided Semantic Estimation and Reconstruction network for cloud removal (VISER-CR), which reformulates CR as a structure-guided completion problem. Specifically, VISER-CR explicitly models cloud interference via spatial masking, encouraging the model to reason beyond pixel-level appearance and enhance scene-level structural understanding. Moreover, to further improve the representation of structural information, we introduce Patch Saliency Encoding, a self-guided mechanism that implicitly models structural alignment among patches, significantly enhancing clustering consistency and semantic separability in the latent space. This adaptive mechanism guides the network to focus on learning and reconstructing structurally important regions, thereby reducing redundancy and improving overall cloud removal performance. Extensive experiments on multiple benchmark datasets demonstrate the superior effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat 等CVPR 2022 · 被引用 3,348 次
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 被引用 522 次
- Self-Guided Masked AutoencoderJeongwoo Shin, Inseo Lee, Junho Lee, Joonseok LeeNeurIPS 2024 · 被引用 18 次
- Effective Cloud Removal for Remote Sensing Images by an Improved Mean-Reverting Denoising Model with Elucidated Design SpaceYi Liu, Wengen Li, Jihong Guan, Shuigeng Zhou 等CVPR 2025
- MixMAE: Mixed and Masked Autoencoder for Efficient Pretraining of Hierarchical Vision TransformersJihao Liu, Xin Huang, Jinliang Zheng, Yu Liu 等CVPR 2023
相关 Paper
- Saliency-Guided Adaptive Random Diffusion for Remote Sensing Images Restoration with Cloud and HazeWanting Zhang, Jingxuan Zhang, Libao ZhangACM MM 2025 · 被引用 1 次
- Generative Adversarial Training for Weakly Supervised Cloud MattingZhengxia Zou, Wenyuan Li, Tianyang Shi, Zhenwei Shi 等ICCV 2019 · 被引用 29 次
- MDFL: Multi-Domain Diffusion-Driven Feature LearningDaixun Li, Weiying Xie, Jiaqing Zhang, Yunsong LiAAAI 2024 · 被引用 19 次
- Self-Learning Video Rain Streak Removal: When Cyclic Consistency Meets Temporal CorrespondenceWenhan Yang, Robby T. Tan, Shiqi Wang, Jiaying LiuCVPR 2020
- Vis2Mesh: Efficient Mesh Reconstruction from Unstructured Point Clouds of Large Scenes with Learned Virtual View VisibilityShuang Song, Zhaopeng Cui, Rongjun QinICCV 2021 · 被引用 13 次
