Revealing the Invisible: Latent Structure Modeling for Semantically Consistent Cloud Removal
Jingwei Xin, Kai Guo, Jie Li, Nannan Wang
Abstract
Cloud removal (CR) in remote sensing imagery is a critical yet challenging task due to complex cloud patterns and diverse underlying ground structures. Despite recent progress in generative models such as diffusion models, CR remains limited by their inadequate capability to perceive and reconstruct structured information beneath cloud-covered areas. In this work, we propose a Visibility-guided Semantic Estimation and Reconstruction network for cloud removal (VISER-CR), which reformulates CR as a structure-guided completion problem. Specifically, VISER-CR explicitly models cloud interference via spatial masking, encouraging the model to reason beyond pixel-level appearance and enhance scene-level structural understanding. Moreover, to further improve the representation of structural information, we introduce Patch Saliency Encoding, a self-guided mechanism that implicitly models structural alignment among patches, significantly enhancing clustering consistency and semantic separability in the latent space. This adaptive mechanism guides the network to focus on learning and reconstructing structurally important regions, thereby reducing redundancy and improving overall cloud removal performance. Extensive experiments on multiple benchmark datasets demonstrate the superior effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 244ef77a-a878-4bb5-8bf1-99ae6c227a44Builds on6
- Restormer: Efficient Transformer for High-Resolution Image RestorationSyed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat et al.CVPR 2022 · 3,348 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Self-Guided Masked AutoencoderJeongwoo Shin, Inseo Lee, Junho Lee, Joonseok LeeNeurIPS 2024 · 18 citations
- Effective Cloud Removal for Remote Sensing Images by an Improved Mean-Reverting Denoising Model with Elucidated Design SpaceYi Liu, Wengen Li, Jihong Guan, Shuigeng Zhou et al.CVPR 2025
- MixMAE: Mixed and Masked Autoencoder for Efficient Pretraining of Hierarchical Vision TransformersJihao Liu, Xin Huang, Jinliang Zheng, Yu Liu et al.CVPR 2023
Related papers
- Saliency-Guided Adaptive Random Diffusion for Remote Sensing Images Restoration with Cloud and HazeWanting Zhang, Jingxuan Zhang, Libao ZhangACM MM 2025 · 1 citation
- Generative Adversarial Training for Weakly Supervised Cloud MattingZhengxia Zou, Wenyuan Li, Tianyang Shi, Zhenwei Shi et al.ICCV 2019 · 29 citations
- MDFL: Multi-Domain Diffusion-Driven Feature LearningDaixun Li, Weiying Xie, Jiaqing Zhang, Yunsong LiAAAI 2024 · 19 citations
- Self-Learning Video Rain Streak Removal: When Cyclic Consistency Meets Temporal CorrespondenceWenhan Yang, Robby T. Tan, Shiqi Wang, Jiaying LiuCVPR 2020
- Vis2Mesh: Efficient Mesh Reconstruction from Unstructured Point Clouds of Large Scenes with Learned Virtual View VisibilityShuang Song, Zhaopeng Cui, Rongjun QinICCV 2021 · 13 citations
