CrossCut: Cross-Patch Aware Interactive Segmentation for Remote Sensing Images
Zheng Lin, Nan Zhou, Yuhan Wang, Bojian Zhang
Abstract
Interactive segmentation aims to delineate a user-specified target in an image by leveraging positive and negative clicks. While effective on natural images, existing methods often fail in remote sensing scenarios, where satellite imagery is characterized by ultra-high resolution, sparse object distribution, and significant scale variation. These factors hinder accurate segmentation of fine-grained targets like roads, buildings, and aircraft. To overcome these problems, we propose CrossCut, a novel interactive segmentation framework tailored for remote sensing imagery. Unlike previous approaches that either process the entire image or treat each patch independently, CrossCut enables simultaneous segmentation across multiple patches by propagating user click information to all patches. This design allows the model to fully utilize click guidance regardless of object location, effectively resolving the challenge of inter-patch information isolation. Furthermore, CrossCut supports flexible inference by allowing segmentation results from different patch configurations to be fused, enhancing both accuracy and robustness. Extensive evaluations across multiple remote sensing datasets demonstrate that CrossCut achieves state-of-the-art performance. Quantitative results and visualizations show that CrossCut significantly advances the field of interactive segmentation for remote sensing imagery.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 18a6e8f0-9c74-49d7-9370-283af6a85b21Builds on11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- SimpleClick: Interactive Image Segmentation with Simple Vision TransformersQin Liu, Zhenlin Xu, Gedas Bertasius, Marc NiethammerICCV 2023 · 161 citations
- FocalClick: Towards Practical Interactive Image SegmentationXi Chen, Zhiyan Zhao, Yilei Zhang, Manni Duan et al.CVPR 2022 · 153 citations
Related papers
- FocusCut: Diving into a Focus View in Interactive SegmentationZheng Lin, Zheng-Peng Duan, Zhao Zhang, Chun-Le Guo et al.CVPR 2022 · 61 citations
- TETRIS: Towards Exploring the Robustness of Interactive SegmentationAndrey Moskalenko, Vlad Shakhuro, Anna Vorontsova, Anton Konushin et al.AAAI 2024 · 3 citations
- NTClick: Achieving Precise Interactive Segmentation With Noise-tolerant ClicksChenyi Zhang, Ting Liu, Xiaochao Qu, Luoqi Liu et al.CVPR 2025
- Focused and Collaborative Feedback Integration for Interactive Image SegmentationQiaoqiao Wei, Hui Zhang, Jun-Hai YongCVPR 2023
- AGILE3D: Attention Guided Interactive Multi-object 3D SegmentationYuanwen Yue, Sabarinath Mahadevan, Jonas Schult, Francis Engelmann et al.ICLR 2024 · 36 citations
