SemStereo: Semantic-Constrained Stereo Matching Network for Remote Sensing
Chen Chen, Liangjin Zhao, Yuanchun He, Yingxuan Long, Kaiqiang Chen, Zhirui Wang, Yanfeng Hu, Xian Sun
Abstract
Semantic segmentation and 3D reconstruction are two fundamental tasks in remote sensing, typically treated as separate or loosely coupled tasks. Despite attempts to integrate them into a unified network, the constraints between the two heterogeneous tasks are not explicitly modeled, since the pioneering studies either utilize a loosely coupled parallel structure or engage in only implicit interactions, failing to capture the inherent connections. In this work, we explore the connections between the two tasks and propose a new network that imposes semantic constraints on the stereo matching task, both implicitly and explicitly. Implicitly, we transform the traditional parallel structure to a new cascade structure termed Semantic-Guided Cascade structure, where the deep features enriched with semantic information are utilized for the computation of initial disparity maps, enhancing semantic guidance. Explicitly, we propose a Semantic Selective Refinement (SSR) module and a Left-Right Semantic Consistency (LRSC) module. The SSR refines the initial disparity map under the guidance of the semantic map. The LRSC ensures semantic consistency between two views via reducing the semantic divergence after transforming the semantic map from one view to the other using the disparity map. Experiments on the US3D and WHU datasets demonstrate that our method achieves state-of-the-art performance for both semantic segmentation and stereo matching.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 72321d49-ac0d-49d5-863b-b59ffb2f87c6Builds on7
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision TransformerSachin Mehta, Mohammad RastegariICLR 2022 · 2,162 citations
- Attention Concatenation Volume for Accurate and Efficient Stereo MatchingGangwei Xu, Junda Cheng, Peng Guo, Xin YangCVPR 2022 · 265 citations
- Adaptive Unimodal Cost Volume Filtering for Deep Stereo MatchingYoumin Zhang, Yimin Chen, Xiao Bai, Suihanjin Yu et al.AAAI 2020 · 201 citations
- Semantic Stereo Matching With Pyramid Cost VolumesZhenyao Wu, Xinyi Wu, Xiaoping Zhang, Song Wang et al.ICCV 2019 · 125 citations
Related papers
- Feedback Network for Mutually Boosted Stereo Image Super-Resolution and Disparity EstimationQinyan Dai, Juncheng Li, Qiaosi Yi, Faming Fang et al.ACM MM 2021 · 68 citations
- EC-MVSNet: Enhanced Cascaded Multi-View Stereo with Cross-Scale Relevance IntegrationShaoqian Wang, Jiadai Sun, Bin Fan, Qiang Wang et al.AAAI 2026
- Generalizable Novel-View Synthesis Using a Stereo CameraHaechan Lee, Wonjoon Jin, Seung-Hwan Baek, Sunghyun ChoCVPR 2024
- UASNet: Uncertainty Adaptive Sampling Network for Deep Stereo MatchingYamin Mao, Zhihua Liu, Weiming Li, Yuchao Dai et al.ICCV 2021 · 34 citations
- Local Similarity Pattern and Cost Self-Reassembling for Deep Stereo Matching NetworksBiyang Liu, Huimin Yu, Yangqi LongAAAI 2022 · 86 citations
