GraftNet: Towards Domain Generalized Stereo Matching with a Broad-Spectrum and Task-Oriented Feature
Biyang Liu, Huimin Yu, Guodong Qi
Abstract
Although supervised deep stereo matching networks have made impressive achievements, the poor generalization ability caused by the domain gap prevents them from being applied to real-life scenarios. In this paper, we propose to leverage the feature of a model trained on large-scale datasets to deal with the domain shift since it has seen various styles of images. With the cosine similarity based cost volume as a bridge, the feature will be grafted to an ordinary cost aggregation module. Despite the broad-spectrum representation, such a low-level feature contains much general information which is not aimed at stereo matching. To recover more task-specific information, the grafted feature is further input into a shallow network to be transformed before calculating the cost. Extensive experiments show that the model generalization ability can be improved significantly with this broad-spectrum and task-oriented feature. Specifically, based on two well-known architectures PSMNet and GANet, our methods are superior to other robust algorithms when transferring from SceneFlow to KITTI 2015, KITTI 2012, and Middlebury. Code is available at https://github.com/SpadeLiu/Graft-PSMNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a6e4ccfd-53eb-4ea1-9c77-eb26f316131fCited by top-tier papers22
- Fast-FoundationStereo: Real-Time Zero-Shot Stereo MatchingBowen Wen, Shaurya Dewan, Stan BirchfieldCVPR 2026 · 36 citations
- Adaptive Multi-Modal Cross-Entropy Loss for Stereo MatchingPeng Xu, Zhiyu Xiang, Chengyu Qiao, Jingyun Fu et al.CVPR 2024 · 28 citations
- Unifying Feature and Cost Aggregation with Transformers for Semantic and Visual CorrespondenceSunghwan Hong, Seokju Cho, Seungryong Kim, Stephen LinICLR 2024 · 16 citations
- Active Stereo Without Pattern ProjectorLuca Bartolomei, Matteo Poggi, Fabio Tosi, Andrea Conti et al.ICCV 2023 · 12 citations
- Federated Online Adaptation for Deep StereoMatteo Poggi, Fabio TosiCVPR 2024 · 11 citations
Builds on9
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai et al.NeurIPS 2020 · 436 citations
- Revisiting Stereo Depth Estimation From a Sequence-to-Sequence Perspective with TransformersZhaoshuo Li, Xingtong Liu, Nathan Drenkow, Andy S. Ding et al.ICCV 2021 · 380 citations
- Learning Across Tasks and DomainsPierluigi Zama Ramirez, Alessio Tonioni, Samuele Salti, Luigi Di StefanoICCV 2019 · 32 citations
- AdaStereo: A Simple and Efficient Approach for Adaptive Stereo MatchingXiao Song, Guorun Yang, Xinge Zhu, Hui Zhou et al.CVPR 2021
Related papers
- Domain Generalized Stereo Matching via Hierarchical Visual TransformationTianyu Chang, Xun Yang, Tianzhu Zhang, Meng WangCVPR 2023
- CFNet: Cascade and Fused Cost Volume for Robust Stereo MatchingZhelun Shen, Yuchao Dai, Zhibo RaoCVPR 2021
- StereoGAN: Bridging Synthetic-to-Real Domain Gap by Joint Optimization of Domain Translation and Stereo MatchingRui Liu, Chengxi Yang, Wenxiu Sun, Xiaogang Wang et al.CVPR 2020
- Local Similarity Pattern and Cost Self-Reassembling for Deep Stereo Matching NetworksBiyang Liu, Huimin Yu, Yangqi LongAAAI 2022 · 86 citations
- Revisiting Non-Parametric Matching Cost Volumes for Robust and Generalizable Stereo MatchingKelvin Cheng, Tianfu Wu, Christopher G. HealeyNeurIPS 2022 · 5 citations
