Revisiting Non-Parametric Matching Cost Volumes for Robust and Generalizable Stereo Matching
Kelvin Cheng, Tianfu Wu, Christopher G. Healey
Abstract
Stereo matching is a classic challenging problem in computer vision, which has recently witnessed remarkable progress by Deep Neural Networks (DNNs). This paradigm shift leads to two interesting and entangled questions that have not been addressed well. First , it is unclear whether stereo matching DNNs that are trained from scratch really learn to perform matching well. This paper studies this problem from the lens of white-box adversarial attacks. It presents a method of learning stereo-constrained photometrically-consistent attacks, which by design are weaker adversarial attacks, and yet can cause catastrophic performance drop for those DNNs. This observation suggests that they may not actually learn to perform matching well in the sense that they should otherwise achieve potentially even better after stereo-constrained perturbations are introduced. Second , stereo matching DNNs are typically trained under the simulation-to-real (Sim2Real) pipeline due to the data hungriness of DNNs. Thus, alleviating the impacts of the Sim2Real photometric gap in stereo matching DNNs becomes a pressing need. Towards joint adversarially robust and domain generalizable stereo matching, this paper proposes to learn DNN-contextualized binary-pattern-driven non-parametric cost-volumes . It leverages the perspective of learning the cost aggregation via DNNs, and presents a simple yet expressive design that is fully end-to-end trainable, without resorting to specific aggregation inductive biases. In experiments, the proposed method is tested in the SceneFlow dataset, the KITTI2015 dataset, and the Middlebury dataset. It significantly improves the adversarial robustness, while retaining accuracy performance comparable to state-of-the-art methods. It also shows a better Sim2Real generalizability. Our code and pretrained models are released at this Github Repo.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7b3ff7da-f841-43c3-bc59-d8f58dd41287Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai et al.NeurIPS 2020 · 436 citations
- Attacking Optical FlowAnurag Ranjan, Joel Janai, Andreas Geiger, Michael J. BlackICCV 2019 · 93 citations
- Stereopagnosia: Fooling Stereo Networks with Adversarial PerturbationsAlex Wong, Mukund Mundhra, Stefano SoattoAAAI 2021 · 33 citations
- CFNet: Cascade and Fused Cost Volume for Robust Stereo MatchingZhelun Shen, Yuchao Dai, Zhibo RaoCVPR 2021
- Physically Realizable Adversarial Examples for LiDAR Object DetectionJames Tu, Mengye Ren, Sivabalan Manivasagam, Ming Liang et al.CVPR 2020
Related papers
- Adaptive Unimodal Cost Volume Filtering for Deep Stereo MatchingYoumin Zhang, Yimin Chen, Xiao Bai, Suihanjin Yu et al.AAAI 2020 · 201 citations
- AdaStereo: A Simple and Efficient Approach for Adaptive Stereo MatchingXiao Song, Guorun Yang, Xinge Zhu, Hui Zhou et al.CVPR 2021
- GraftNet: Towards Domain Generalized Stereo Matching with a Broad-Spectrum and Task-Oriented FeatureBiyang Liu, Huimin Yu, Guodong QiCVPR 2022 · 52 citations
- Semantic Stereo Matching With Pyramid Cost VolumesZhenyao Wu, Xinyi Wu, Xiaoping Zhang, Song Wang et al.ICCV 2019 · 125 citations
- Patchmatch Stereo++: Patchmatch Binocular Stereo with Continuous Disparity OptimizationWenjia Ren, Qingmin Liao, Zhijing Shao, Xiangru Lin et al.ACM MM 2023 · 5 citations
