Open Challenges in Deep Stereo: the Booster Dataset
Pierluigi Zama Ramirez, Fabio Tosi, Matteo Poggi, Samuele Salti, Stefano Mattoccia, Luigi Di Stefano
Abstract
We present a novel high-resolution and challenging stereo dataset framing indoor scenes annotated with dense and accurate ground-truth disparities. Peculiar to our dataset is the presence of several specular and transparent surfaces, i.e. the main causes of failures for state-of-the-art stereo networks. Our acquisition pipeline leverages a novel deep space-time stereo framework which allows for easy and accurate labeling with sub-pixel precision. We re-lease a total of 419 samples collected in 64 different scenes and annotated with dense ground-truth disparities. Each sample include a high-resolution pair (12 Mpx) as well as an unbalanced pair (Left: 12 Mpx, Right: 1.1 Mpx). Additionally, we provide manually annotated material segmentation masks and 15K unlabeled samples. We evaluate state-of-the-art deep networks based on our dataset, highlighting their limitations in addressing the open challenges in stereo and drawing hints for future research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 779983cf-58c4-479c-8ff7-1dea04ea2312Cited by top-tier papers22
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
- CroCo v2: Improved Cross-view Completion Pre-training for Stereo Matching and Optical FlowPhilippe Weinzaepfel, Thomas Lucas, Vincent Leroy, Yohann Cabon et al.ICCV 2023 · 181 citations
- Learning Depth Estimation for Transparent and Mirror SurfacesAlex Costanzino, Pierluigi Zama Ramirez, Matteo Poggi, Fabio Tosi et al.ICCV 2023 · 41 citations
- Parameterized Cost Volume for Stereo MatchingJiaxi Zeng, Chengtang Yao, Lidong Yu, Yuwei Wu et al.ICCV 2023 · 35 citations
- Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface DetectionJiaying Lin, Yuen Hei Yeung, Shuquan Ye, Rynson W. H. LauAAAI 2025 · 15 citations
Builds on6
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai et al.NeurIPS 2020 · 436 citations
- On the Over-Smoothing Problem of CNN Based Disparity EstimationChuangrong Chen, Xiaozhi Chen, Hui ChengICCV 2019 · 24 citations
- CFNet: Cascade and Fused Cost Volume for Robust Stereo MatchingZhelun Shen, Yuchao Dai, Zhibo RaoCVPR 2021
- SMD-Nets: Stereo Mixture Density NetworksFabio Tosi, Yiyi Liao, Carolin Schmitt, Andreas GeigerCVPR 2021
Related papers
- StereOBJ-1M: Large-scale Stereo Image Dataset for 6D Object Pose EstimationXingyu Liu, Shun Iwase, Kris M. KitaniICCV 2021 · 58 citations
- RGB-Multispectral Matching: Dataset, Learning Methodology, EvaluationFabio Tosi, Pierluigi Zama Ramirez, Matteo Poggi, Samuele Salti et al.CVPR 2022 · 5 citations
- DiLiGenT102: A Photometric Stereo Benchmark Dataset with Controlled Shape and Material VariationJieji Ren, Feishi Wang, Jiahao Zhang, Qian Zheng et al.CVPR 2022 · 26 citations
- Practical Stereo Matching via Cascaded Recurrent Network with Adaptive CorrelationJiankun Li, Peisen Wang, Pengfei Xiong, Tao Cai et al.CVPR 2022 · 294 citations
- BlendedMVS: A Large-Scale Dataset for Generalized Multi-View Stereo NetworksYao Yao, Zixin Luo, Shiwei Li, Jingyang Zhang et al.CVPR 2020
