RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo
Victor Oei, Jenny Schmalfuss, Lukas Mehl, Madlen Bartsch, Shashank Agnihotri, Margret Keuper, Andreas Bulling, Andrés Bruhn
Abstract
Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image corruptions like noise or rain. Hence, the resilience of models to such real-world perturbations is largely unquantified. To address this, we present RobustSpring, a comprehensive dataset and benchmark for evaluating robustness to image corruptions for optical flow, scene flow, and stereo models. RobustSpring applies 20 different image corruptions, including noise, blur, color changes, quality degradations, and weather distortions, in a time-, stereo-, and depth-consistent manner to the high-resolution Spring dataset, creating a suite of 20,000 corrupted images that reflect challenging conditions. RobustSpring enables comparisons of model robustness via a new corruption robustness metric. Integration with the Spring benchmark enables two-axis evaluations of both accuracy and robustness. We benchmark a curated selection of initial models, observing that robustness varies widely by corruption type, and experimentally show that evaluations on RobustSpring indicate real-world robustness. RobustSpring is a new computer vision benchmark to treat robustness as a first-class citizen, fostering models that are accurate and resilient. It is available at https://spring-benchmark.org.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 45693b76-465f-49e4-88ee-e9602181cd2dCited by top-tier papers2
- PARC: A Quantitative Framework Uncovering the Symmetries within Vision Language ModelsJenny Schmalfuss, Nadine Chang, Vibashan VS, Maying Shen et al.CVPR 2025
- Instant Video Models: Universal Adapters for Stabilizing Image-Based NetworksMatthew Dutson, Nathan Labiosa, Yin Li, Mohit GuptaNeurIPS 2025
Builds on20
- Measuring Robustness to Natural Distribution Shifts in Image ClassificationRohan Taori, Achal Dave, Vaishaal Shankar, Nicholas Carlini et al.NeurIPS 2020 · 731 citations
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai et al.NeurIPS 2020 · 436 citations
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li et al.ICCV 2021 · 402 citations
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi et al.CVPR 2022 · 353 citations
- Attention Concatenation Volume for Accurate and Efficient Stereo MatchingGangwei Xu, Junda Cheng, Peng Guo, Xin YangCVPR 2022 · 265 citations
Related papers
- Spring: A High-Resolution High-Detail Dataset and Benchmark for Scene Flow, Optical Flow and StereoLukas Mehl, Jenny Schmalfuss, Azin Jahedi, Yaroslava Nalivayko et al.CVPR 2023
- Robo3D: Towards Robust and Reliable 3D Perception against CorruptionsLingdong Kong, Youquan Liu, Xin Li, Runnan Chen et al.ICCV 2023 · 151 citations
- VLM-RobustBench: A Comprehensive Benchmark for Robustness of Vision-Language ModelsRohit Saxena, Alessandro Suglia, Pasquale MinerviniICML 2026 · 7 citations
- RCP-Bench: Benchmarking Robustness for Collaborative Perception Under Diverse CorruptionsShihang Du, Sanqing Qu, Tianhang Wang, Xudong Zhang et al.CVPR 2025
- Evaluating Robustness of Monocular Depth Estimation with Procedural Scene PerturbationsJack Nugent, Siyang Wu, Zeyu Ma, Beining Han et al.NeurIPS 2025 · 6 citations
