Adaptive Multi-Modal Cross-Entropy Loss for Stereo Matching
Peng Xu, Zhiyu Xiang, Chengyu Qiao, Jingyun Fu, Tianyu Pu
摘要
Despite the great success of deep learning in stereo matching, recovering accurate disparity maps is still challenging. Currently, L1 and cross-entropy are the two most widely used losses for stereo network training. Compared with the former, the latter usually performs better thanks to its probability modeling and direct supervision to the cost volume. However, how to accurately model the stereo ground-truth for cross-entropy loss remains largely under-explored. Existing works simply assume that the ground-truth distributions are uni-modal, which ignores the fact that most of the edge pixels can be multimodal. In this paper, a novel adaptive multimodal cross-entropy loss (ADL) is proposed to guide the networks to learn different distribution patterns for each pixel. Moreover, we optimize the disparity estimator to further alleviate the bleeding or mis-alignment artifacts in inference. Extensive experimental results show that our method is generic and can help classic stereo networks regain state-of-the-art performance. In particular, GANet with our method ranks 1 <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">st</sup> on both the KITTI 2015 and 2012 benchmarks among the published methods. Meanwhile, excellent synthetic-to-realistic generalization performance can be achieved by simply replacing the traditional loss with ours. Code is available at https://github.com/xxxupeng/AdL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- MDP-Omni: Parameter-Free Multimodal Depth Prior-Based Sampling for Omnidirectional Stereo MatchingEunjin Son, HyungGi Jo, Wookyong Kwon, Sang Jun LeeICCV 2025 · 被引用 1 次
- Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono FailLuca Bartolomei, Fabio Tosi, Matteo Poggi, Stefano MattocciaCVPR 2025
- DEFOM-Stereo: Depth Foundation Model Based Stereo MatchingHualie Jiang, Zhiqiang Lou, Laiyan Ding, Rui Xu 等CVPR 2025
它引用的顶会 Paper18
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima 等ICCV 2019 · 被引用 1,411 次
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai 等NeurIPS 2020 · 被引用 436 次
- Practical Stereo Matching via Cascaded Recurrent Network with Adaptive CorrelationJiankun Li, Peisen Wang, Pengfei Xiong, Tao Cai 等CVPR 2022 · 被引用 294 次
- Attention Concatenation Volume for Accurate and Efficient Stereo MatchingGangwei Xu, Junda Cheng, Peng Guo, Xin YangCVPR 2022 · 被引用 265 次
- Adaptive Unimodal Cost Volume Filtering for Deep Stereo MatchingYoumin Zhang, Yimin Chen, Xiao Bai, Suihanjin Yu 等AAAI 2020 · 被引用 201 次
相关 Paper
- Stereo Risk: A Continuous Modeling Approach to Stereo MatchingCe Liu, Suryansh Kumar, Shuhang Gu, Radu Timofte 等ICML 2024 · 被引用 8 次
- On the Over-Smoothing Problem of CNN Based Disparity EstimationChuangrong Chen, Xiaozhi Chen, Hui ChengICCV 2019 · 被引用 24 次
- GraftNet: Towards Domain Generalized Stereo Matching with a Broad-Spectrum and Task-Oriented FeatureBiyang Liu, Huimin Yu, Guodong QiCVPR 2022 · 被引用 52 次
- Local Similarity Pattern and Cost Self-Reassembling for Deep Stereo Matching NetworksBiyang Liu, Huimin Yu, Yangqi LongAAAI 2022 · 被引用 86 次
- StereoGAN: Bridging Synthetic-to-Real Domain Gap by Joint Optimization of Domain Translation and Stereo MatchingRui Liu, Chengxi Yang, Wenxiu Sun, Xiaogang Wang 等CVPR 2020
