Causes and Effects of Unanticipated Numerical Deviations in Neural Network Inference Frameworks
Alexander Schlögl, Nora Hofer, Rainer Böhme
摘要
Hardware-specific optimizations in machine learning (ML) frameworks can cause numerical deviations of inference results. Quite surprisingly, despite using a fixed trained model and fixed input data, inference results are not consistent across platforms, and sometimes not even deterministic on the same platform. We study the causes of these numerical deviations for convolutional neural networks (CNN) on realistic end-to-end inference pipelines and in isolated experiments. Results from 75 distinct platforms suggest that the main causes of deviations on CPUs are differences in SIMD use, and the selection of convolution algorithms at runtime on GPUs. We link the causes and propagation effects to properties of the ML model and evaluate potential mitigations. We make our research code publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Synchronizing Probabilities in Model-Driven Lossless CompressionAviv Adler, Jennifer TangICLR 2026 · 被引用 1 次
- Adversarial Inputs for Linear Algebra BackendsJonas Möller, Lukas Pirch, Felix Weissberg, Sebastian Baunsgaard 等ICML 2025
- Hardware and Software Platform InferenceCheng Zhang, Hanna Foerster, Robert D. Mullins, Yiren Zhao 等ICML 2025
- No Soundness in the Real World: On the Challenges of the Verification of Deployed Neural NetworksAttila Szász, Balázs Bánhelyi, Márk JelasityICML 2025
- Towards a Re-evaluation of Data Forging Attacks in PracticeMohamed Suliman, Anisa Halimi, Swanand Ravindra Kadhe, Nathalie Baracaldo 等USENIX Security 2025
它引用的顶会 Paper5
- Problems and Opportunities in Training Deep Learning Software Systems: An Analysis of VarianceHung Viet Pham, Shangshu Qian, Jiannan Wang, Thibaud Lutellier 等ASE 2020 · 被引用 91 次
- Optimizing batched Winograd convolution on GPUsDa Yan, Wei Wang, Xiaowen ChuPPoPP 2020 · 被引用 64 次
- Fooling a Complete Neural Network VerifierDániel Zombori, Balázs Bánhelyi, Tibor Csendes, István Megyeri 等ICLR 2021 · 被引用 21 次
- Widespread Underestimation of Sensitivity in Differentially Private Libraries and How to Fix ItSílvia Casacuberta, Michael Shoemate, Salil P. Vadhan, Connor WagamanCCS 2022 · 被引用 12 次
- Dos and Don'ts of Machine Learning in Computer SecurityDaniel Arp, Erwin Quiring, Feargus Pendlebury, Alexander Warnecke 等USENIX Security 2022
相关 Paper
- An Investigation on Numerical Bugs in GPU Programs Towards Automated Bug DetectionRavishka Rathnasuriya, Nidhi Majoju, Zihe Song, Wei YangISSTA 2025
- Efficient Algorithms for Device Placement of DNN Graph OperatorsJakub Tarnawski, Amar Phanishayee, Nikhil R. Devanur, Divya Mahajan 等NeurIPS 2020 · 被引用 84 次
- pLiner: isolating lines of floating-point code for compiler-induced variabilityHui Guo, Ignacio Laguna, Cindy Rubio-GonzálezSC 2020 · 被引用 15 次
- DeepCache: Revisiting Cache Side-Channel Attacks in Deep Neural Networks ExecutablesZhibo Liu, Yuanyuan Yuan, Yanzuo Chen, Sihang Hu 等CCS 2024 · 被引用 3 次
- LaLaRAND: Flexible Layer-by-Layer CPU/GPU Scheduling for Real-Time DNN TasksWoosung Kang, Kilho Lee, Jinkyu Lee, Insik Shin 等RTSS 2021 · 被引用 68 次
