Distribution-Aware Testing of Neural Networks Using Generative Models
Swaroopa Dola, Matthew B. Dwyer, Mary Lou Soffa
摘要
The reliability of software that has a Deep Neural Network (DNN) as a component is urgently important today given the increasing number of critical applications being deployed with DNNs. The need for reliability raises a need for rigorous testing of the safety and trustworthiness of these systems. In the last few years, there have been a number of research efforts focused on testing DNNs. However the test generation techniques proposed so far lack a check to determine whether the test inputs they are generating are valid, and thus invalid inputs are produced. To illustrate this situation, we explored three recent DNN testing techniques. Using deep generative model based input validation, we show that all the three techniques generate significant number of invalid test inputs. We further analyzed the test coverage achieved by the test inputs generated by the DNN testing techniques and showed how invalid test inputs can falsely inflate test coverage metrics. To overcome the inclusion of invalid inputs in testing, we propose a technique to incorporate the valid input space of the DNN model under test in the test generation process. Our technique uses a deep generative model-based algorithm to generate only valid inputs. Results of our empirical studies show that our technique is effective in eliminating invalid tests and boosting the number of valid test inputs generated.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- ThirdEye: Attention Maps for Safe Autonomous Driving SystemsAndrea Stocco, Paulo J. Nunes, Marcelo d'Amorim, Paolo TonellaASE 2022 · 被引用 43 次
- DeepMetis: Augmenting a Deep Learning Test Set to Increase its Mutation ScoreVincenzo Riccio, Nargiz Humbatova, Gunel Jahangirova, Paolo TonellaASE 2021 · 被引用 41 次
- When and Why Test Generators for Deep Learning Produce Invalid Inputs: an Empirical StudyVincenzo Riccio, Paolo TonellaICSE 2023 · 被引用 29 次
- Regression Fuzzing for Deep Learning SystemsHanmo You, Zan Wang, Junjie Chen, Shuang Liu 等ICSE 2023 · 被引用 28 次
- DistXplore: Distribution-Guided Testing for Evaluating and Enhancing Deep Learning SystemsLongtian Wang, Xiaofei Xie, Xiaoning Du, Meng Tian 等FSE 2023 · 被引用 15 次
它引用的顶会 Paper1
相关 Paper
- DeepGini: prioritizing massive tests to enhance the robustness of deep neural networksYang Feng, Qingkai Shi, Xinyu Gao, Jun Wan 等ISSTA 2020 · 被引用 206 次
- CIT4DNN: Generating Diverse and Rare Inputs for Neural Networks Using Latent Space Combinatorial TestingSwaroopa Dola, Rory McDaniel, Matthew B. Dwyer, Mary Lou SoffaICSE 2024 · 被引用 12 次
- Evaluating Deep Neural Networks in Deployment: A Comparative Study (Replicability Study)Eduard Pinconschi, Divya Gopinath, Rui Abreu, Corina S. PasareanuISSTA 2024
- DeepState: Selecting Test Suites to Enhance the Robustness of Recurrent Neural NetworksZixi Liu, Yang Feng, Yining Yin, Zhenyu ChenICSE 2022 · 被引用 17 次
- Efficient Online Testing for DNN-Enabled Systems using Surrogate-Assisted and Many-Objective OptimizationFitash Ul Haq, Donghwan Shin, Lionel C. BriandICSE 2022 · 被引用 77 次
