ASRTest: automated testing for deep-neural-network-driven speech recognition systems
Pin Ji, Yang Feng, Jia Liu, Zhihong Zhao, Zhenyu Chen
摘要
With the rapid development of deep neural networks and end-to-end learning techniques, automatic speech recognition (ASR) systems have been deployed into our daily and assist in various tasks. However, despite their tremendous progress, ASR systems could also suffer from software defects and exhibit incorrect behaviors. While the nature of DNN makes conventional software testing techniques inapplicable for ASR systems, lacking diverse tests and oracle information further hinders their testing. In this paper, we propose and implement a testing approach, namely ASR, specifically for the DNN-driven ASR systems. ASRTest is built upon the theory of metamorphic testing. We first design the metamorphic relation for ASR systems and then implement three families of transformation operators that can simulate practical application scenarios to generate speeches. Furthermore, we adopt Gini impurity to guide the generation process and improve the testing efficiency. To validate the effectiveness of ASRTest, we apply ASRTest to four ASR models with four widely-used datasets. The results show that ASRTest can detect erroneous behaviors under different realistic application conditions efficiently and improve 19.1% recognition performance on average via retraining with the generated data. Also, we conduct a case study on an industrial ASR system to investigate the performance of ASRTest under the real usage scenario. The study shows that ASRTest can detect errors and improve the performance of DNN-driven ASR systems effectively.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- FedSlice: Protecting Federated Learning Models from Malicious Participants with Model SlicingZiqi Zhang, Yuanchun Li, Bingyan Liu, Yifeng Cai 等ICSE 2023 · 被引用 8 次
- Synthesizing Speech Test Cases with Text-to-Speech? An Empirical Study on the False Alarms in Automated Speech Recognition TestingJulia Kaiwen Lau, Kelvin Kai Wen Kong, Julian Hao Yong, Per Hoong Tan 等ISSTA 2023 · 被引用 5 次
- ROME: Testing Image Captioning Systems via Recursive Object MeltingBoxi Yu, Zhiqing Zhong, Jiaqi Li, Yixing Yang 等ISSTA 2023 · 被引用 3 次
- COSTELLO: Contrastive Testing for Embedding-Based Large Language Model as a Service EmbeddingsWeipeng Jiang, Juan Zhai, Shiqing Ma, Xiaoyu Zhang 等FSE 2024 · 被引用 1 次
相关 Paper
- DialTest: automated testing for recurrent-neural-network-driven dialogue systemsZixi Liu, Yang Feng, Zhenyu ChenISSTA 2021 · 被引用 26 次
- Metamorphic Object Insertion for Testing Object Detection SystemsShuai Wang, Zhendong SuASE 2020 · 被引用 69 次
- DeepGini: prioritizing massive tests to enhance the robustness of deep neural networksYang Feng, Qingkai Shi, Xinyu Gao, Jun Wan 等ISSTA 2020 · 被引用 206 次
- Unveiling Hidden DNN Defects with Decision-Based Metamorphic TestingYuanyuan Yuan, Qi Pang, Shuai WangASE 2022 · 被引用 16 次
- Natural Test Generation for Precise Testing of Question Answering SoftwareQingchao Shen, Junjie Chen, Jie M. Zhang, Haoyu Wang 等ASE 2022 · 被引用 27 次
