TRADER: trace divergence analysis and embedding regulation for debugging recurrent neural networks
Guanhong Tao, Shiqing Ma, Yingqi Liu, Qiuling Xu, Xiangyu Zhang
Abstract
Recurrent Neural Networks (RNN) can deal with (textual) input with various length and hence have a lot of applications in software systems and software engineering applications. RNNs depend on word embeddings that are usually pre-trained by third parties to encode textual inputs to numerical values. It is well known that problematic word embeddings can lead to low model accuracy. In this paper, we propose a new technique to automatically diagnose how problematic embeddings impact model performance, by comparing model execution traces from correctly and incorrectly executed samples. We then leverage the diagnosis results as guidance to harden/repair the embeddings. Our experiments show that TRADER can consistently and effectively improve accuracy for real world models and datasets by 5.37% on average, which represents substantial improvement in the literature of RNN models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6665d66a-e91d-46e4-b0c4-1d0adaaeb316Cited by top-tier papers6
- AUTOTRAINER: An Automatic DNN Training Problem Detection and Repair SystemXiaoyu Zhang, Juan Zhai, Shiqing Ma, Chao ShenICSE 2021 · 62 citations
- RULER: discriminative and iterative adversarial training for deep neural network fairnessGuanhong Tao, Weisong Sun, Tingxu Han, Chunrong Fang et al.FSE 2022 · 29 citations
- Improving Binary Code Similarity Transformer Models by Semantics-Driven Instruction DeemphasisXiangzhe Xu, Shiwei Feng, Yapeng Ye, Guangyu Shen et al.ISSTA 2023 · 25 citations
- MTTM: Metamorphic Testing for Textual Content Moderation SoftwareWenxuan Wang, Jen-tse Huang, Weibin Wu, Jianping Zhang et al.ICSE 2023 · 23 citations
- An Image is Worth a Thousand Toxic Words: A Metamorphic Testing Framework for Content Moderation SoftwareWenxuan Wang, Jingyuan Huang, Jen-tse Huang, Chang Chen et al.ASE 2023 · 7 citations
Builds on1
Related papers
- RNNRepair: Automatic RNN Repair via Model-based AnalysisXiaofei Xie, Wenbo Guo, Lei Ma, Wei Le et al.ICML 2021 · 21 citations
- TRACED: Execution-aware Pre-training for Source CodeYangruibo Ding, Benjamin Steenhoek, Kexin Pei, Gail E. Kaiser et al.ICSE 2024 · 29 citations
- AI-Lancet: Locating Error-inducing Neurons to Optimize Neural NetworksYue Zhao, Hong Zhu, Kai Chen, Shengzhi ZhangCCS 2021 · 17 citations
- StateTree: A Tree-Based Modeling Approach for Fault Detection in Recurrent Neural NetworksXinyu Gao, Shuoxiao Zhang, Minghui Wei, Xiao Zhang et al.ISSTA 2026
- DeepState: Selecting Test Suites to Enhance the Robustness of Recurrent Neural NetworksZixi Liu, Yang Feng, Yining Yin, Zhenyu ChenICSE 2022 · 17 citations
