Error Correction Output Codes for Robust Neural Networks against Weight-errors: A Neural Tangent Kernel Point of View
Anlan Yu, Shusen Jing, Ning Lyu, Wujie Wen, Zhiyuan Yan
Abstract
Error correcting output code (ECOC) is a classic method that encodes binary classifiers to tackle the multi-class classification problem in decision trees and neural networks. Among ECOCs, the one-hot code has become the default choice in modern deep neural networks (DNNs) due to its simplicity in decision making. However, it suffers from a significant limitation in its ability to achieve high robust accuracy, particularly in the presence of weight-errors. While recent studies have experimentally demonstrated that the non-one-hot ECOCs with multi-bits error correction ability, could be a better solution, there is a notable absence of theoretical foundations that can elucidate the relationship between codeword design, weighterror magnitude, and network characteristics, so as to provide robustness guarantees. This work is positioned to bridge this gap through the lens of neural tangent kernel (NTK). We have two important theoretical findings: 1) In clean models (without weight-errors), utilizing one-hot code and non-one-hot ECOC is akin to altering decoding metrics from l 2 distance to Mahalanobis distance. 2) There exists a threshold, determined by the normalized distance among codewords, the DNN architecture, and the scale of weight-errors. If the distance between a clean output (without weight-errors) and its nearest codewords is smaller than this threshold, then the DNN can make predictions as if it is free of weight-errors. Based on these findings, we further demonstrate how to practically use them to identify optimal ECOCs for simple tasks (small number of classes) and complex tasks (large number of classes), by balancing the code orthogonality (as per finding 1) and code distance (as per finding 2). Extensive experimental results across four datasets and four DNN models validate the superior performance of constructed codes, guided by our findings, compared to existing ECOCs. To our best knowledge, this is the first work providing theoretical explanations for the effectiveness of ECOCs and offers associated design guidance for optimal ECOCs specifically tailored to DNNs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Error-Correcting Output Codes with Ensemble Diversity for Robust Learning in Neural NetworksYang Song, Qiyu Kang, Wee Peng TayAAAI 2021 · 23 citations
- Scalable design of Error-Correcting Output Codes using Discrete Optimization with Graph ColoringSamarth Gupta, Saurabh AminNeurIPS 2022 · 9 citations
- COLA: Orchestrating Error Coding and Learning for Robust Neural Network Inference Against Hardware DefectsAnlan Yu, Ning Lyu, Jieming Yin, Zhiyuan Yan et al.ICML 2023 · 3 citations
Related papers
- Understanding Square Loss in Training Overparametrized Neural Network ClassifiersTianyang Hu, Jun Wang, Wenjia Wang, Zhenguo LiNeurIPS 2022 · 23 citations
- PoP-ECC: Robust and Flexible Error Correction against Multi-Bit Upsets in DNN AcceleratorsTaewon Park, Saeid Gorgin, Dongwhee Kim, Jaeho Shin et al.DAC 2025 · 2 citations
- A Generalized Neural Tangent Kernel Analysis for Two-layer Neural NetworksZixiang Chen, Yuan Cao, Quanquan Gu, Tong ZhangNeurIPS 2020 · 82 citations
- Improving Robustness Against Stealthy Weight Bit-Flip Attacks by Output Code MatchingOzan Özdenizci, Robert LegensteinCVPR 2022 · 11 citations
- Bipolar vector classifier for fault-tolerant deep neural networksSuyong Lee, Insu Choi, Joon-Sung YangDAC 2022 · 4 citations
