Model Doctor: A Simple Gradient Aggregation Strategy for Diagnosing and Treating CNN Classifiers
Zunlei Feng, Jiacong Hu, Sai Wu, Xiaotian Yu, Jie Song, Mingli Song
摘要
Recently, Convolutional Neural Network (CNN) has achieved excellent performance in the classification task. It is widely known that CNN is deemed as a 'black-box', which is hard for understanding the prediction mechanism and debugging the wrong prediction. Some model debugging and explanation works are developed for solving the above drawbacks. However, those methods focus on explanation and diagnosing possible causes for model prediction, based on which the researchers handle the following optimization of models manually. In this paper, we propose the first completely automatic model diagnosing and treating tool, termed as Model Doctor. Based on two discoveries that 1) each category is only correlated with sparse and specific convolution kernels, and 2) adversarial samples are isolated while normal samples are successive in the feature space, a simple aggregate gradient constraint is devised for effectively diagnosing and optimizing CNN classifiers. The aggregate gradient strategy is a versatile module for mainstream CNN classifiers. Extensive experiments demonstrate that the proposed Model Doctor applies to all existing CNN classifiers, and improves the accuracy of 16 mainstream CNN classifiers by 1% ∼ 5%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Vision Mamba MenderJiacong Hu, Anda Cao, Zunlei Feng, Shengxuming Zhang 等NeurIPS 2024 · 被引用 8 次
- Transformer Doctor: Diagnosing and Treating Vision TransformersJiacong Hu, Hao Chen, Kejia Chen, Yang Gao 等NeurIPS 2024 · 被引用 5 次
- ViT-Calibrator: Decision Stream Calibration for Vision TransformerLin Chen, Zhijie Jia, Lechao Cheng, Yang Gao 等AAAI 2024 · 被引用 3 次
- Schema Inference for Interpretable Image ClassificationHaofei Zhang, Mengqi Xue, Xiaokang Liu, Kaixuan Chen 等ICLR 2023 · 被引用 1 次
- Model LEGO: Creating Models Like Disassembling and Assembling Building BlocksJiacong Hu, Jing Gao, Jingwen Ye, Yang Gao 等NeurIPS 2024 · 被引用 1 次
它引用的顶会 Paper2
相关 Paper
- DOCTOR: A Simple Method for Detecting Misclassification ErrorsFederica Granese, Marco Romanelli, Daniele Gorla, Catuscia Palamidessi 等NeurIPS 2021 · 被引用 65 次
- Towards Trustable Skin Cancer Diagnosis via Rewriting Model's DecisionSiyuan Yan, Zhen Yu, Xuelin Zhang, Dwarikanath Mahapatra 等CVPR 2023
- Patching Weak Convolutional Neural Network Models through Modularization and CompositionBinhang Qi, Hailong Sun, Xiang Gao, Hongyu ZhangASE 2022 · 被引用 13 次
- AUTOTRAINER: An Automatic DNN Training Problem Detection and Repair SystemXiaoyu Zhang, Juan Zhai, Shiqing Ma, Chao ShenICSE 2021 · 被引用 62 次
- Agreement-Discrepancy-Selection: Active Learning with Progressive Distribution AlignmentMengying Fu, Tianning Yuan, Fang Wan, Songcen Xu 等AAAI 2021 · 被引用 13 次
