Architecture Disentanglement for Deep Neural Networks
Jie Hu, Liujuan Cao, Tong Tong, Qixiang Ye, Shengchuan Zhang, Ke Li, Feiyue Huang, Ling Shao, Rongrong Ji
摘要
Understanding the inner workings of deep neural networks (DNNs) is essential to provide trustworthy artificial intelligence techniques for practical applications. Existing studies typically involve linking semantic concepts to units or layers of DNNs, but fail to explain the inference process. In this paper, we introduce neural architecture disentanglement (NAD) to fill the gap. Specifically, NAD learns to disentangle a pre-trained DNN into sub-architectures according to independent tasks, forming information flows that describe the inference processes. We investigate whether, where, and how the disentanglement occurs through experiments conducted with handcrafted and automatically-searched network architectures, on both object-based and scene-based datasets. Based on the experimental results, we present three new findings that provide fresh insights into the inner logic of DNNs. First, DNNs can be divided into sub-architectures for independent tasks. Second, deeper layers do not always correspond to higher semantics. Third, the connection type in a DNN affects how the information flows across layers, leading to different disentanglement behaviors. With NAD, we further explain why DNNs sometimes give wrong predictions. Experimental results show that misclassified images have a high probability of being assigned to task sub-architectures similar to the correct ones. Our code is available at https://github.com/hujiecpp/NAD.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Aha! Adaptive History-driven Attack for Decision-based Black-box ModelsJie Li, Rongrong Ji, Peixian Chen, Baochang Zhang 等ICCV 2021 · 被引用 25 次
- Pareto Deep Long-Tailed Recognition: A Conflict-Averse SolutionZhipeng Zhou, Liu Liu, Peilin Zhao, Wei GongICLR 2024 · 被引用 12 次
- Model LEGO: Creating Models Like Disassembling and Assembling Building BlocksJiacong Hu, Jing Gao, Jingwen Ye, Yang Gao 等NeurIPS 2024 · 被引用 1 次
- Layer-Centric Factors of Variation Disentanglement for Task- and Model-Agnostic GeneralizationHee-Jun Jung, Jongmin Park, Minwoo Kang, Hoyong Kim 等ICML 2026
相关 Paper
- Towards Disentangling Information Paths with Coded ResNeXtApostolos Avranas, Marios KountourisNeurIPS 2022 · 被引用 1 次
- Knowledge Consistency between Neural Networks and BeyondRuofan Liang, Tianlin Li, Longfei Li, Jing Wang 等ICLR 2020 · 被引用 30 次
- Learning Protein Structure-Function Relationships through Knowledge-guided Representation DecompositionMingqing Wang, Zhiwei Nie, ATHANASIOS VASILAKOS, Yonghong He 等ICML 2026
- A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual ConceptsYunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng 等CVPR 2021
- A Disentangling Invertible Interpretation Network for Explaining Latent RepresentationsPatrick Esser, Robin Rombach, Björn OmmerCVPR 2020
