Towards Theoretical Analysis of Transformation Complexity of ReLU DNNs
Jie Ren, Mingjie Li, Meng Zhou, Shih-Han Chan, Quanshi Zhang
摘要
This paper aims to theoretically analyze the complexity of feature transformations encoded in piecewise linear DNNs with ReLU layers. We propose metrics to measure three types of complexities of transformations based on the information theory. We further discover and prove the strong correlation between the complexity and the disentanglement of transformations. Based on the proposed metrics, we analyze two typical phenomena of the change of the transformation complexity during the training process, and explore the ceiling of a DNN's complexity. The proposed metrics can also be used as a loss to learn a DNN with the minimum complexity, which also controls the over-fitting level of the DNN and influences adversarial robustness, adversarial transferability, and knowledge consistency. Comprehensive comparative studies have provided new perspectives to understand the DNN. The code is released at https://github.com/sjtu-XAI-lab/transformationcomplexity .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Deep Double Descent: Where Bigger Models and More Data HurtPreetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang 等ICLR 2020 · 被引用 1,108 次
- Learning Latent Space Energy-Based Prior ModelBo Pang, Tian Han, Erik Nijkamp, Song-Chun Zhu 等NeurIPS 2020 · 被引用 152 次
- A Unified Approach to Interpreting and Boosting Adversarial TransferabilityXin Wang, Jie Ren, Shuyun Lin, Xiangming Zhu 等ICLR 2021 · 被引用 113 次
- Early Stopping in Deep Networks: Double Descent and How to Eliminate itReinhard Heckel, Fatih Furkan YilmazICLR 2021 · 被引用 55 次
- Knowledge Consistency between Neural Networks and BeyondRuofan Liang, Tianlin Li, Longfei Li, Jing Wang 等ICLR 2020 · 被引用 30 次
相关 Paper
- Interpreting and Disentangling Feature Components of Various Complexity from DNNsJie Ren, Mingjie Li, Zexu Liu, Quanshi ZhangICML 2021 · 被引用 20 次
- On the Local Complexity of Linear Regions in Deep ReLU NetworksNiket Patel, Guido MontúfarICML 2025
- Interpreting Representation Quality of DNNs for 3D Point Cloud ProcessingWen Shen, Qihan Ren, Dongrui Liu, Quanshi ZhangNeurIPS 2021 · 被引用 23 次
- Quantification and Analysis of Layer-wise and Pixel-wise Information DiscardingHaotian Ma, Hao Zhang, Fan Zhou, Yinqing Zhang 等ICML 2022 · 被引用 2 次
- Understanding Model Ensemble in Transferable Adversarial AttackWei Yao, Zeliang Zhang, Huayi Tang, Yong LiuICML 2025
