HyDRA: Hypergradient Data Relevance Analysis for Interpreting Deep Neural Networks
Yuanyuan Chen, Boyang Li, Han Yu, Pengcheng Wu, Chunyan Miao
摘要
The behaviors of deep neural networks (DNNs) are notoriously resistant to human interpretations. In this paper, we propose Hypergradient Data Relevance Analysis, or HY-DRA, which interprets the predictions made by DNNs as effects of their training data. Existing approaches generally estimate data contributions around the final model parameters and ignore how the training data shape the optimization trajectory. By unrolling the hypergradient of test loss w.r.t. the weights of training data, HYDRA assesses the contribution of training data toward test data points throughout the training trajectory. In order to accelerate computation, we remove the Hessian from the calculation and prove that, under moderate conditions, the approximation error is bounded. Corroborating this theoretical claim, empirical results indicate the error is indeed small. In addition, we quantitatively demonstrate that HYDRA outperforms influence functions in accurately estimating data contribution and detecting noisy data labels. The source code is available at https://github.com/cyyever/aaai hydra.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Intriguing Properties of Data Attribution on Diffusion ModelsXiaosen Zheng, Tianyu Pang, Chao Du, Jing Jiang 等ICLR 2024 · 被引用 41 次
- Training Data Attribution via Approximate UnrollingJuhan Bae, Wu Lin, Jonathan Lorraine, Roger B. GrosseNeurIPS 2024 · 被引用 41 次
- "What Data Benefits My Classifier?" Enhancing Model Performance and Interpretability through Influence-Based Data SelectionAnshuman Chhabra, Peizhao Li, Prasant Mohapatra, Hongfu LiuICLR 2024 · 被引用 32 次
- Rethinking Influence Functions of Neural Networks in the Over-Parameterized RegimeRui Zhang, Shihua ZhangAAAI 2022 · 被引用 32 次
- LayerIF: Estimating Layer Quality for Large Language Models using Influence FunctionsHadi Askari, Shivanshu Gupta, Fei Wang, Anshuman Chhabra 等NeurIPS 2025 · 被引用 16 次
它引用的顶会 Paper3
- Explaining Neural Networks Semantically and QuantitativelyRunjin Chen, Hao Chen, Ge Huang, Jie Ren 等ICCV 2019 · 被引用 60 次
- Frequentist Uncertainty in Recurrent Neural Networks via Blockwise Influence FunctionsAhmed M. Alaa, Mihaela van der SchaarICML 2020 · 被引用 26 次
- Multi-Stage Influence FunctionHongge Chen, Si Si, Yang Li, Ciprian Chelba 等NeurIPS 2020 · 被引用 21 次
相关 Paper
- Outlier Gradient Analysis: Efficiently Identifying Detrimental Training Samples for Deep Learning ModelsAnshuman Chhabra, Bo Li, Jian Chen, Prasant Mohapatra 等ICML 2025
- Understanding Influence Functions and Datamodels via Harmonic AnalysisNikunj Saunshi, Arushi Gupta, Mark Braverman, Sanjeev AroraICLR 2023 · 被引用 1 次
- Towards Robust Influence Functions with Flat Validation MinimaXichen Ye, Yifan Wu, Weizhong Zhang, Cheng Jin 等ICML 2025
- Influence Functions for Scalable Data Attribution in Diffusion ModelsBruno Kacper Mlodozeniec, Runa Eschenhagen, Juhan Bae, Alexander Immer 等ICLR 2025
- Learning Trajectories are Generalization IndicatorsJingwen Fu, Zhizheng Zhang, Dacheng Yin, Yan Lu 等NeurIPS 2023 · 被引用 6 次
