Estimating informativeness of samples with Smooth Unique Information
Hrayr Harutyunyan, Alessandro Achille, Giovanni Paolini, Orchid Majumder, Avinash Ravichandran, Rahul Bhotika, Stefano Soatto
Abstract
We define a notion of information that an individual sample provides to the training of a neural network, and we specialize it to measure both how much a sample informs the final weights and how much it informs the function computed by the weights. Though related, we show that these quantities have a qualitatively different behavior. We give efficient approximations of these quantities using a linearized network and demonstrate empirically that the approximation is accurate for real-world architectures, such as pre-trained ResNets. We apply these measures to several problems, such as dataset summarization, analysis of under-sampled classes, comparison of informativeness of different data sources, and detection of adversarial and corrupted examples. Our work generalizes existing frameworks but enjoys better computational properties for heavily over-parametrized models, which makes it possible to apply it to real-world networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b0617f39-ab41-40dc-b6e9-42958ec3f8d7Cited by top-tier papers10
- Beyond neural scaling laws: beating power law scaling via data pruningBen Sorscher, Robert Geirhos, Shashank Shekhar, Surya Ganguli et al.NeurIPS 2022 · 720 citations
- Information-theoretic generalization bounds for black-box learning algorithmsHrayr Harutyunyan, Maxim Raginsky, Greg Ver Steeg, Aram GalstyanNeurIPS 2021 · 61 citations
- Estimating Example Difficulty using Variance of GradientsChirag Agarwal, Daniel D'souza, Sara HookerCVPR 2022 · 57 citations
- Understanding Instance-based Interpretability of Variational Auto-EncodersZhifeng Kong, Kamalika ChaudhuriNeurIPS 2021 · 32 citations
- GEX: A flexible method for approximating influence via Geometric EnsembleSungyub Kim, Kyungsu Kim, Eunho YangNeurIPS 2023 · 18 citations
Builds on6
- DeltaGrad: Rapid retraining of machine learning modelsYinjun Wu, Edgar Dobriban, Susan B. DavidsonICML 2020 · 262 citations
- Data Valuation using Reinforcement LearningJinsung Yoon, Sercan Ömer Arik, Tomas PfisterICML 2020 · 236 citations
- Evaluation of Neural Architectures trained with square Loss vs Cross-Entropy in Classification TasksLike Hui, Mikhail BelkinICLR 2021 · 199 citations
- Gradients as Features for Deep Representation LearningFangzhou Mu, Yingyu Liang, Yin LiICLR 2020 · 46 citations
- Predicting Training Time Without TrainingLuca Zancato, Alessandro Achille, Avinash Ravichandran, Rahul Bhotika et al.NeurIPS 2020 · 32 citations
Related papers
- Laplace Sample Information: Data Informativeness Through a Bayesian LensJohannes Kaiser, Kristian Schwethelm, Daniel Rueckert, Georgios KaissisICLR 2025
- Deep Linear Probe Generators for Weight Space LearningJonathan Kahana, Eliahu Horwitz, Imri Shuval, Yedid HoshenICLR 2025
- Efficient Parametric Approximations of Neural Network Function Space DistanceNikita Dhawan, Sicong Huang, Juhan Bae, Roger Baker GrosseICML 2023 · 7 citations
- Datamodels: Understanding Predictions with Data and Data with PredictionsAndrew Ilyas, Sung Min Park, Logan Engstrom, Guillaume Leclerc et al.ICML 2022 · 66 citations
- If Influence Functions are the Answer, Then What is the Question?Juhan Bae, Nathan Ng, Alston Lo, Marzyeh Ghassemi et al.NeurIPS 2022 · 185 citations
