Stochastic Amortization: A Unified Approach to Accelerate Feature and Data Attribution
Ian Covert, Chanwoo Kim, Su-In Lee, James Y. Zou, Tatsunori B. Hashimoto
摘要
Many tasks in explainable machine learning, such as data valuation and feature attribution, perform expensive computation for each data point and are intractable for large datasets. These methods require efficient approximations, and although amortizing the process by learning a network to directly predict the desired output is a promising solution, training such models with exact labels is often infeasible. We therefore explore training amortized models with noisy labels, and we find that this is inexpensive and surprisingly effective. Through theoretical analysis of the label noise and experiments with various models and datasets, we show that this approach tolerates high noise levels and significantly accelerates several feature attribution and data valuation methods, often yielding an order of magnitude speedup over existing approaches.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Enhancing Training Data Attribution with Representational OptimizationWeiwei Sun, Haokun Liu, Nikhil Kandpal, Colin A. Raffel 等NeurIPS 2025 · 被引用 9 次
- SHAP zero Explains Biological Sequence Models with Near-zero Marginal Cost for Future QueriesDarin Tsui, Aryan Musharaf, Yigit Efe Erginbas, Justin Singh Kang 等NeurIPS 2025 · 被引用 5 次
- On-device Content-based Recommendation with Single-shot Embedding Pruning: A Cooperative Game PerspectiveHung Vinh Tran, Tong Chen, Guanhua Ye, Quoc Viet Hung Nguyen 等WWW 2025 · 被引用 4 次
- Selective ExplanationsLucas Monteiro Paes, Dennis Wei, Flávio P. CalmonNeurIPS 2024 · 被引用 4 次
- MAnchors: Memorization-Based Acceleration of Anchors via Rule Reuse and TransformationHaonan Yu, Junhao Liu, Xin ZhangICML 2026 · 被引用 2 次
它引用的顶会 Paper27
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- What Neural Networks Memorize and Why: Discovering the Long Tail via Influence EstimationVitaly Feldman, Chiyuan ZhangNeurIPS 2020 · 被引用 674 次
- Understanding Deep Networks via Extremal Perturbations and Smooth MasksRuth Fong, Mandela Patrick, Andrea VedaldiICCV 2019 · 被引用 480 次
- The Shapley Taylor Interaction IndexMukund Sundararajan, Kedar Dhamdhere, Ashish AgarwalICML 2020 · 被引用 199 次
- FastSHAP: Real-Time Shapley Value EstimationNeil Jethani, Mukund Sudarshan, Ian Connick Covert, Su-In Lee 等ICLR 2022 · 被引用 186 次
相关 Paper
- Efficient Shapley Values Estimation by Amortization for Text ClassificationChenghao Yang, Fan Yin, He He, Kai-Wei Chang 等ACL 2023 · 被引用 2 次
- Transfer and Marginalize: Explaining Away Label Noise with Privileged InformationMark Collier, Rodolphe Jenatton, Effrosyni Kokiopoulou, Jesse BerentICML 2022 · 被引用 19 次
- SAVA: Scalable Learning-Agnostic Data ValuationSamuel Kessler, Tam Le, Vu NguyenICLR 2025
- EcoVal: An Efficient Data Valuation Framework for Machine LearningAyush K. Tarun, Vikram S. Chundawat, Murari Mandal, Hong Ming Tan 等KDD 2024 · 被引用 3 次
- Amortized Variational Inference for Partial-Label Learning: A Probabilistic Approach to Label DisambiguationTobias Fuchs, Nadja KleinICML 2026
