One Sample Fits All: Approximating All Probabilistic Values Simultaneously and Efficiently
Weida Li, Yaoliang Yu
Abstract
The concept of probabilistic values, such as Beta Shapley values and weighted Banzhaf values, has gained recent attention in applications like feature attribution and data valuation. However, exact computation of these values is often exponentially expensive, necessitating approximation techniques. Prior research has shown that the choice of probabilistic values significantly impacts downstream performance, with no universally superior option. Consequently, one may have to approximate multiple candidates and select the best-performing one. Although there have been many efforts to develop efficient estimators, none are intended to approximate all probabilistic values both simultaneously and efficiently. In this work, we embark on the first exploration of achieving this goal. Adhering to the principle of maximum sample reuse, we propose a one-sample-fits-all framework parameterized by a sampling vector to approximate intermediate terms that can be converted to any probabilistic value without amplifying scalars. Leveraging the concept of -approximation, we theoretically identify a key formula that effectively determines the convergence rate of our framework. By optimizing the sampling vector using this formula, we obtain i) a one-for-all estimator that achieves the currently best time complexity for all probabilistic values on average, and ii) a faster generic estimator with the sampling vector optimally tuned for each probabilistic value. Particularly, our one-for-all estimator achieves the fastest convergence rate on Beta Shapley values, including the well-known Shapley value, both theoretically and empirically. Finally, we establish a connection between probabilistic values and the least square regression used in (regularized) datamodels, showing that our one-for-all estimator can solve a family of datamodels simultaneously.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0daad5f9-7f6e-41bf-b126-c64e158a83c6Cited by top-tier papers6
- Regression-adjusted Monte Carlo Estimators for Shapley Values and Probabilistic ValuesR. Teal Witter, Yurong Liu, Christopher MuscoNeurIPS 2025 · 22 citations
- Explaining Similarity in Vision-Language Encoders with Weighted Banzhaf InteractionsHubert Baniecki, Maximilian Muschalik, Fabian Fumagalli, Barbara Hammer et al.NeurIPS 2025 · 6 citations
- TreeGrad-Ranker: Feature Ranking via O(L)-Time Gradients for Decision TreesWeida Li, Yaoliang Yu, Bryan Kian Hsiang LowICLR 2026 · 5 citations
- Faithful Group Shapley ValueKiljae Lee, Ziqi Liu, Weijing Tang, Yuan ZhangNeurIPS 2025 · 4 citations
- Priority-Aware Shapley ValueKiljae Lee, Ziqi Liu, Weijing Tang, Yuan ZhangICML 2026 · 2 citations
Builds on10
- SHAP-IQ: Unified Approximation of any-order Shapley InteractionsFabian Fumagalli, Maximilian Muschalik, Patrick Kolpaczki, Eyke Hüllermeier et al.NeurIPS 2023 · 80 citations
- Measuring the Effect of Training Data on Deep Learning Predictions via Randomized ExperimentsJinkun Lin, Anqi Zhang, Mathias Lécuyer, Jinyang Li et al.ICML 2022 · 70 citations
- WeightedSHAP: analyzing and improving Shapley based feature attributionsYongchan Kwon, James Y. ZouNeurIPS 2022 · 60 citations
- SHAQ: Incorporating Shapley Value Theory into Multi-Agent Q-LearningJianhong Wang, Yuan Zhang, Yunjie Gu, Tae-Kyun KimNeurIPS 2022 · 50 citations
- Approximating the Shapley Value without Marginal ContributionsPatrick Kolpaczki, Viktor Bengs, Maximilian Muschalik, Eyke HüllermeierAAAI 2024 · 43 citations
Related papers
- Faster Approximation of Probabilistic and Distributional Values via Least SquaresWeida Li, Yaoliang YuICLR 2024 · 13 citations
- Robust Data Valuation with Weighted Banzhaf ValuesWeida Li, Yaoliang YuNeurIPS 2023 · 30 citations
- Shapley-Based Data Valuation for Weighted -Nearest NeighborsGuangyi Zhang, Qiyu Liu, Aristides GionisNeurIPS 2025 · 2 citations
- CaSh: Shapley Value Computation with Cache OptimizationJiajun Tang, Xiaokai Mao, Ning Liu, Jinfei Liu et al.VLDB 2026
- Efficient Banzhaf-Based Data Valuation for k-Nearest Neighbors ClassificationGuangyi Zhang, Lutz Oettershagen, Lixu Wang, Aristides GionisVLDB 2026 · 1 citation
