Efficiently Controlling Multiple Risks with Pareto Testing
Bracha Laufer-Goldshtein, Adam Fisch, Regina Barzilay, Tommi S. Jaakkola
Abstract
Machine learning applications frequently come with multiple diverse objectives and constraints that can change over time. Accordingly, trained models can be tuned with sets of hyper-parameters that affect their predictive behavior (e.g., their run-time efficiency versus error rate). As the number of constraints and hyper-parameter dimensions grow, naively selected settings may lead to sub-optimal and/or unreliable results. We develop an efficient method for calibrating models such that their predictions provably satisfy multiple explicit and simultaneous statistical guarantees (e.g., upper-bounded error rates), while also optimizing any number of additional, unconstrained objectives (e.g., total run-time cost). Building on recent results in distribution-free, finite-sample risk control for general losses, we propose Pareto Testing: a two-stage process which combines multi-objective optimization with multiple hypothesis testing. The optimization stage constructs a set of promising combinations on the Pareto frontier. We then apply statistical testing to this frontier only to identify configurations that have (i) high utility with respect to our objectives, and (ii) guaranteed risk levels with respect to our constraints, with specifiable high probability. We demonstrate the effectiveness of our approach to reliably accelerate the execution of large-scale Transformer models in natural language processing (NLP) applications. In particular, we show how Pareto Testing can be used to dynamically configure multiple inter-dependent model attributes -- including the number of layers computed before exiting, number of attention heads pruned, or number of text tokens considered -- to simultaneously control and optimize various accuracy and cost metrics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 863af83c-2ff2-4eb4-85d6-7a2e657040bdCited by top-tier papers8
- Conformal Language ModelingVictor Quach, Adam Fisch, Tal Schuster, Adam Yala et al.ICLR 2024 · 132 citations
- How to Trust Your Diffusion Model: A Convex Optimization Approach to Conformal Risk ControlJacopo Teneggi, Matthew Tivnan, J. Webster Stayman, Jeremias SulamICML 2023 · 49 citations
- PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial AttacksChen Feng, Ziquan Liu, Zhuo Zhi, Ilija Bogunovic et al.AAAI 2025 · 15 citations
- Early Time Classification with Accumulated Accuracy Gap ControlLiran Ringel, Regev Cohen, Daniel Freedman, Michael Elad et al.ICML 2024 · 9 citations
- Multi-Objective Hyperparameter Selection via Hypothesis Testing on Reliability GraphsAmirmohammad Farzaneh, Osvaldo SimeoneNeurIPS 2025 · 1 citation
Builds on14
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited DevicesZhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu et al.ACL 2020 · 660 citations
- DynaBERT: Dynamic BERT with Adaptive Width and DepthLu Hou, Zhiqi Huang, Lifeng Shang, Xin Jiang et al.NeurIPS 2020 · 401 citations
- Confident Adaptive Language ModelingTal Schuster, Adam Fisch, Jai Gupta, Mostafa Dehghani et al.NeurIPS 2022 · 394 citations
- FastBERT: a Self-distilling BERT with Adaptive Inference TimeWeijie Liu, Peng Zhou, Zhiruo Wang, Zhe Zhao et al.ACL 2020 · 257 citations
Related papers
- Conformal Arbitrage: Risk-Controlled Balancing of Competing Objectives in Language ModelsWilliam Overman, Mohsen BayatiNeurIPS 2025 · 12 citations
- Conformal Thinking: Risk Control for Reasoning on a Compute BudgetXi Wang, Anushri Suresh, Alvin Zhang, Rishi More et al.ICML 2026 · 10 citations
- PASHA: Efficient HPO and NAS with Progressive Resource AllocationOndrej Bohdal, Lukas Balles, Martin Wistuba, Beyza Ermis et al.ICLR 2023 · 4 citations
- Landmark-Guided Policy Optimization for Multi-Objective Language Model SelectionMarcio Monteiro, Weichen Li, Puyu Wang, Marius Kloft et al.ICML 2026
- Consistent Accelerated Inference via Confident Adaptive TransformersTal Schuster, Adam Fisch, Tommi S. Jaakkola, Regina BarzilayEMNLP 2021 · 30 citations
