Functional Bilevel Optimization for Machine Learning
Ieva Petrulionyte, Julien Mairal, Michael Arbel
摘要
In this paper, we introduce a new functional point of view on bilevel optimization problems for machine learning, where the inner objective is minimized over a function space. These types of problems are most often solved by using methods developed in the parametric setting, where the inner objective is strongly convex with respect to the parameters of the prediction function. The functional point of view does not rely on this assumption and notably allows using over-parameterized neural networks as the inner prediction function. We propose scalable and efficient algorithms for the functional bilevel optimization problem and illustrate the benefits of our approach on instrumental regression and reinforcement learning tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Cautious Weight DecayLizhang Chen, Jonathan Li, Kaizhao Liang, Baiyu Su 等ICLR 2026 · 被引用 14 次
- Beyond Value Functions: Single-Loop Bilevel Optimization under Flatness ConditionsLiuyuan Jiang, Quan Xiao, Lisha Chen, Tianyi ChenNeurIPS 2025 · 被引用 11 次
- CoBo: Collaborative Learning via Bilevel OptimizationDiba Hashemi, Lie He, Martin JaggiNeurIPS 2024 · 被引用 8 次
- Demystifying Spectral Feature Learning for Instrumental Variable RegressionDimitri Meunier, Antoine Moulin, Jakub Wornbard, Vladimir Kostic 等NeurIPS 2025 · 被引用 5 次
- Efficiency Follows Global-Local DecouplingZhenyu Yang, Gensheng Pei, Tao Chen, Yichao Zhou 等CVPR 2026 · 被引用 3 次
它引用的顶会 Paper23
- Efficient and Modular Implicit DifferentiationMathieu Blondel, Quentin Berthet, Marco Cuturi, Roy Frostig 等NeurIPS 2022 · 被引用 386 次
- Bilevel Optimization: Convergence Analysis and Enhanced DesignKaiyi Ji, Junjie Yang, Yingbin LiangICML 2021 · 被引用 343 次
- Symplectic ODE-Net: Learning Hamiltonian Dynamics with ControlYaofeng Desmond Zhong, Biswadip Dey, Amit ChakrabortyICLR 2020 · 被引用 319 次
- Invariant Risk Minimization GamesKartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, Amit DhurandharICML 2020 · 被引用 289 次
- On the Iteration Complexity of Hypergradient ComputationRiccardo Grazzi, Luca Franceschi, Massimiliano Pontil, Saverio SalzoICML 2020 · 被引用 241 次
相关 Paper
- Learning Theory for Kernel Bilevel OptimizationFares El Khoury, Edouard Pauwels, Samuel Vaiter, Michael ArbelNeurIPS 2025 · 被引用 2 次
- Neur2BiLO: Neural Bilevel OptimizationJustin Dumouchelle, Esther Julien, Jannis Kurtz, Elias B. KhalilNeurIPS 2024 · 被引用 10 次
- On the Global Optimality of Model-Agnostic Meta-LearningLingxiao Wang, Qi Cai, Zhuoran Yang, Zhaoran WangICML 2020 · 被引用 48 次
- BOME! Bilevel Optimization Made Easy: A Simple First-Order ApproachBo Liu, Mao Ye, Stephen Wright, Peter Stone 等NeurIPS 2022 · 被引用 170 次
- Linearly Constrained Bilevel Optimization: A Smoothed Implicit Gradient ApproachPrashant Khanduri, Ioannis C. Tsaknakis, Yihua Zhang, Jia Liu 等ICML 2023 · 被引用 28 次
