Incorporating Surrogate Gradient Norm to Improve Offline Optimization Techniques
Cuong Dao, Phi Le Nguyen, Truong Thao Nguyen, Nghia Hoang
Abstract
Offline optimization has recently emerged as an increasingly popular approach to mitigate the prohibitively expensive cost of online experimentation. The key idea is to learn a surrogate of the black-box function that underlines the target experiment using a static (offline) dataset of its previous input-output queries. Such an approach is, however, fraught with an out-of-distribution issue where the learned surrogate becomes inaccurate outside the offline data regimes. To mitigate this, existing offline optimizers have proposed numerous conditioning techniques to prevent the learned surrogate from being too erratic. Nonetheless, such conditioning strategies are often specific to particular surrogate or search models, which might not generalize to a different model choice. This motivates us to develop a model-agnostic approach instead, which incorporates a notion of model sharpness into the training loss of the surrogate as a regularizer. Our approach is supported by a new theoretical analysis demonstrating that reducing surrogate sharpness on the offline dataset provably reduces its generalized sharpness on unseen data. Our analysis extends existing theories from bounding generalized prediction loss (on unseen data) with loss sharpness to bounding the worst-case generalized surrogate sharpness with its empirical estimate on training data, providing a new perspective on sharpness regularization. Our extensive experimentation on a diverse range of optimization tasks also shows that reducing surrogate sharpness often leads to significant improvement, marking (up to) a noticeable 9.6% performance boost. Our code is publicly available at https://github.com/cuong-dm/IGNITE
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b72edeba-4f95-4941-b011-1f6c26cfbf93Cited by top-tier papers2
- ROOT: Rethinking Offline Optimization as Distributional Translation via Probabilistic BridgeCuong Dao, The Hung Tran, Phi Le Nguyen, Truong Thao Nguyen et al.NeurIPS 2025 · 4 citations
- Offline Model-Based Optimization by Learning to RankRong-Xi Tan, Ke Xue, Shen-Huan Lyu, Haopu Shang et al.ICLR 2025
Builds on14
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- Penalizing Gradient Norm for Efficiently Improving Generalization in Deep LearningYang Zhao, Hao Zhang, Xiuyuan HuICML 2022 · 165 citations
- Model Inversion Networks for Model-Based OptimizationAviral Kumar, Sergey LevineNeurIPS 2020 · 129 citations
- Design-Bench: Benchmarks for Data-Driven Offline Model-Based OptimizationBrandon Trabucco, Xinyang Geng, Aviral Kumar, Sergey LevineICML 2022 · 126 citations
- Conservative Objective Models for Effective Offline Model-Based OptimizationBrandon Trabucco, Aviral Kumar, Xinyang Geng, Sergey LevineICML 2021 · 119 citations
Related papers
- Boosting Offline Optimizers with Surrogate SensitivityManh Cuong Dao, Phi Le Nguyen, Truong Thao Nguyen, Trong Nghia HoangICML 2024 · 10 citations
- Offline Model-Based Optimization via Policy-Guided Gradient SearchYassine Chemingui, Aryan Deshwal, Trong Nghia Hoang, Janardhan Rao DoppaAAAI 2024 · 22 citations
- Learning Surrogates for Offline Black-Box Optimization via Gradient MatchingMinh Hoang, Azza Fadhel, Aryan Deshwal, Jana Doppa et al.ICML 2024 · 18 citations
- Generative Adversarial Model-Based Optimization via Source Critic RegularizationMichael S. Yao, Yimeng Zeng, Hamsa Bastani, Jacob R. Gardner et al.NeurIPS 2024 · 14 citations
- Q-SAM: Unlocking Sharpness-Aware Minimization for Generalization in Offline Reinforcement LearningDa Wang, Yi Ma, Ting Guo, Lin Li et al.ICML 2026
