Monotonic Variational Gaussian Process for Efficient Data Collection
Donghyun Lee, Young Myoung Ko
Abstract
Modeling the learning curve is critical for cost-effective data collection in deep learning systems. Most prior approaches assume a specific parametric learning curve, but these can be inappropriate when no reliable parametric form can be assumed for the learning curve. While Gaussian processes offer flexible nonparametric modeling, existing GP approaches that enforce monotonicity typically introduce intractable factors or require derivative observations. To address this, we propose a Monotonic Variational Gaussian Process for Efficient Data Collection (MOVE), which (i) introduces a novel monotonic variational GP formulation with virtual-derivative factors to enable tractable posterior inference, and (ii) develops an expected shortfall based objective for target-driven data collection. Furthermore, our theoretical analysis shows that expected shortfall provides non-vanishing gradient signals that enable reliable gradient-based optimization. Extensive experiments on classification, segmentation, and detection benchmarks demonstrate consistent improvements over the prior method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aafbe09c-38ff-46ba-a0b9-9081f1d60c2dBuilds on20
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Beyond neural scaling laws: beating power law scaling via data pruningBen Sorscher, Robert Geirhos, Shashank Shekhar, Surya Ganguli et al.NeurIPS 2022 · 720 citations
- Coresets for Data-efficient Training of Machine Learning ModelsBaharan Mirzasoleiman, Jeff A. Bilmes, Jure LeskovecICML 2020 · 494 citations
- GRAD-MATCH: Gradient Matching based Data Subset Selection for Efficient Deep Model TrainingKrishnaTeja Killamsetty, Durga Sivasubramanian, Ganesh Ramakrishnan, Abir De et al.ICML 2021 · 305 citations
- Revisiting Neural Scaling Laws in Language and VisionIbrahim M. Alabdulmohsin, Behnam Neyshabur, Xiaohua ZhaiNeurIPS 2022 · 171 citations
Related papers
- Conditioning Sparse Variational Gaussian Processes for Online Decision-makingWesley J. Maddox, Samuel Stanton, Andrew Gordon WilsonNeurIPS 2021 · 44 citations
- Empirical Gaussian ProcessesJihao Andreas Lin, Sebastian Ament, Louis Tiao, David Eriksson et al.ICML 2026
- Improving Hyperparameter Learning under Approximate Inference in Gaussian Process ModelsRui Li, S. T. John, Arno SolinICML 2023 · 4 citations
- SigGPDE: Scaling Sparse Gaussian Processes on Sequential DataMaud Lemercier, Cristopher Salvi, Thomas Cass, Edwin V. Bonilla et al.ICML 2021 · 30 citations
- Fast and Efficient DNN Deployment via Deep Gaussian Transfer LearningQi Sun, Chen Bai, Tinghuan Chen, Hao Geng et al.ICCV 2021 · 7 citations
