Handling Varied Objectives by Online Decision Making
Lanjihong Ma, Zhen-Yu Zhang, Yao-Xiang Ding, Zhi-Hua Zhou
摘要
Conventional machine learning typically assume a fixed learning objective throughout the learning process.However, for real-world tasks in open and dynamic environments, objectives can change frequently.For example, in autonomous driving, a car has several default modes, but a user's concern for speed and fuel consumption varies depending on road conditions and personal needs.We formulate this problem as learning with varied objectives (LVO), where the goal is to optimize a dynamic weighted combination of multiple sub-objectives by sequentially selecting actions that incur different losses on these sub-objectives.We propose the VaRons algorithm, which estimates the action-wise performance on each sub-objective and adaptively selects decisions according to the dynamic requirements on different sub-objectives.Further, we extend our approach to cases involving contextual representations and propose the Con-VaRons algorithm, assuming parameterized linear structure that links contextual features to the main objective.Both the VaRons and ConVaRons are provably minimax optimal with respect to the time horizon , with ConVaRons showing better dependency with the number of sub-objectives .Experiments on dynamic classifier and real-world cluster service allocation tasks validate the effectiveness of our methods and support our theoretical findings. CCS Concepts Computing methodologies Online learning settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Beyond UCB: Optimal and Efficient Contextual Bandits with Regression OraclesDylan J. Foster, Alexander RakhlinICML 2020 · 被引用 241 次
- Learning with Feature and Distribution Evolvable StreamsZhenyu Zhang, Peng Zhao, Yuan Jiang, Zhi-Hua ZhouICML 2020 · 被引用 48 次
- An Unbiased Risk Estimator for Learning with Augmented ClassesYu-Jie Zhang, Peng Zhao, Lanjihong Ma, Zhi-Hua ZhouNeurIPS 2020 · 被引用 31 次
- Exploratory Machine Learning with Unknown UnknownsPeng Zhao, Yu-Jie Zhang, Zhi-Hua ZhouAAAI 2021 · 被引用 29 次
- Towards Driving-Oriented Metric for Lane Detection ModelsTakami Sato, Qi Alfred ChenCVPR 2022 · 被引用 15 次
相关 Paper
- Distributional Pareto-Optimal Multi-Objective Reinforcement LearningXin-Qiang Cai, Pushi Zhang, Li Zhao, Jiang Bian 等NeurIPS 2023 · 被引用 46 次
- Bayesian Optimization for Unknown Cost-Varying Variable Subsets with No-Regret CostsVu Viet Hoang, Quoc Anh Hoang Nguyen, Hung Tran TheAAAI 2025
- Agnostic Learning with Multiple ObjectivesCorinna Cortes, Mehryar Mohri, Javier Gonzalvo, Dmitry StorcheusNeurIPS 2020 · 被引用 25 次
- Blind Optimal User Association in Small-Cell NetworksLivia Elena Chatzieleftheriou, Apostolos Destounis, Georgios S. Paschos, Iordanis KoutsopoulosINFOCOM 2021 · 被引用 1 次
- A distributional view on multi-objective policy optimizationAbbas Abdolmaleki, Sandy H. Huang, Leonard Hasenclever, Michael Neunert 等ICML 2020 · 被引用 93 次
