Gradient-Based Nonlinear Rehearsal Learning with Multivariate Alterations
Tian Qin, Tian-Zuo Wang, Zhi-Hua Zhou
摘要
Machine learning (ML) has made significant advancements across various domains, with a shifting focus from purely predictive tasks to decision-making. The recent proposal by Zhou (2022) introduced a line of research known as rehearsal learning, which provides a novel perspective on modeling decision-making tasks. However, previous studies mainly focused on the linear Gaussian setting to constrain the modeling complexity. Furthermore, it has been demonstrated that finding exact optimal multivariate decisions within the sampling-based rehearsal framework is computationally infeasible in polynomial time, necessitating the development of approximate methods. In this work, we present Grad-Rh, the first gradient-based rehearsal learning method that can efficiently find multivariate decisions under non-linear and non-Gaussian settings. We address the uncertainty in decision-making tasks using flexible and expressive conditional normalizing flow models and derive four surrogate loss functions to enable efficient gradient-based optimization. Experimental results show that Grad-Rh performs comparably to exact baselines on linear data and significantly outperforms them on non-linear data in both decision quality and running time.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Structural Causal Bandits under Markov EquivalenceMin Woo Park, Andy Arditi, Elias Bareinboim, Sanghack LeeNeurIPS 2025 · 被引用 3 次
- Counterfactual Structural Causal BanditsMin Woo Park, Sanghack LeeICLR 2026 · 被引用 1 次
- Variance-Reduced Long-Term Rehearsal Learning with Quadratic Programming ReformulationWen-Bo Du, Tian Qin, Tian-Zuo Wang, Zhi-Hua ZhouNeurIPS 2025 · 被引用 1 次
- On Measuring Influence in Avoiding Undesired FutureLue Tao, Tian-Zuo Wang, Yuan Jiang, Zhi-Hua ZhouICLR 2026
它引用的顶会 Paper6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Sound and Complete Causal Identification with Latent Variables Given Local Background KnowledgeTian-Zuo Wang, Tian Qin, Zhi-Hua ZhouNeurIPS 2022 · 被引用 22 次
- Estimating Possible Causal Effects with Latent Variables via AdjustmentTian-Zuo Wang, Tian Qin, Zhi-Hua ZhouICML 2023 · 被引用 16 次
- Rehearsal Learning for Avoiding Undesired FutureTian Qin, Tian-Zuo Wang, Zhi-Hua ZhouNeurIPS 2023 · 被引用 8 次
- Avoiding Undesired Future with Minimal Cost in Non-Stationary EnvironmentsWen-Bo Du, Tian Qin, Tian-Zuo Wang, Zhi-Hua ZhouNeurIPS 2024 · 被引用 6 次
相关 Paper
- Enabling Optimal Decisions in Rehearsal Learning under CARE ConditionWen-Bo Du, Hao-Yi Lei, Lue Tao, Tian-Zuo Wang 等ICML 2025
- Policy Rehearsing: Training Generalizable Policies for Reinforcement LearningChengxing Jia, Chenxiao Gao, Hao Yin, Fuxiang Zhang 等ICLR 2024 · 被引用 6 次
- Locally Convex Global Loss Network for Decision-Focused LearningHaeun Jeon, Hyunglip Bae, Minsu Park, Chanyeong Kim 等AAAI 2025 · 被引用 6 次
- Reliable Off-Policy Learning for Dosage CombinationsJonas Schweisthal, Dennis Frauen, Valentyn Melnychuk, Stefan FeuerriegelNeurIPS 2023 · 被引用 22 次
- Probabilistic Forecasting of Irregularly Sampled Time Series with Missing Values via Conditional Normalizing FlowsVijaya Krishna Yalavarthi, Randolf Scholz, Stefan Born, Lars Schmidt-ThiemeAAAI 2025 · 被引用 6 次
