When to Intervene: Learning Optimal Intervention Policies for Critical Events
Niranjan Damera Venkata, Chiranjib Bhattacharyya
摘要
Providing a timely intervention before the onset of a critical event, such as a system failure, is of importance in many industrial settings. Before the onset of the critical event, systems typically exhibit behavioral changes which often manifest as stochastic co-variate observations which may be leveraged to trigger intervention. In this paper, for the first time, we formulate the problem of finding an optimally timed intervention (OTI) policy as minimizing the expected residual time to event, subject to a constraint on the probability of missing the event. Existing machine learning approaches to intervention on critical events focus on predicting event occurrence within a pre-defined window (a classification problem) or predicting time-to-event (a regression problem). Interventions are then triggered by setting model thresholds. These are heuristic-driven, lacking guarantees regarding optimality. To model the evolution of system behavior, we introduce the concept of a hazard rate process. We show that the OTI problem is equivalent to an optimal stopping problem on the associated hazard rate process. This key link has not been explored in literature. Under Markovian assumptions on the hazard rate process, we show that an OTI policy at any time can be analytically determined from the conditional hazard rate function at that time. Further, we show that our theory includes, as a special case, the important class of neural hazard rate processes generated by recurrent neural networks (RNNs). To model such processes, we propose a dynamic deep recurrent survival analysis (DDRSA) architecture, introducing an RNN encoder into the static DRSA setting. Finally, we demonstrate RNN-based OTI policies with experiments and show that they outperform popular intervention methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- mRisk: Continuous Risk Estimation for Smoking Lapse from Noisy Sensor Data with Incomplete and Positive-Only LabelsMd. Azim Ullah, Soujanya Chatterjee, Christopher P. Fagundes, Cho Lam 等UbiComp 2022 · 被引用 2 次
- Bellman Meets Hawkes: Model-Based Reinforcement Learning via Temporal Point ProcessesChao Qu, Xiaoyu Tan, Siqiao Xue, Xiaoming Shi 等AAAI 2023 · 被引用 23 次
- A Multi-Channel Neural Graphical Event Model with Negative EvidenceTian Gao, Dharmashankar Subramanian, Karthikeyan Shanmugam, Debarun Bhattacharjya 等AAAI 2020 · 被引用 9 次
- Beyond Average Value Function in Precision Medicine: Maximum Probability-Driven Reinforcement Learning for Survival AnalysisJianqi Feng, Chengchun Shi, Zhenke Wu, Xiaodong Yan 等NeurIPS 2025 · 被引用 1 次
- Temporal Label Smoothing for Early Event PredictionHugo Yèche, Alizée Pace, Gunnar Rätsch, Rita KuznetsovaICML 2023 · 被引用 16 次
