When to Intervene: Learning Optimal Intervention Policies for Critical Events
Niranjan Damera Venkata, Chiranjib Bhattacharyya
Abstract
Providing a timely intervention before the onset of a critical event, such as a system failure, is of importance in many industrial settings. Before the onset of the critical event, systems typically exhibit behavioral changes which often manifest as stochastic co-variate observations which may be leveraged to trigger intervention. In this paper, for the first time, we formulate the problem of finding an optimally timed intervention (OTI) policy as minimizing the expected residual time to event, subject to a constraint on the probability of missing the event. Existing machine learning approaches to intervention on critical events focus on predicting event occurrence within a pre-defined window (a classification problem) or predicting time-to-event (a regression problem). Interventions are then triggered by setting model thresholds. These are heuristic-driven, lacking guarantees regarding optimality. To model the evolution of system behavior, we introduce the concept of a hazard rate process. We show that the OTI problem is equivalent to an optimal stopping problem on the associated hazard rate process. This key link has not been explored in literature. Under Markovian assumptions on the hazard rate process, we show that an OTI policy at any time can be analytically determined from the conditional hazard rate function at that time. Further, we show that our theory includes, as a special case, the important class of neural hazard rate processes generated by recurrent neural networks (RNNs). To model such processes, we propose a dynamic deep recurrent survival analysis (DDRSA) architecture, introducing an RNN encoder into the static DRSA setting. Finally, we demonstrate RNN-based OTI policies with experiments and show that they outperform popular intervention methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e9f0dbc8-28cd-4e7d-ac7c-e8fe4480d8a3Cited by top-tier papers1
Ask how each one uses itRelated papers
- mRisk: Continuous Risk Estimation for Smoking Lapse from Noisy Sensor Data with Incomplete and Positive-Only LabelsMd. Azim Ullah, Soujanya Chatterjee, Christopher P. Fagundes, Cho Lam et al.UbiComp 2022 · 2 citations
- Bellman Meets Hawkes: Model-Based Reinforcement Learning via Temporal Point ProcessesChao Qu, Xiaoyu Tan, Siqiao Xue, Xiaoming Shi et al.AAAI 2023 · 23 citations
- A Multi-Channel Neural Graphical Event Model with Negative EvidenceTian Gao, Dharmashankar Subramanian, Karthikeyan Shanmugam, Debarun Bhattacharjya et al.AAAI 2020 · 9 citations
- Beyond Average Value Function in Precision Medicine: Maximum Probability-Driven Reinforcement Learning for Survival AnalysisJianqi Feng, Chengchun Shi, Zhenke Wu, Xiaodong Yan et al.NeurIPS 2025 · 1 citation
- Temporal Label Smoothing for Early Event PredictionHugo Yèche, Alizée Pace, Gunnar Rätsch, Rita KuznetsovaICML 2023 · 16 citations
