REDS: Rule Extraction for Discovering Scenarios
Vadim Arzamasov, Klemens Böhm
Abstract
Scenario discovery is the process of finding areas of interest, known as scenarios, in data spaces resulting from simulations. For instance, one might search for conditions, i.e., inputs of the simulation model, where the system is unstable. Subgroup discovery methods are commonly used for scenario discovery. They find scenarios in the form of hyperboxes, which are easy to comprehend. Given a computational budget, results tend to get worse as the number of inputs of the simulation model and the cost of simulations increase. We propose a new procedure for scenario discovery from few simulations, dubbed REDS. A key ingredient is using an intermediate machine learning model to label data for subsequent use by conventional subgroup discovery methods. We provide statistical arguments why this is an improvement. In our experiments, REDS reduces the number of simulations required by 50--75% on average, depending on the quality measure. It is also useful as a semi-supervised subgroup discovery method and for discovering better scenarios from third-party data, when a simulation model is not available.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c8b69163-24f7-4338-b236-91e2b92e63daCited by top-tier papers2
- Fast Search-By-Classification for Large-Scale Databases Using Index-Aware Decision Trees and Random ForestsChristian Lülf, Denis Mayr Lima Martins, Marcos Antonio Vaz Salles, Yongluan Zhou et al.VLDB 2023 · 5 citations
- Subgroup Discovery with Small and Alternative Feature SetsJakob BachSIGMOD 2025 · 4 citations
Builds on1
Related papers
- Subgroup Discovery with the Cox ModelZachary Izzo, Iain MelvinICML 2026
- DISCES: Systematic Discovery of Event Stream QueriesRebecca Sattler, Sarah Kleest-Meißner, Steven Lange, Markus L. Schmid et al.SIGMOD 2025 · 3 citations
- "What makes my queries slow?": Subgroup Discovery for SQL Workload AnalysisYoucef Remil, Anes Bendimerad, Romain Mathonat, Philippe Chaleat et al.ASE 2021 · 12 citations
- Learning Subgroups with Maximum Treatment Effects Without Causal HeuristicsLincen Yang, Zhong Li, Matthijs van Leeuwen, Saber SalehkaleybarAAAI 2026
- MARIOH: Multiplicity-Aware Hypergraph ReconstructionKyuhan Lee, Geon Lee, Kijung ShinICDE 2025
