TCD-Arena: Assessing Robustness of Time Series Causal Discovery Methods Against Assumption Violations
Gideon Stein, Niklas Penzel, Tristan Piater, Joachim Denzler
Abstract
Causal Discovery (CD) is a powerful framework for scientific inquiry. Yet, its practical adoption is hindered by a reliance on strong, often unverifiable assumptions and a lack of robust performance assessment. To address these limitations and advance empirical CD evaluation, we present TCD-Arena, a modularized, highly customizable, and extendable testing kit to assess the robustness of time series CD algorithms against stepwise more severe assumption violations. For demonstration, we conduct an extensive empirical study comprising around 30 million individual CD attempts and reveal nuanced robustness profiles for 33 distinct assumption violations. Further, we investigate CD ensembles and find that they have the potential to improve general robustness, which has implications for real-world applications. With this, we strive to ultimately facilitate the development of CD methods that are reliable for a diverse range of synthetic and potentially real-world data conditions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 97dcf15a-835c-4feb-864c-5264213e40cbBuilds on16
- Ensemble of Averages: Improving Model Selection and Boosting Performance in Domain GeneralizationDevansh Arpit, Huan Wang, Yingbo Zhou, Caiming XiongNeurIPS 2022 · 232 citations
- Beware of the Simulated DAG! Causal Discovery Benchmarks May Be Easy to GameAlexander G. Reisach, Christof Seiler, Sebastian WeichwaldNeurIPS 2021 · 213 citations
- High-recall causal discovery for autocorrelated time series with latent confoundersAndreas Gerhardus, Jakob RungeNeurIPS 2020 · 159 citations
- CausalTime: Realistically Generated Time-series for Benchmarking of Causal DiscoveryYuxiao Cheng, Ziqian Wang, Tingxiong Xiao, Qin Zhong et al.ICLR 2024 · 35 citations
- Assumption violations in causal discovery and the robustness of score matchingFrancesco Montagna, Atalanti-Anastasia Mastakouri, Elias Eulig, Nicoletta Noceti et al.NeurIPS 2023 · 35 citations
Related papers
- The robustness of differentiable Causal Discovery in misspecified ScenariosHuiyang Yi, Yanyan He, Duxin Chen, Mingyu Kang et al.ICLR 2025
- CausalRivers - Scaling up benchmarking of causal discovery for real-world time-seriesGideon Stein, Maha Shadaydeh, Jan Blunk, Niklas Penzel et al.ICLR 2025
- Discovering Mixtures of Structural Causal Models from Time Series DataSumanth Varambally, Yian Ma, Rose YuICML 2024 · 11 citations
- Causal Discovery in the Wild: A Voting-Theoretic Ensemble ApproachVy Vo, Haoxuan Li, Mingming GongICLR 2026
- Robust Causal Discovery in Real-World Time Series with Power-LawsMatteo Tusoni, Giuseppe Masi, Andrea Coletta, Aldo Glielmo et al.ICML 2026
