Yet Another ICU Benchmark: A Flexible Multi-Center Framework for Clinical ML
Robin Van De Water, Hendrik Schmidt, Paul W. G. Elbers, Patrick Thoral, Bert Arnrich, Patrick Rockenschaub
摘要
Medical applications of machine learning (ML) have experienced a surge in popularity in recent years. The intensive care unit (ICU) is a natural habitat for ML given the abundance of available data from electronic health records. Models have been proposed to address numerous ICU prediction tasks like the early detection of complications. While authors frequently report state-of-the-art performance, it is challenging to verify claims of superiority. Datasets and code are not always published, and cohort definitions, preprocessing pipelines, and training setups are difficult to reproduce. This work introduces Yet Another ICU Benchmark (YAIB), a modular framework that allows researchers to define reproducible and comparable clinical ML experiments; we offer an end-to-end solution from cohort definition to model evaluation. The framework natively supports most open-access ICU datasets (MIMIC III/IV, eICU, HiRID, AUMCdb) and is easily adaptable to future ICU datasets. Combined with a transparent preprocessing pipeline and extensible training code for multiple ML and deep learning models, YAIB enables unified model development. Our benchmark comes with five predefined established prediction tasks (mortality, acute kidney injury, sepsis, kidney function, and length of stay) developed in collaboration with clinicians. Adding further tasks is straightforward by design. Using YAIB, we demonstrate that the choice of dataset, cohort definition, and preprocessing have a major impact on the prediction performance - often more so than model class - indicating an urgent need for YAIB as a holistic benchmarking tool. We provide our work to the clinical ML community to accelerate method development and enable real-world clinical implementations. Software Repository: https://github.com/rvandewater/YAIB.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- FEDKIM: Adaptive Federated Knowledge Injection into Medical Foundation ModelsXiaochen Wang, Jiaqi Wang, Houping Xiao, Jinghui Chen 等EMNLP 2024 · 被引用 6 次
- Can we generate portable representations for clinical time series data using LLMs?Zongliang Ji, Yifei Sun, Andre Carlos Kajdacsy-Balla Amaral, Anna Goldenberg 等ICLR 2026 · 被引用 5 次
- PyHealth 2.0: A Comprehensive Open-Source Toolkit for Accessible and Reproducible Clinical Deep LearningJohn Wu, Yongda Fan, Zhenbang Wu, Paul Landes 等ICML 2026 · 被引用 1 次
- ACES: Automatic Cohort Extraction System for Event-Stream DatasetsJustin Xu, Jack Gallifant, Alistair E. W. Johnson, Matthew B. A. McDermottICLR 2025
- Benchmarking Reinforcement Learning Algorithms for ICU Ventilator Settings: An Interpretable and Probabilistic Patient Environment for Doctor AgentsYa-Hsi Chang, Po-Chih KuoAAAI 2026
它引用的顶会 Paper4
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- DiffWave: A Versatile Diffusion Model for Audio SynthesisZhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao 等ICLR 2021 · 被引用 1,902 次
- CSDI: Conditional Score-based Diffusion Models for Probabilistic Time Series ImputationYusuke Tashiro, Jiaming Song, Yang Song, Stefano ErmonNeurIPS 2021 · 被引用 1,245 次
- HyperImpute: Generalized Iterative Imputation with Automatic Model SelectionDaniel Jarrett, Bogdan Cebere, Tennison Liu, Alicia Curth 等ICML 2022 · 被引用 129 次
相关 Paper
- Distilling Knowledge from Publicly Available Online EMR Data to Emerging Epidemic for PrognosisLiantao Ma, Xinyu Ma, Junyi Gao, Xianfeng Jiao 等WWW 2021 · 被引用 32 次
- CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation ModelsWei Dai, Peilin Chen, Malinda Lu, Daniel Li 等ICML 2025
- REACT-LLM: A Benchmark for Evaluating LLM Integration with Causal Features in Clinical Prognostic TasksLinna Wang, Zhixuan You, Qihui Zhang, Jiunan Wen 等AAAI 2026
- Clairvoyance: A Pipeline Toolkit for Medical Time SeriesDaniel Jarrett, Jinsung Yoon, Ioana Bica, Zhaozhi Qian 等ICLR 2021 · 被引用 43 次
- Reasoning-Enhanced Healthcare Predictions with Knowledge Graph Community RetrievalPengcheng Jiang, Cao Xiao, Minhao Jiang, Parminder Bhatia 等ICLR 2025
