PyHealth 2.0: A Comprehensive Open-Source Toolkit for Accessible and Reproducible Clinical Deep Learning
John Wu, Yongda Fan, Zhenbang Wu, Paul Landes, Eric Schrock, Sayeed Sajjad Razin, Arjun Chatterjee, Naveen Baskaran, Joshua Steier, Andrea Fitzpatrick, Bilal Arif, Rian Atri
摘要
Difficulty replicating baselines, high computational costs, and required domain expertise create persistent barriers to clinical AI research. To address these challenges, we introduce PyHealth 2.0, an enhanced clinical deep learning toolkit that enables predictive modeling in as few as 7 lines of code. PyHealth 2.0 offers three key contributions: (1) a comprehensive toolkit addressing reproducibility and compatibility challenges by unifying 15+ datasets, 20+ clinical tasks, 25+ models, 5+ interpretability methods, and 5+ uncertainty quantification methods within a single framework that supports diverse clinical data modalities-signals, text, imaging, and electronic health records-with translation of 5+ medical coding standards; (2) accessibility-focused design accommodating multimodal data and diverse computational resources with up to 39× faster processing and 20× lower memory usage, enabling work from 16GB laptops to production systems; and (3) an active open-source community of 400+ members lowering domain expertise barriers through extensive documentation, reproducible research contributions, and collaborations with academic health systems and industry partners, including multi-language support via RHealth. PyHealth 2.0 establishes an open-source foundation and community advancing accessible, reproducible healthcare AI. Project details at https://pyhealth.dev/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- StageNet: Stage-Aware Neural Networks for Health Risk PredictionJunyi Gao, Cao Xiao, Yasha Wang, Wen Tang 等WWW 2020 · 被引用 131 次
- GRASP: Generic Framework for Health Status Representation Learning Based on Incorporating Knowledge from Similar PatientsChaohe Zhang, Xin Gao, Liantao Ma, Yasha Wang 等AAAI 2021 · 被引用 77 次
- Yet Another ICU Benchmark: A Flexible Multi-Center Framework for Clinical MLRobin Van De Water, Hendrik Schmidt, Paul W. G. Elbers, Patrick Thoral 等ICLR 2024 · 被引用 43 次
- CoDrug: Conformal Drug Property Prediction with Density Estimation under Covariate ShiftSiddhartha Laghuvarapu, Zhen Lin, Jimeng SunNeurIPS 2023 · 被引用 14 次
- SCRIB: Set-Classifier with Class-Specific Risk Bounds for Blackbox ModelsZhen Lin, Lucas Glass, M. Brandon Westover, Cao Xiao 等AAAI 2022 · 被引用 14 次
相关 Paper
- CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation ModelsWei Dai, Peilin Chen, Malinda Lu, Daniel Li 等ICML 2025
- PyTDC: A multimodal machine learning training, evaluation, and inference platform for biomedical foundation modelsAlejandro Velez-Arce, Marinka ZitnikICML 2025
- Measuring Cross-Modal Interactions in Multimodal ModelsLaura Wenderoth, Konstantin Hemker, Nikola Simidjievski, Mateja JamnikAAAI 2025 · 被引用 12 次
- DeepCoDA: personalized interpretability for compositional health dataThomas P. Quinn, Dang Nguyen, Santu Rana, Sunil Gupta 等ICML 2020 · 被引用 15 次
- VBridge: Connecting the Dots Between Features and Data to Explain Healthcare ModelsFurui Cheng, Dongyu Liu, Fan Du, Yanna Lin 等IEEE VIS 2021 · 被引用 54 次
