OSF: On Pre-training and Scaling of Sleep Foundation Models
Zitao Shuai, Zongzhe Xu, David Yang, Wei Wang, Yuzhe Yang
Abstract
Polysomnography (PSG) provides the gold standard for sleep assessment but suffers from substantial heterogeneity across recording devices and cohorts. There have been growing efforts to build general-purpose foundation models (FMs) for sleep physiology, but lack an in-depth understanding of the pre-training process and scaling patterns that lead to more generalizable sleep FMs. To fill this gap, we curate a massive corpus of 166,500 hours of sleep recordings from nine public sources and establish SleepBench, a comprehensive, fully open-source benchmark. Leveraging SleepBench, we systematically evaluate four families of self-supervised pre-training objectives and uncover three critical findings: (1) existing FMs fail to generalize to missing channels at inference; (2) channel-invariant feature learning is essential for pre-training; and (3) scaling sample size, model capacity, and multi-source data mixture consistently improves downstream performance. With an enhanced pre-training and scaling recipe, we introduce OSF, a family of sleep FMs that achieves state-of-the-art performance across nine datasets on diverse sleep and disease prediction tasks. Further analysis of OSF also reveals intriguing properties in sample efficiency, hierarchical aggregation, and cross-dataset scaling. Codes are available at: https://github.com/yang-ai-lab/OSF-Open-Sleep-FM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e2b65fce-ed5e-4e6b-ab29-fce25d7da9cdBuilds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- TS2Vec: Towards Universal Representation of Time SeriesZhihan Yue, Yujing Wang, Juanyong Duan, Tianmeng Yang et al.AAAI 2022 · 938 citations
- Guiding Masked Representation Learning to Capture Spatio-Temporal Relationship of ElectrocardiogramYeongyeon Na, Minje Park, Yunwon Tae, Sunghoon JooICLR 2024 · 92 citations
- This Time is Different: An Observability Perspective on Time Series Foundation ModelsBen Cohen, Emaad Khwaja, Youssef Doubli, Salahidine Lemaachi et al.NeurIPS 2025 · 68 citations
Related papers
- SleepMaMi: A Universal Sleep Foundation Model for Integrating Macro- and Micro-structuresKeondo Park, Younghoon Na, Yourim Choi, Hyunwoo Ryu et al.ICML 2026
- SleepLM: Natural-Language Intelligence for Human SleepZongzhe Xu, Zitao Shuai, Eideen Mozaffari, Ravi Aysola et al.ICML 2026 · 10 citations
- SleepFM: Multi-modal Representation Learning for Sleep Across Brain Activity, ECG and Respiratory SignalsRahul Thapa, Bryan He, Magnus Ruud Kjær, Hyatt E. Moore IV et al.ICML 2024 · 48 citations
- EEG-FM-Bench: A Comprehensive Benchmark for the Systematic Evaluation and Diagnostic Analyses of EEG Foundation ModelsWei Xiong, Jiangtong Li, Jie Li, Kun Zhu et al.ICML 2026 · 15 citations
- sleep2vec: Unified Cross-Modal Alignment for Heterogeneous Nocturnal BiosignalsWeixuan Yuan, Zengrui Jin, Yichen Wang, Donglin Xie et al.ICLR 2026 · 4 citations
