Exploring Self-Supervised Representation Ensembles for COVID-19 Cough Classification
Hao Xue, Flora D. Salim
Abstract
The usage of smartphone-collected respiratory sound, trained with deep learning models, for detecting and classifying COVID-19 becomes popular recently. It removes the need for in-person testing procedures especially for rural regions where related medical supplies, experienced workers, and equipment are limited. However, existing sound-based diagnostic approaches are trained in a fully-supervised manner, which requires large scale well-labelled data. It is critical to discover new methods to leverage unlabelled respiratory data, which can be obtained more easily. In this paper, we propose a novel self-supervised learning enabled framework for COVID-19 cough classification. A contrastive pre-training phase is introduced to train a Transformer-based feature encoder with unlabelled data. Specifically, we design a random masking mechanism to learn robust representations of respiratory sounds. The pre-trained feature encoder is then fine-tuned in the downstream phase to perform cough classification. In addition, different ensembles with varied random masking rates are also explored in the downstream phase. Through extensive evaluations, we demonstrate that the proposed contrastive pre-training, the random masking mechanism, and the ensemble architecture contribute to improving cough classification performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4dedb614-4bc1-4dfb-8c40-29adcdc6afe8Cited by top-tier papers1
Ask how each one uses itBuilds on3
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Hierarchically Structured Transformer Networks for Fine-Grained Spatial Event ForecastingXian Wu, Chao Huang, Chuxu Zhang, Nitesh V. ChawlaWWW 2020 · 62 citations
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie et al.CVPR 2020
Related papers
- Learning Low-Rank Feature for Thorax Disease ClassificationYancheng Wang, Rajeev Goel, Utkarsh Nath, Alvin C. Silva et al.NeurIPS 2024 · 12 citations
- AudioMosaic: Contrastive Masked Audio Representation LearningHanxun Huang, Qizhou Wang, Xingjun Ma, Cihang Xie et al.ICML 2026 · 2 citations
- Context Matters: Graph-based Self-supervised Representation Learning for Medical ImagesLi Sun, Ke Yu, Kayhan BatmanghelichAAAI 2021 · 27 citations
- RF-HeartSSL: Self-Supervised Learning for RF-Based Cardiac SensingXinmeng Cai, Jinbo Chen, Guixin Xu, Haoyu Wang et al.UbiComp 2026 · 1 citation
- Self-supervised object detection from audio-visual correspondenceTriantafyllos Afouras, Yuki M. Asano, Francois Fagan, Andrea Vedaldi et al.CVPR 2022 · 50 citations
