ICML2026
FairSSL: Fair Multimodal Self-Supervised Learning
Jiaee Cheong, Abtin Mogharabin, Paul Pu Liang, Hatice Gunes, Sinan Kalkan
1 citation
Abstract
Early efforts on leveraging self-supervised learning (SSL) to improve machine learning (ML) fairness has proven promising. However, such an approach has yet to be explored within a multimodal context. Prior work has shown that, within a multimodal setting, different modalities contain modalityunique information that can complement information of other modalities. Leveraging on this, we propose a novel subjectlevel loss function to learn fairer representations via the following three mechanisms, adapting the variance-invariancecovariance regularization (VICReg) method: (i) the variance term, which reduces reliance on the protected attribute as a trivial solution; (ii) the invariance term, which ensures consistent predictions for similar individuals; and (iii) the covariance term, which minimizes correlational dependence on the protected attribute. Consequently, our loss function, coined as FAIRWELL, aims to obtain subject-independent representations, enforcing fairness in multimodal prediction tasks. We evaluate our method on three challenging real-world heterogeneous healthcare datasets (i.e. D-Vlog, MIMIC and MODMA) which contain different modalities of varying length and different prediction tasks. Our findings indicate that our framework improves overall fairness performance with minimal reduction in classification performance and significantly improves on the performance-fairness Pareto frontier. Code and trained models will be made available at: https://is.gd/FAIRWELL