FADES: Fair Disentanglement with Sensitive Relevance
Taeuk Jang, Xiaoqian Wang
Abstract
Learning fair representation in deep learning is essential to mitigate discriminatory outcomes and enhance trustworthiness. However, previous research has been commonly established on inappropriate assumptions prone to unrealistic counterfactuals and performance degradation. Although some proposed alternative approaches, such as employing correlation-aware causal graphs or proxies for mutual information, these methods are less practical and not applicable in general. In this work, we propose FAir DisEntanglement with Sensitive relevance (FADES), a novel approach that leverages conditional mutual information from the information theory perspective to address these challenges. We employ sensitive relevant code to direct correlated information between target labels and sensitive attributes by imposing conditional independence, allowing better separation of the features of interest in the latent space. Utilizing an intuitive disentangling approach, FADES consistently achieves superior performance and fairness both quantitatively and qualitatively with its straightforward structure. Specifically, the proposed method outperforms existing works in downstream classification and counterfactual generations on various benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d71edbd-5255-45f4-abbd-bcd07854dd7eCited by top-tier papers2
- CAD-VAE: Leveraging Correlation-Aware Latents for Comprehensive Fair DisentanglementChenrui Ma, Xi Xiao, Tianyang Wang, Xiao Wang et al.AAAI 2026 · 9 citations
- On the Alignment between Fairness and Accuracy: from the Perspective of Adversarial RobustnessJunyi Chai, Taeuk Jang, Jing Gao, Xiaoqian WangICML 2025
Builds on13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or et al.ICCV 2021 · 1,437 citations
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan et al.ICML 2021 · 683 citations
- Learning from Failure: De-biasing Classifier from Biased ClassifierJun Hyun Nam, Hyuntak Cha, Sungsoo Ahn, Jaeho Lee et al.NeurIPS 2020 · 428 citations
- Fairness without Demographics through Adversarially Reweighted LearningPreethi Lahoti, Alex Beutel, Jilin Chen, Kang Lee et al.NeurIPS 2020 · 406 citations
Related papers
- Counterfactual Fairness with Disentangled Causal Effect Variational AutoencoderHyemi Kim, Seungjae Shin, JoonHo Jang, Kyungwoo Song et al.AAAI 2021 · 72 citations
- Learning Fair Representation via Distributional Contrastive DisentanglementChangdae Oh, Heeji Won, Junhyuk So, Taero Kim et al.KDD 2022 · 29 citations
- Learning Disentangled Representation for Fair Facial Attribute Classification via Fairness-aware Information AlignmentSungho Park, Sunhee Hwang, Dohyung Kim, Hyeran ByunAAAI 2021 · 68 citations
- Fair Representation Learning: An Alternative to Mutual InformationJi Liu, Zenan Li, Yuan Yao, Feng Xu et al.KDD 2022 · 14 citations
- Scalable Infomin LearningYanzhi Chen, Weihao Sun, Yingzhen Li, Adrian WellerNeurIPS 2022 · 10 citations
