Learning Disentangled Textual Representations via Statistical Measures of Similarity
Pierre Colombo, Guillaume Staerman, Nathan Noiry, Pablo Piantanida
Abstract
When working with textual data, a natural application of disentangled representations is the fair classification where the goal is to make predictions without being biased (or influenced) by sensible attributes that may be present in the data (e.g., age, gender or race). Dominant approaches to disentangle a sensitive attribute from textual representations rely on learning simultaneously a penalization term that involves either an adversary loss (e.g., a discriminator) or an information measure (e.g., mutual information). However, these methods require the training of a deep neural network with several parameter updates for each update of the representation model. As a matter of fact, the resulting nested optimization loop is both times consuming, adding complexity to the optimization dynamic, and requires a fine hyperparameter selection (e.g., learning rates, architecture). In this work, we introduce a family of regularizers for learning disentangled representations that do not require training. These regularizers are based on statistical measures of similarity between the conditional probability distributions with respect to the sensible attributes. Our novel regularizers do not require additional training, are faster and do not involve additional tuning while achieving better results both when combined with pretrained and randomly initialized text encoders.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 00a38e00-4215-48c2-b48d-fa0dac402ed6Cited by top-tier papers6
- Beyond Mahalanobis Distance for Textual OOD DetectionPierre Colombo, Eduardo Dadalto Câmara Gomes, Guillaume Staerman, Nathan Noiry et al.NeurIPS 2022 · 24 citations
- Enhancing Feature Diversity Boosts Channel-Adaptive Vision TransformersChau Pham, Bryan A. PlummerNeurIPS 2024 · 15 citations
- TACIT: A Target-Agnostic Feature Disentanglement Framework for Cross-Domain Text ClassificationRui Song, Fausto Giunchiglia, Yingji Li, Mingjie Tian et al.AAAI 2024 · 10 citations
- Transductive Learning for Textual Few-Shot Classification in API-based Embedding ModelsPierre Colombo, Victor Pellegrain, Malik Boudiaf, Myriam Tami et al.EMNLP 2023 · 7 citations
- Hypothesis Transfer Learning with Surrogate Classification Losses: Generalization Bounds through Algorithmic StabilityAnass Aghbalou, Guillaume StaermanICML 2023 · 3 citations
Builds on13
- CLUB: A Contrastive Log-ratio Upper Bound of Mutual InformationPengyu Cheng, Weituo Hao, Shuyang Dai, Jiachang Liu et al.ICML 2020 · 512 citations
- Towards Understanding and Mitigating Social Biases in Language ModelsPaul Pu Liang, Chiyu Wu, Louis-Philippe Morency, Ruslan SalakhutdinovICML 2021 · 495 citations
- Understanding the Limitations of Variational Mutual Information EstimatorsJiaming Song, Stefano ErmonICLR 2020 · 243 citations
- Guiding Attention in Sequence-to-Sequence Models for Dialogue Act PredictionPierre Colombo, Emile Chapuis, Matteo Manica, Emmanuel Vignon et al.AAAI 2020 · 69 citations
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
Related papers
- A Novel Estimator of Mutual Information for Learning to Disentangle Textual RepresentationsPierre Colombo, Pablo Piantanida, Chloé ClavelACL 2021
- Learning Disentangled Representation for Fair Facial Attribute Classification via Fairness-aware Information AlignmentSungho Park, Sunhee Hwang, Dohyung Kim, Hyeran ByunAAAI 2021 · 68 citations
- Fair Representation Learning: An Alternative to Mutual InformationJi Liu, Zenan Li, Yuan Yao, Feng Xu et al.KDD 2022 · 14 citations
- Fair Text Classification with Wasserstein IndependenceThibaud Leteno, Antoine Gourru, Charlotte Laclau, Rémi Emonet et al.EMNLP 2023 · 3 citations
- Counterfactual Fairness with Disentangled Causal Effect Variational AutoencoderHyemi Kim, Seungjae Shin, JoonHo Jang, Kyungwoo Song et al.AAAI 2021 · 72 citations
