COFFEE: Counterfactual Fairness for Personalized Text Generation in Explainable Recommendation
Nan Wang, Qifan Wang, Yi-Chia Wang, Maziar Sanjabi, Jingzhou Liu, Hamed Firooz, Hongning Wang, Shaoliang Nie
Abstract
As language models become increasingly integrated into our digital lives, Personalized Text Generation (PTG) has emerged as a pivotal component with a wide range of applications. However, the bias inherent in user written text, often used for PTG model training, can inadvertently associate different levels of linguistic quality with users' protected attributes. The model can inherit the bias and perpetuate inequality in generating text w.r.t. users' protected attributes, leading to unfair treatment when serving users. In this work, we investigate fairness of PTG in the context of personalized explanation generation for recommendations. We first discuss the biases in generated explanations and their fairness implications. To promote fairness, we introduce a general framework to achieve measure-specific counterfactual fairness in explanation generation. Extensive experiments and human evaluations demonstrate the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1c42c74d-4d8d-464a-9c99-68287af9cb5eCited by top-tier papers2
- Counterfactually Fair RepresentationZhiqun Zuo, Mahdi Khalili, Xueru ZhangNeurIPS 2023 · 17 citations
- Metrics for What, Metrics for Whom: Assessing Actionability of Bias Evaluation Metrics in NLPPieter Delobelle, Giuseppe Attanasio, Debora Nozza, Su Lin Blodgett et al.EMNLP 2024 · 4 citations
Builds on8
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung et al.ICLR 2020 · 1,166 citations
- Towards Understanding and Mitigating Social Biases in Language ModelsPaul Pu Liang, Chiyu Wu, Louis-Philippe Morency, Ruslan SalakhutdinovICML 2021 · 495 citations
- Disentangling User Interest and Conformity for Recommendation with Causal EmbeddingYu Zheng, Chen Gao, Xiang Li, Xiangnan He et al.WWW 2021 · 392 citations
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
Related papers
- Personalized Transformer for Explainable RecommendationLei Li, Yongfeng Zhang, Li ChenACL 2021
- Towards Personalized Fairness based on Causal NotionYunqi Li, Hanxiong Chen, Shuyuan Xu, Yingqiang Ge et al.SIGIR 2021 · 139 citations
- Constructing Fair Latent Space for Intersection of Fairness and ExplainabilityHyungjun Joo, Hyeonggeun Han, Sehwan Kim, Sangwoo Hong et al.AAAI 2025 · 2 citations
- Can LLMs Enhance Fairness in Recommendation Systems? A Data Augmentation ApproachHanzhe Li, Dazhong Shen, Chao Wang, Yuting Liu et al.SIGIR 2025 · 2 citations
- ReXPlug: Explainable Recommendation using Plug-and-Play Language ModelDeepesh V. Hada, Vijaikumar M, Shirish K. ShevadeSIGIR 2021 · 55 citations
