Generate Your Counterfactuals: Towards Controlled Counterfactual Generation for Text
Nishtha Madaan, Inkit Padhi, Naveen Panwar, Diptikalyan Saha
Abstract
Machine Learning has seen tremendous growth recently, which has led to a larger adaptation of ML systems for educational assessments, credit risk, healthcare, employment, criminal justice, to name a few. The trustworthiness of ML and NLP systems is a crucial aspect and requires a guarantee that the decisions they make are fair and robust. Aligned with this, we propose a novel framework GYC, to generate a set of exhaustive counterfactual text, which are crucial for testing these ML systems. Our main contributions include a) We introduce GYC, a framework to generate counterfactual samples such that the generation is plausible, diverse, goal-oriented, and effective, b) We generate counterfactual samples, that can direct the generation towards a corresponding condition such as named-entity tag, semantic role label, or sentiment. Our experimental results on various domains show that GYC generates counterfactual text samples exhibiting the above four properties. GYC generates counterfactuals that can act as test cases to evaluate a model and any text debiasing algorithm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers28
- A Causal Lens for Controllable Text GenerationZhiting Hu, Li Erran LiNeurIPS 2021 · 77 citations
- Towards Understanding the Working Mechanism of Text-to-Image Diffusion ModelMingyang Yi, Aoxue Li, Yi Xin, Zhenguo LiNeurIPS 2024 · 55 citations
- MT-Teql: Evaluating and Augmenting Neural NLIDB on Real-world Linguistic and Schema VariationsPingchuan Ma, Shuai WangVLDB 2022 · 38 citations
- The King Is Naked: On the Notion of Robustness for Natural Language ProcessingEmanuele La Malfa, Marta KwiatkowskaAAAI 2022 · 31 citations
- DISCO: Distilling Counterfactuals with Large Language ModelsZeming Chen, Qiyue Gao, Antoine Bosselut, Ashish Sabharwal et al.ACL 2023 · 27 citations
Builds on3
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung et al.ICLR 2020 · 1,166 citations
- Beyond Accuracy: Behavioral Testing of NLP Models with CheckListMarco Túlio Ribeiro, Tongshuang Wu, Carlos Guestrin, Sameer SinghACL 2020 · 51 citations
Related papers
- Counterfactual Inference for Text Classification DebiasingChen Qian, Fuli Feng, Lijie Wen, Chunping Ma et al.ACL 2021
- Do LLMs Behave as Claimed? Investigating How LLMs Follow Their Own Claims using Counterfactual QuestionsHaochen Shi, Shaobo Li, Guoqing Chao, Xiaoliang Shi et al.EMNLP 2025
- GeCo: Quality Counterfactual Explanations in Real TimeMaximilian Schleich, Zixuan Geng, Yihong Zhang, Dan SuciuVLDB 2021 · 77 citations
- COFFEE: Counterfactual Fairness for Personalized Text Generation in Explainable RecommendationNan Wang, Qifan Wang, Yi-Chia Wang, Maziar Sanjabi et al.EMNLP 2023 · 2 citations
- CREST: A Joint Framework for Rationalization and Counterfactual Text GenerationMarcos V. Treviso, Alexis Ross, Nuno Miguel Guerreiro, André F. T. MartinsACL 2023 · 7 citations
