Improving Commonsense Causal Reasoning by Adversarial Training and Data Augmentation
Ieva Staliunaite, Philip John Gorinski, Ignacio Iacobacci
Abstract
Determining the plausibility of causal relations between clauses is a commonsense reasoning task that requires complex inference ability. The general approach to this task is to train a large pretrained language model on a specific dataset. However, the available training data for the task is often scarce, which leads to instability of model training or reliance on the shallow features of the dataset. This paper presents a number of techniques for making models more robust in the domain of causal reasoning. Firstly, we perform adversarial training by generating perturbed inputs through synonym substitution. Secondly, based on a linguistic theory of discourse connectives, we perform data augmentation using a discourse parser for detecting causally linked clauses in large text, and a generative language model for generating distractors. Both methods boost model performance on the Choice of Plausible Alternatives (COPA) dataset, as well as on a Balanced COPA dataset, which is a modified version of the original data that has been developed to avoid superficial cues, leading to a more challenging benchmark. We show a statistically significant improvement in performance and robustness on both datasets, even with only a small number of additionally generated data points.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 17d7eb56-a2cc-483d-924c-521ead48bf24Cited by top-tier papers5
- ROCK: Causal Inference Principles for Reasoning about Commonsense CausalityJiayao Zhang, Hongming Zhang, Weijie J. Su, Dan RothICML 2022 · 28 citations
- LogiGAN: Learning Logical Reasoning via Adversarial Pre-trainingXinyu Pi, Wanjun Zhong, Yan Gao, Nan Duan et al.NeurIPS 2022 · 19 citations
- COLA: Contextualized Commonsense Causal Reasoning from the Causal Inference PerspectiveZhaowei Wang, Quyet V. Do, Hongming Zhang, Jiayao Zhang et al.ACL 2023 · 8 citations
- Connective Prediction for Implicit Discourse Relation Recognition via Knowledge DistillationHongyi Wu, Hao Zhou, Man Lan, Yuanbin Wu et al.ACL 2023 · 7 citations
- Expert-guided Clinical Text Augmentation via Query-Based Model CollaborationDongkyu Cho, Miao Zhang, Gregory Lyng, Rumi ChunaraICML 2026
Builds on4
- Word-level Textual Adversarial Attacking as Combinatorial OptimizationYuan Zang, Fanchao Qi, Chenghao Yang, Zhiyuan Liu et al.ACL 2020 · 188 citations
- Pretraining with Contrastive Sentence Objectives Improves Discourse Performance of Language ModelsDan Iter, Kelvin Guu, Larry Lansing, Dan JurafskyACL 2020 · 72 citations
- Pre-training Is (Almost) All You Need: An Application to Commonsense ReasoningAlexandre Tamborrino, Nicola Pellicanò, Baptiste Pannier, Pascal Voitot et al.ACL 2020 · 30 citations
- Unsupervised Commonsense Question Answering with Self-TalkVered Shwartz, Peter West, Ronan Le Bras, Chandra Bhagavatula et al.EMNLP 2020 · 25 citations
Related papers
- DiscoSense: Commonsense Reasoning with Discourse ConnectivesPrajjwal Bhargava, Vincent NgEMNLP 2022 · 1 citation
- XCOPA: A Multilingual Dataset for Causal Commonsense ReasoningEdoardo Maria Ponti, Goran Glavas, Olga Majewska, Qianchu Liu et al.EMNLP 2020 · 6 citations
- DISCO: Distilling Counterfactuals with Large Language ModelsZeming Chen, Qiyue Gao, Antoine Bosselut, Ashish Sabharwal et al.ACL 2023 · 27 citations
- Dually Self-Improved Counterfactual Data Augmentation Using Large Language ModelLuhao Zhang, Xinyu Zhang, Linmei Hu, Dandan Song et al.ACL 2025 · 1 citation
- CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language ModelsYuefei Chen, Vivek K. Singh, Jing Ma, Ruixiang TangAAAI 2026 · 1 citation
