Causal Direction of Data Collection Matters: Implications of Causal and Anticausal Learning for NLP
Zhijing Jin, Julius von Kügelgen, Jingwei Ni, Tejas Vaidhya, Ayush Kaushal, Mrinmaya Sachan, Bernhard Schölkopf
摘要
The principle of independent causal mechanisms (ICM) states that generative processes of real world data consist of independent modules which do not influence or inform each other. While this idea has led to fruitful developments in the field of causal inference, it is not widely-known in the NLP community. In this work, we argue that the causal direction of the data collection process bears nontrivial implications that can explain a number of published NLP findings, such as differences in semi-supervised learning (SSL) and domain adaptation (DA) performance across different settings. We categorize common NLP tasks according to their causal direction and empirically assay the validity of the ICM principle for text data using minimum description length. We conduct an extensive meta-analysis of over 100 published SSL and 30 DA studies, and find that the results are consistent with our expectations based on causal insights. This work presents the first attempt to analyze the ICM principle in NLP, and provides constructive suggestions for future modeling choices. 1 * Equal contribution. 1 The codes are at https://github.com/zhijing-jin/icm4nlp .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- CauSSL: Causality-inspired Semi-supervised Learning for Medical Image SegmentationJuzheng Miao, Cheng Chen, Furui Liu, Hao Wei 等ICCV 2023 · 被引用 88 次
- A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language ModelsAlessandro Stolfo, Zhijing Jin, Kumar Shridhar, Bernhard Schölkopf 等ACL 2023 · 被引用 15 次
- Causal vs. Anticausal merging of predictorsSergio Hernan Garrido Mejia, Patrick Blöbaum, Bernhard Schölkopf, Dominik JanzingNeurIPS 2024 · 被引用 1 次
- Sequential Learning of Neural Networks for Prequential MDLJörg Bornschein, Yazhe Li, Marcus HutterICLR 2023 · 被引用 1 次
- Whose Boat Does it Float? Improving Personalization in Preference Tuning via Inferred User PersonasNishant Balepur, Vishakh Padmakumar, Fumeng Yang, Shi Feng 等ACL 2025
它引用的顶会 Paper6
- Statistical Power and Translationese in Machine Translation EvaluationYvette Graham, Barry Haddow, Philipp KoehnEMNLP 2020 · 被引用 82 次
- Discovering Fully Oriented Causal NetworksOsman Mian, Alexander Marx, Jilles VreekenAAAI 2021 · 被引用 37 次
- Information-Theoretic Probing with Minimum Description LengthElena Voita, Ivan TitovEMNLP 2020 · 被引用 34 次
- Text and Causal Inference: A Review of Using Text to Remove Confounding from Causal EstimatesKatherine A. Keith, David D. Jensen, Brendan O'ConnorACL 2020 · 被引用 16 次
- On The Evaluation of Machine Translation SystemsTrained With Back-TranslationSergey Edunov, Myle Ott, Marc'Aurelio Ranzato, Michael AuliACL 2020 · 被引用 15 次
相关 Paper
- Can Large Language Models Learn Independent Causal Mechanisms?Gaël Gendron, Bao Trung Nguyen, Alex Yuxuan Peng, Michael J. Witbrock 等EMNLP 2024 · 被引用 2 次
- Can Large Language Models Infer Causation from Correlation?Zhijing Jin, Jiarui Liu, Zhiheng Lyu, Spencer Poff 等ICLR 2024 · 被引用 186 次
- Diverse Distributions of Self-Supervised Tasks for Meta-Learning in NLPTrapit Bansal, Karthick Prasad Gunasekaran, Tong Wang, Tsendsuren Munkhdalai 等EMNLP 2021 · 被引用 27 次
- Cross-Lingual Transfer with Class-Weighted Language-Invariant RepresentationsRuicheng Xian, Heng Ji, Han ZhaoICLR 2022 · 被引用 5 次
- iTAG: Inverse Design for Natural Text Generation with Accurate Causal Graph AnnotationsWenshuo Wang, Boyu Cao, Nan Zhuang, Wei LiACL 2026
