Linguistic Dependencies and Statistical Dependence
Jacob Louis Hoover, Wenyu Du, Alessandro Sordoni, Timothy J. O'Donnell
摘要
Are pairs of words that tend to occur together also likely to stand in a linguistic dependency? This empirical question is motivated by a long history of literature in cognitive science, psycholinguistics, and NLP. In this work we contribute an extensive analysis of the relationship between linguistic dependencies and statistical dependence between words. Improving on previous work, we introduce the use of large pretrained language models to compute contextualized estimates of the pointwise mutual information between words (CPMI). For multiple models and languages, we extract dependency trees which maximize CPMI, and compare to gold standard linguistic dependencies. Overall, we find that CPMI dependencies achieve an unlabelled undirected attachment score of at most ≈ 0.5. While far above chance, and consistently above a non-contextualized PMI baseline, this score is generally comparable to a simple baseline formed by connecting adjacent words. We analyze which kinds of linguistic dependencies are best captured in CPMI dependencies, and also find marked differences between the estimates of the large pretrained language models, illustrating how their different training schemes affect the type of dependencies they capture.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Constructions are Revealed in Word DistributionsJoshua Rozner, Leonie Weissweiler, Kyle Mahowald, Cory ShainEMNLP 2025 · 被引用 8 次
- Similarity-weighted Construction of Contextualized Commonsense Knowledge Graphs for Knowledge-intense Argumentation TasksMoritz Plenz, Juri Opitz, Philipp Heinisch, Philipp Cimiano 等ACL 2023 · 被引用 4 次
- Syntactic Substitutability as Unsupervised Dependency SyntaxJasper Jian, Siva ReddyEMNLP 2023
它引用的顶会 Paper4
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Perturbed Masking: Parameter-free Probing for Analyzing and Interpreting BERTZhiyong Wu, Yun Chen, Ben Kao, Qun LiuACL 2020 · 被引用 158 次
- Are Pre-trained Language Models Aware of Phrases? Simple but Strong Baselines for Grammar InductionTaeuk Kim, Jihun Choi, Daniel Edmiston, Sang-goo LeeICLR 2020 · 被引用 92 次
- Exploiting Syntactic Structure for Better Language Modeling: A Syntactic Distance ApproachWenyu Du, Zhouhan Lin, Yikang Shen, Timothy J. O'Donnell 等ACL 2020 · 被引用 15 次
相关 Paper
- PMIScore: An Unsupervised Approach to Quantify Dialogue EngagementYongkang Guo, Zhihuan Huang, Yuqing KongWWW 2026
- Neural Methods for Point-wise Dependency EstimationYao-Hung Hubert Tsai, Han Zhao, Makoto Yamada, Louis-Philippe Morency 等NeurIPS 2020 · 被引用 41 次
- An Information-theoretical Framework for Understanding Out-of-distribution Detection with Pretrained Vision-Language ModelsBo Peng, Jie Lu, Guangquan Zhang, Zhen FangNeurIPS 2025 · 被引用 9 次
- Language models and brains align due to more than next-word prediction and word-level informationGabriele Merlin, Mariya TonevaEMNLP 2024 · 被引用 2 次
- Fine-grained Analysis of Brain-LLM Alignment through Input AttributionMichela Proietti, Roberto Capobianco, Mariya TonevaICML 2026
