An Analysis of the Utility of Explicit Negative Examples to Improve the Syntactic Abilities of Neural Language Models
Hiroshi Noji, Hiroya Takamura
摘要
We explore the utilities of explicit negative examples in training neural language models. Negative examples here are incorrect words in a sentence, such as barks in *The dogs barks. Neural language models are commonly trained only on positive examples, a set of sentences in the training data, but recent studies suggest that the models trained in this way are not capable of robustly handling complex syntactic constructions, such as long-distance agreement. In this paper, we first demonstrate that appropriately using negative examples about particular constructions (e.g., subject-verb agreement) will boost the model’s robustness on them in English, with a negligible loss of perplexity. The key to our success is an additional margin loss between the log-likelihoods of a correct word and an incorrect word. We then provide a detailed analysis of the trained models. One of our findings is the difficulty of object-relative clauses for RNNs. We find that even with our direct learning signals the models still suffer from resolving agreement across an object-relative clause. Augmentation of training sentences involving the constructions somewhat helps, but the accuracy still does not reach the level of subject-relative clauses. Although not directly cognitively appealing, our method can be a tool to analyze the true architectural limitation of neural models on challenging linguistic constructions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Cross-Linguistic Syntactic Evaluation of Word Prediction ModelsAaron Mueller, Garrett Nicolai, Panayiota Petrou-Zeniou, Natalia Talmina 等ACL 2020 · 被引用 2 次
- Learning with Instance Bundles for Reading ComprehensionDheeru Dua, Pradeep Dasigi, Sameer Singh, Matt GardnerEMNLP 2021 · 被引用 1 次
它引用的顶会 Paper2
相关 Paper
- NegatER: Unsupervised Discovery of Negatives in Commonsense Knowledge BasesTara Safavi, Jing Zhu, Danai KoutraEMNLP 2021 · 被引用 9 次
- Influence Paths for Characterizing Subject-Verb Number Agreement in LSTM Language ModelsKaiji Lu, Piotr Mardziel, Klas Leino, Matt Fredrikson 等ACL 2020 · 被引用 8 次
- CLINE: Contrastive Learning with Semantic Negative Examples for Natural Language UnderstandingDong Wang, Ning Ding, Piji Li, Haitao ZhengACL 2021
- Recurrent Neural Network Language Models Always Learn English-Like Relative Clause AttachmentForrest Davis, Marten van SchijndelACL 2020 · 被引用 15 次
- Negative Pre-activations Differentiate SyntaxLinghao Kong, Angelina Ning, Micah Adler, Nir N ShavitICLR 2026 · 被引用 2 次
