An Analysis of the Utility of Explicit Negative Examples to Improve the Syntactic Abilities of Neural Language Models
Hiroshi Noji, Hiroya Takamura
Abstract
We explore the utilities of explicit negative examples in training neural language models. Negative examples here are incorrect words in a sentence, such as barks in *The dogs barks. Neural language models are commonly trained only on positive examples, a set of sentences in the training data, but recent studies suggest that the models trained in this way are not capable of robustly handling complex syntactic constructions, such as long-distance agreement. In this paper, we first demonstrate that appropriately using negative examples about particular constructions (e.g., subject-verb agreement) will boost the model’s robustness on them in English, with a negligible loss of perplexity. The key to our success is an additional margin loss between the log-likelihoods of a correct word and an incorrect word. We then provide a detailed analysis of the trained models. One of our findings is the difficulty of object-relative clauses for RNNs. We find that even with our direct learning signals the models still suffer from resolving agreement across an object-relative clause. Augmentation of training sentences involving the constructions somewhat helps, but the accuracy still does not reach the level of subject-relative clauses. Although not directly cognitively appealing, our method can be a tool to analyze the true architectural limitation of neural models on challenging linguistic constructions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6a4536ed-ec7e-4141-a04d-0f416147b34eCited by top-tier papers2
- Cross-Linguistic Syntactic Evaluation of Word Prediction ModelsAaron Mueller, Garrett Nicolai, Panayiota Petrou-Zeniou, Natalia Talmina et al.ACL 2020 · 2 citations
- Learning with Instance Bundles for Reading ComprehensionDheeru Dua, Pradeep Dasigi, Sameer Singh, Matt GardnerEMNLP 2021 · 1 citation
Builds on2
Related papers
- NegatER: Unsupervised Discovery of Negatives in Commonsense Knowledge BasesTara Safavi, Jing Zhu, Danai KoutraEMNLP 2021 · 9 citations
- Influence Paths for Characterizing Subject-Verb Number Agreement in LSTM Language ModelsKaiji Lu, Piotr Mardziel, Klas Leino, Matt Fredrikson et al.ACL 2020 · 8 citations
- CLINE: Contrastive Learning with Semantic Negative Examples for Natural Language UnderstandingDong Wang, Ning Ding, Piji Li, Haitao ZhengACL 2021
- Recurrent Neural Network Language Models Always Learn English-Like Relative Clause AttachmentForrest Davis, Marten van SchijndelACL 2020 · 15 citations
- Negative Pre-activations Differentiate SyntaxLinghao Kong, Angelina Ning, Micah Adler, Nir N ShavitICLR 2026 · 2 citations
