Incremental Processing in the Age of Non-Incremental Encoders: An Empirical Assessment of Bidirectional Models for Incremental NLU
Brielen Madureira, David Schlangen
摘要
While humans process language incrementally, the best language encoders currently used in NLP do not. Both bidirectional LSTMs and Transformers assume that the sequence that is to be encoded is available in full, to be processed either forwards and backwards (BiL-STMs) or as a whole (Transformers). We investigate how they behave under incremental interfaces, when partial output must be provided based on partial input seen up to a certain time step, which may happen in interactive systems. We test five models on various NLU datasets and compare their performance using three incremental evaluation metrics. The results support the possibility of using bidirectional encoders in incremental mode while retaining most of their non-incremental quality. The "omni-directional" BERT model, which achieves better non-incremental performance, is impacted more by the incremental access. This can be alleviated by adapting the training regime (truncated training), or the testing procedure, by delaying the output until some right context is available or by incorporating hypothetical right contexts generated by a language model like GPT-2.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Sentence-Incremental Neural Coreference ResolutionMatt Grenander, Shay B. Cohen, Mark SteedmanEMNLP 2022 · 被引用 4 次
- Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection IncrementalMorteza Rohanian, Julian HoughACL 2021
- When Only Time Will Tell: Interpreting How Transformers Process Local Ambiguities Through the Lens of Restart-IncrementalityBrielen Madureira, Patrick Kahardipraja, David SchlangenACL 2024
相关 Paper
- On the Robustness of Language Encoders against Grammatical ErrorsFan Yin, Quanyu Long, Tao Meng, Kai-Wei ChangACL 2020 · 被引用 32 次
- Linear Attention for Efficient Bidirectional Sequence ModelingArshia Afzal, Elías Abad-Rocamora, Leyla Naz Candogan, Pol Puigdemont 等NeurIPS 2025 · 被引用 8 次
- Incremental BPE TokenizationShenghu Jiang, Ruihao GongICML 2026 · 被引用 12 次
- BERT, mBERT, or BiBERT? A Study on Contextualized Embeddings for Neural Machine TranslationHaoran Xu, Benjamin Van Durme, Kenton W. MurrayEMNLP 2021 · 被引用 55 次
- A Targeted Assessment of Incremental Processing in Neural Language Models and HumansEthan Wilcox, Pranali Vani, Roger LevyACL 2021
