Morphological Inflection: A Reality Check
Jordan Kodner, Sarah R. B. Payne, Salam Khalifa, Zoey Liu
Abstract
Morphological inflection is a popular task in sub-word NLP with both practical and cognitive applications. For years now, state-of-the-art systems have reported high, but also highly variable, performance across data sets and languages. We investigate the causes of this high performance and high variability; we find several aspects of data set creation and evaluation which systematically inflate performance and obfuscate differences between languages. To improve generalizability and reliability of results, we propose new data sampling and evaluation strategies that better reflect likely use-cases. Using these new strategies, we make new observations on the generalization abilities of current inflection systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 38c6dc88-963c-474b-bb2c-c48c8907f2bfCited by top-tier papers3
- Extracting Linguistic Information from Large Language Models: Syntactic Relations and Derivational KnowledgeTsedeniya Kinfe Temesgen, Marion Di Marco, Alexander FraserEMNLP 2025 · 2 citations
- Understanding Compositional Data Augmentation in Typologically Diverse Morphological InflectionFarhan Samir, Miikka SilfverbergEMNLP 2023
- Getting The Most Out of Your Training Data: Exploring Unsupervised Tasks for Morphological InflectionAbhishek Purushothama, Adam Wiemerslage, Katharina von der WenseEMNLP 2024
Builds on2
- Not always about you: Prioritizing community needs when developing endangered language technologyZoey Liu, Crystal Richardson, Richard J. Hatcher, Emily Prud'hommeauxACL 2022 · 36 citations
- Systematic Inequalities in Language Technology Performance across the World's LanguagesDamián E. Blasi, Antonios Anastasopoulos, Graham NeubigACL 2022
Related papers
- Subword Segmentation in LLMs: Looking at Inflection and ConsistencyMarion Di Marco, Alexander FraserEMNLP 2024
- What is "Typological Diversity" in NLP?Esther Ploeger, Wessel Poelman, Miryam de Lhoneux, Johannes BjervaEMNLP 2024 · 2 citations
- Word Frequency Does Not Predict Grammatical Knowledge in Language ModelsCharles Yu, Ryan Sie, Nico Tedeschi, Leon BergenEMNLP 2020 · 6 citations
- Unsupervised Morphological Paradigm CompletionHuiming Jin, Liwei Cai, Yihui Peng, Chen Xia et al.ACL 2020 · 20 citations
- A Comprehensive Comparison of Neural Networks as Cognitive Models of InflectionAdam Wiemerslage, Shiran Dudy, Katharina KannEMNLP 2022 · 3 citations
