Evaluating the Morphosyntactic Well-formedness of Generated Texts
Adithya Pratapa, Antonios Anastasopoulos, Shruti Rijhwani, Aditi Chaudhary, David R. Mortensen, Graham Neubig, Yulia Tsvetkov
Abstract
Text generation systems are ubiquitous in natural language processing applications. However, evaluation of these systems remains a challenge, especially in multilingual settings. In this paper, we propose L'AMBRE -a metric to evaluate the morphosyntactic wellformedness of text using its dependency parse and morphosyntactic rules of the language. We present a way to automatically extract various rules governing morphosyntax directly from dependency treebanks. To tackle the noisy outputs from text generation systems, we propose a simple methodology to train robust parsers. We show the effectiveness of our metric on the task of machine translation through a diachronic study of systems translating into morphologically-rich languages. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e07ad57f-9ecd-4aa3-8a27-d9de5acff859Cited by top-tier papers2
- Few-shot Controllable Style Transfer for Low-Resource Multilingual SettingsKalpesh Krishna, Deepak Nathani, Xavier Garcia, Bidisha Samanta et al.ACL 2022 · 28 citations
- CaMEL: Case Marker Extraction without LabelsLeonie Weissweiler, Valentin Hofmann, Masoud Jalili Sabet, Hinrich SchützeACL 2022 · 3 citations
Builds on7
- It's Morphin' Time! Combating Linguistic Discrimination with Inflectional PerturbationsSamson Tan, Shafiq R. Joty, Min-Yen Kan, Richard SocherACL 2020 · 88 citations
- Beyond User Self-Reported Likert Scale Ratings: A Comparison Model for Automatic Dialog EvaluationWeixin Liang, James Zou, Zhou YuACL 2020 · 25 citations
- A Latent Morphology Model for Open-Vocabulary Neural Machine TranslationDuygu Ataman, Wilker Aziz, Alexandra BirchICLR 2020 · 18 citations
- Tangled up in BLEU: Reevaluating the Evaluation of Automatic Machine Translation Evaluation MetricsNitika Mathur, Timothy Baldwin, Trevor CohnACL 2020 · 14 citations
- Cross-Linguistic Syntactic Evaluation of Word Prediction ModelsAaron Mueller, Garrett Nicolai, Panayiota Petrou-Zeniou, Natalia Talmina et al.ACL 2020 · 2 citations
Related papers
- Automatic Extraction of Rules Governing Morphological AgreementAditi Chaudhary, Antonios Anastasopoulos, Adithya Pratapa, David R. Mortensen et al.EMNLP 2020
- Unsupervised Morphological Paradigm CompletionHuiming Jin, Liwei Cai, Yihui Peng, Chen Xia et al.ACL 2020 · 20 citations
- The Paradigm Discovery ProblemAlexander Erdmann, Micha Elsner, Shijie Wu, Ryan Cotterell et al.ACL 2020 · 1 citation
- Communicating in Emergent Language with an Induced Morphological PhrasebookBrendon Boldt, David R. MortensenACL 2026
- Extracting Linguistic Information from Large Language Models: Syntactic Relations and Derivational KnowledgeTsedeniya Kinfe Temesgen, Marion Di Marco, Alexander FraserEMNLP 2025 · 2 citations
