Lune

ACL2025Top-tier venue

BelarusianGLUE: Towards a Natural Language Understanding Benchmark for Belarusian

Maksim Aparovich, Volha Harytskaya, Vladislav Poritski, Oksana Volchek, Pavel Smrz

2025Year

Abstract

In the epoch of multilingual large language models (LLMs), it is still challenging to evaluate the models' understanding of lowerresourced languages, which motivates further development of expert-crafted natural language understanding benchmarks. We introduce Be-larusianGLUE -a natural language understanding benchmark for Belarusian, an East Slavic language, with ≈15K instances in five tasks: sentiment analysis, linguistic acceptability, word in context, Winograd schema challenge, textual entailment. A systematic evaluation of BERT models and LLMs against this novel benchmark reveals that both types of models approach human-level performance on easier tasks, such as sentiment analysis, but there is a significant gap in performance between machine and human on a harder task -Winograd schema challenge. We find the optimal choice of model type to be task-specific: e.g. BERT models underperform on textual entailment task but are competitive for linguistic acceptability. We release the datasets 1 and evaluation code. 2

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 33b582d6-329f-459e-bedc-47f0083c7b89

Builds on19

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines