Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text
Amr Mohamed, Yang Zhang, Michalis Vazirgiannis, Guokan Shang
摘要
Code-switching (CSW) is the act of alternating between two or more languages within a single discourse. This phenomenon is widespread in multilingual communities, and increasingly prevalent in online content, where users naturally mix languages in everyday communication. As a result, Large Language Models (LLMs), now central to content processing and generation, are frequently exposed to code-switched inputs. Given their widespread use, it is crucial to understand how LLMs process and reason about such mixed-language text. This paper presents a systematic evaluation of LLM comprehension under codeswitching by generating CSW variants of established reasoning and comprehension benchmarks. While degradation is evident when foreign tokens disrupt English text-even under linguistic constraints-embedding English into other languages often improves comprehension. Though prompting yields mixed results, finetuning offers a more stable path to degradation mitigation. : (D) : Hume says that beauty is _____. : Hume says that اﻟﺠﻤﺎل is _____. : Hume says that la beauté is _____. : Hume says that Schönheit is _____. : Hume says that 美 is _____. (A) a quality in things themselves (B) a matter of a priori knowledge (C) judged by logical standards (D) no quality in things themselves : (A) : (D) : (C) : (B)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou 等ICLR 2021 · 被引用 7,905 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- Attention-Informed Mixed-Language Training for Zero-Shot Cross-Lingual Task-Oriented Dialogue SystemsZihan Liu, Genta Indra Winata, Zhaojiang Lin, Peng Xu 等AAAI 2020 · 被引用 105 次
- The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language VariantsLucas Bandarkar, Davis Liang, Benjamin Muller, Mikel Artetxe 等ACL 2024 · 被引用 30 次
- Multilingual Large Language Models Are Not (Yet) Code-SwitchersRuochen Zhang, Samuel Cahyawijaya, Jan Christian Blaise Cruz, Genta Indra Winata 等EMNLP 2023 · 被引用 19 次
相关 Paper
- SASFT: Sparse Autoencoder-guided Supervised Finetuning to Mitigate Unexpected Code-Switching in LLMsBoyi Deng, Yu Wan, Baosong Yang, Fei Huang 等ICLR 2026 · 被引用 2 次
- OLA: Output Language Alignment in Code-Switched LLM InteractionsJuhyun Oh, Haneul Yoo, Faiz Ghifari Haznitrama, Alice OhACL 2026 · 被引用 1 次
- Minimal Pair-Based Evaluation of Code-SwitchingIgor Sterner, Simone TeufelACL 2025 · 被引用 8 次
- Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual UnderstandingHaneul Yoo, Yongjin Yang, Hwaran LeeACL 2025 · 被引用 27 次
- AfroCS-xs: Creating a Compact, High-Quality, Human-Validated Code-Switched Dataset for African LanguagesKayode Olaleye, Arturo Oncevay, Mathieu Sibue, Nombuyiselo Zondi 等ACL 2025 · 被引用 5 次
