DADA: Dialect Adaptation via Dynamic Aggregation of Linguistic Rules
Yanchen Liu, William Barr Held, Diyi Yang
Abstract
Existing large language models (LLMs) that mainly focus on Standard American English (SAE) often lead to significantly worse performance when being applied to other English dialects. While existing mitigations tackle discrepancies for individual target dialects, they assume access to high-accuracy dialect identification systems. The boundaries between dialects are inherently flexible, making it difficult to categorize language into discrete predefined categories. In this work, we propose DADA (Dialect Adaptation via Dynamic Aggregation), a modular approach to imbue SAE-trained models with multi-dialectal robustness by composing adapters which handle specific linguistic features. The compositional architecture of DADA allows for both targeted adaptation to specific dialect variants and simultaneous adaptation to various dialects. We show that DADA is effective for both single task and instruction finetuned language models, offering an extensible and interpretable framework for adapting existing LLMs to different English dialects. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8e9687e5-fdf3-4646-87a5-6cda0bfe5e12Cited by top-tier papers4
- Task-Agnostic Low-Rank Adapters for Unseen English DialectsZedian Xiao, William Barr Held, Yanchen Liu, Diyi YangEMNLP 2023 · 3 citations
- MC²: Towards Transparent and Culturally-Aware NLP for Minority Languages in ChinaChen Zhang, Mingxu Tao, Quzhe Huang, Jiuheng Lin et al.ACL 2024
- A Multi-Agent Framework for Mitigating Dialect Biases in Privacy Policy Question-Answering SystemsDorde Klisura, Astrid R. Bernaga Torres, Anna Karen Gárate-Escamilla, Rajesh Roshan Biswal et al.ACL 2025
- Would LLMs be Good Historical Linguists and Chinese Dialect Learners?Yicheng Liu, Shumin Shi, Youchao Zhou, Xingchen ZhangACL 2026
Builds on12
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach et al.ICLR 2022 · 1,976 citations
- Towards a Unified View of Parameter-Efficient Transfer LearningJunxian He, Chunting Zhou, Xuezhe Ma, Taylor Berg-Kirkpatrick et al.ICLR 2022 · 1,182 citations
- Process for Adapting Language Models to Society (PALMS) with Values-Targeted DatasetsIrene Solaiman, Christy DennisonNeurIPS 2021 · 276 citations
Related papers
- DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to ModelsNiyati Bafna, Emily Chang, Nathaniel Romney Robinson, David R. Mortensen et al.ACL 2025
- DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English DialectsJason S. Lucas, Matt Murtagh-White, Ali Al-Lawati, Uchendu Uchendu et al.ACL 2026 · 1 citation
- Assessing Dialect Fairness and Robustness of Large Language Models in Reasoning TasksFangru Lin, Shaoguang Mao, Emanuele La Malfa, Valentin Hofmann et al.ACL 2025 · 14 citations
- It's Morphin' Time! Combating Linguistic Discrimination with Inflectional PerturbationsSamson Tan, Shafiq R. Joty, Min-Yen Kan, Richard SocherACL 2020 · 88 citations
- Dialect-robust Evaluation of Generated TextJiao Sun, Thibault Sellam, Elizabeth Clark, Tu Vu et al.ACL 2023 · 11 citations
