Language Drift in Multilingual Retrieval-Augmented Generation: Characterization and Decoding-Time Mitigation
Bo Li, Zhenghua Xu, Rui Xie
摘要
Multilingual Retrieval-Augmented Generation (RAG) enables large language models (LLMs) to perform knowledge-intensive tasks in multilingual settings by leveraging retrieved documents as external evidence. However, when the retrieved evidence differs in language from the user query and in-context exemplars, the model often exhibits language drift by generating responses in an unintended language. This phenomenon is especially pronounced during reasoning-intensive decoding, such as Chain-of-Thought (CoT) generation, where intermediate steps introduce further language instability. In this paper, we systematically study output language drift in multilingual RAG across multiple datasets, languages, and LLM backbones. Our controlled experiments reveal that the drift results not from comprehension failure but from decoder-level collapse, where dominant token distributions and high-frequency English patterns dominate the intended generation language. We further observe that English serves as a semantic attractor under cross-lingual conditions, emerging as both the strongest interference source and the most frequent fallback language.
To mitigate this, we propose Soft Constrained Decoding (SCD), a lightweight, training-free decoding strategy that gently steers generation toward the target language by penalizing non-target-language tokens. SCD is model-agnostic and can be applied to any generation algorithm without modifying the architecture or requiring additional data. Experiments across three multilingual datasets and multiple typologically diverse languages show that SCD consistently improves language alignment and task performance, providing an effective and generalizable solution in multilingual RAG.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Retrieval as Generation: A Unified Framework with Self-Triggered Information PlanningBo Li, Mingda Wang, Gexiang Fang, Shikun Zhang 等ACL 2026 · 被引用 8 次
- Learning from AVA: Early Lessons from a Curated and Trustworthy Generative AI for Policy and Development ResearchNimisha Karnatak, Mohamad Chatila, Daniel Alejandro Pinzón Hernández, Reza Yazdanfar 等CHI 2026 · 被引用 1 次
它引用的顶会 Paper6
- Language models are multilingual chain-of-thought reasonersFreda Shi, Mirac Suzgun, Markus Freitag, Xuezhi Wang 等ICLR 2023 · 被引用 52 次
- Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial TrainingFeiteng Fang, Yuelin Bai, Shiwen Ni, Min Yang 等ACL 2024 · 被引用 18 次
- Landmark Embedding: A Chunking-Free Embedding Method For Retrieval Augmented Long-Context Large Language ModelsKun Luo, Zheng Liu, Shitao Xiao, Tong Zhou 等ACL 2024 · 被引用 10 次
- MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented GenerationMaría Andrea Cruz Blandón, Jayasimha Talur, Bruno Charron, Dong Liu 等ACL 2025
- Improving Multilingual Retrieval-Augmented Language Models through Dialectic Reasoning ArgumentationsLeonardo Ranaldi, Federico Ranaldi, Fabio Massimo Zanzotto, Barry Haddow 等EMNLP 2025
相关 Paper
- All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAGDan Wang, Guozhao Mo, Yafei Shi, Cheng Zhang 等ACL 2026
- Linguistic Nepotism: Trading-off Quality for Language Preference in Multilingual RAGDayeon Ki, Marine Carpuat, Paul McNamee, Daniel Khashabi 等ICML 2026 · 被引用 7 次
- LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint GuidanceYuchun Fan, Bei Li, Peiguang Li, Yilin Wang 等ACL 2026 · 被引用 1 次
- LiR3AG: A Lightweight Rerank Reasoning Strategy Framework for Retrieval-Augmented GenerationGuo Chen, Junjie Huang, Huaijin Xie, Fei Sun 等AAAI 2026
- M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAGDavid Anugraha, Patrick Amadeus Irawan, Anshul Singh, En-Shiun Annie Lee 等CVPR 2026 · 被引用 2 次
