AdaDPI: Document-level Translation Adaptive Agent via Dynamic Parametric Internalization
Hong Ren, Liting Deng, Shaolin Zhu, Deyi Xiong
摘要
Large Language Models (LLMs) have demonstrated remarkable capabilities in machine translation. However, maintaining discourse coherence and terminological consistency remains a persistent challenge in documentlevel translation (DocMT). Existing solutions, such as memory-based agents, predominantly rely on explicit context concatenation. This paradigm treats historical context as a static external resource, which often leads to context dilution, high inference latency, and superficial knowledge integration. To address these limitations, we propose AdaDPI, an adaptive agentic framework that shifts the DocMT paradigm from static retrieval to dynamic parametric internalization. Specifically, we design a linguistic uncertainty monitor (LUM) to actively detect critical discourse discontinuities by the model's epistemic uncertainty. Upon detection, a context-to-parameter integrator (CPI) compiles retrieved external constraints directly into the model's intrinsic state via an online parameter adaptation mechanism. Through the online parameter adaptation on a lightweight adapter, AdaDPI internalizes document-specific norms into the model's intrinsic representations, enabling a progressive evolution of the translation strategy as the discourse unfolds. Extensive experiments on the discourse-rich GuoFeng and IWSLT2017 datasets demonstrate that AdaDPI significantly outperforms the SoTA baselines by more than 5 points on the consistency metric.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon AgentsZijian Zhou, Ao Qu, Zhaoxuan Wu, Sunghwan Kim 等ICLR 2026 · 被引用 223 次
- Encouraging Lexical Translation Consistency for Document-Level Neural Machine TranslationXinglin Lyu, Junhui Li, Zhengxian Gong, Min ZhangEMNLP 2021 · 被引用 17 次
- Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language ModelsLei Tang, Jinghui Qin, Wenxuan Ye, Hao Tan 等AAAI 2025 · 被引用 9 次
- Long Context is Not Long at All: A Prospector of Long-Dependency Data for Large Language ModelsLongze Chen, Ziqiang Liu, Wanwei He, Yinhe Zheng 等ACL 2024 · 被引用 4 次
- G-Transformer for Document-Level Machine TranslationGuangsheng Bao, Yue Zhang, Zhiyang Teng, Boxing Chen 等ACL 2021
相关 Paper
- DelTA: An Online Document-Level Translation Agent Based on Multi-Level MemoryYutong Wang, Jiali Zeng, Xuebo Liu, Derek F. Wong 等ICLR 2025
- GAM: Hierarchical Graph-based Agentic Memory for LLM AgentsZhaofen Wu, Hanrong Zhang, Fulin Lin, Wujiang Xu 等ACL 2026 · 被引用 8 次
- Seeing through the Conflict: Transparent Knowledge Conflict Handling in Retrieval-Augmented GenerationHua Ye, Siyuan Chen, Ziqi Zhong, Canran Xiao 等AAAI 2026 · 被引用 1 次
- DICE: Dynamic In-Context Example Selection in LLM Agents via Efficient Knowledge TransferRuoyu Wang, Junda Wu, Yu Xia, Tong Yu 等KDD 2026 · 被引用 6 次
- Synergistic Multi-Agent Framework with Trajectory Learning for Knowledge-Intensive TasksShengbin Yue, Siyuan Wang, Wei Chen, Xuanjing Huang 等AAAI 2025 · 被引用 25 次
