Exploring Discourse Structure in Document-level Machine Translation
Xinyu Hu, Xiaojun Wan
Abstract
Neural machine translation has achieved great success in the past few years with the help of transformer architectures and large-scale bilingual corpora. However, when the source text gradually grows into an entire document, the performance of current methods for document-level machine translation (DocMT) is less satisfactory. Although the context is beneficial to the translation in general, it is difficult for traditional methods to utilize such long-range information. Previous studies on DocMT have concentrated on extra contents such as multiple surrounding sentences and input instances divided by a fixed length. We suppose that they ignore the structure inside the source text, which leads to under-utilization of the context. In this paper, we present a more sound paragraph-to-paragraph translation mode and explore whether discourse structure can improve DocMT. We introduce several methods from different perspectives, among which our RST-Att model with a multi-granularity attention mechanism based on the RST parsing tree works best. The experiments show that our method indeed utilizes discourse information and performs better than previous work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 15f6aa36-5e6e-40ed-8ca3-aa638f8ffac1Cited by top-tier papers4
- DeMPT: Decoding-enhanced Multi-phase Prompt Tuning for Making LLMs Be Better Context-aware TranslatorsXinglin Lyu, Junhui Li, Yanqing Zhao, Min Zhang et al.EMNLP 2024 · 4 citations
- Improving Implicit Discourse Relation Recognition with Natural Language Explanations from LLMsHeng Wang, Changxing WuAAAI 2026
- RST-Guarder: Enhancing Long-Context Robustness for Safeguards via RST Parsing and Probabilistic InferenceXu Zhang, Xiaojun WanACL 2026
- Fine-Grained Modeling of Narrative Context: A Coherence Perspective via Retrospective QuestionsLiyan Xu, Jiangnan Li, Mo Yu, Jie ZhouACL 2024
Builds on10
- Discourse-Aware Neural Extractive Text SummarizationJiacheng Xu, Zhe Gan, Yu Cheng, Jingjing LiuACL 2020 · 264 citations
- GraphFormers: GNN-nested Transformers for Representation Learning on Textual GraphJunhan Yang, Zheng Liu, Shitao Xiao, Chaozhuo Li et al.NeurIPS 2021 · 262 citations
- Time Travel in LLMs: Tracing Data Contamination in Large Language ModelsShahriar Golchin, Mihai SurdeanuICLR 2024 · 165 citations
- Document-Level Machine Translation with Large Language ModelsLongyue Wang, Chenyang Lyu, Tianbo Ji, Zhirui Zhang et al.EMNLP 2023 · 129 citations
- Dynamic Context Selection for Document-level Neural Machine Translation via Reinforcement LearningXiaomian Kang, Yang Zhao, Jiajun Zhang, Chengqing ZongEMNLP 2020 · 61 citations
Related papers
- Top-Down RST Parsing Utilizing Granularity Levels in DocumentsNaoki Kobayashi, Tsutomu Hirao, Hidetaka Kamigaito, Manabu Okumura et al.AAAI 2020 · 48 citations
- A Top-down Neural Architecture towards Text-level Parsing of Discourse Rhetorical StructureLongyin Zhang, Yuqing Xing, Fang Kong, Peifeng Li et al.ACL 2020 · 39 citations
- Document-Level Machine Translation with Large-Scale Public Parallel CorporaProyag Pal, Alexandra Birch, Kenneth HeafieldACL 2024
- Document Graph for Neural Machine TranslationMingzhou Xu, Liangyou Li, Derek F. Wong, Qun Liu et al.EMNLP 2021
- Hierarchical Macro Discourse Parsing Based on Topic SegmentationFeng Jiang, Yaxin Fan, Xiaomin Chu, Peifeng Li et al.AAAI 2021 · 14 citations
