On the Inference Calibration of Neural Machine Translation
Shuo Wang, Zhaopeng Tu, Shuming Shi, Yang Liu
摘要
Confidence calibration, which aims to make model predictions equal to the true correctness measures, is important for neural machine translation (NMT) because it is able to offer useful indicators of translation errors in the generated output. While prior studies have shown that NMT models trained with label smoothing are well-calibrated on the groundtruth training data, we find that miscalibration still remains a severe challenge for NMT during inference due to the discrepancy between training and inference. By carefully designing experiments on three language pairs, our work provides in-depth analyses of the correlation between calibration and translation performance as well as linguistic properties of miscalibration and reports a number of interesting findings that might help humans better analyze, understand and improve NMT models. Based on these observations, we further propose a new graduated label smoothing method that can improve both inference calibration and translation performance. 1 * Work was done when Shuo Wang was interning at Tencent AI Lab under the Rhino-Bird Elite Training Program. 1 The source code is available at https://github. com/shuo-git/InfECE .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper31
- BRIO: Bringing Order to Abstractive SummarizationYixin Liu, Pengfei Liu, Dragomir R. Radev, Graham NeubigACL 2022 · 被引用 329 次
- On the Calibration of Pre-trained Language Models using Mixup Guided by Area Under the Margin and SaliencySeoyeon Park, Cornelia CarageaACL 2022 · 被引用 44 次
- Token-level Adaptive Training for Neural Machine TranslationShuhao Gu, Jinchao Zhang, Fandong Meng, Yang Feng 等EMNLP 2020 · 被引用 32 次
- Don't Hallucinate, Abstain: Identifying LLM Knowledge Gaps via Multi-LLM CollaborationShangbin Feng, Weijia Shi, Yike Wang, Wenxuan Ding 等ACL 2024 · 被引用 30 次
- Navigating the Grey Area: How Expressions of Uncertainty and Overconfidence Affect Language ModelsKaitlyn Zhou, Dan Jurafsky, Tatsunori HashimotoEMNLP 2023 · 被引用 29 次
相关 Paper
- Learning Confidence for Transformer-based Neural Machine TranslationYu Lu, Jiali Zeng, Jiajun Zhang, Shuangzhi Wu 等ACL 2022 · 被引用 9 次
- A Close Look into the Calibration of Pre-trained Language ModelsYangyi Chen, Lifan Yuan, Ganqu Cui, Zhiyuan Liu 等ACL 2023 · 被引用 12 次
- Calibrating Zero-shot Cross-lingual (Un-)structured PredictionsZhengping Jiang, Anqi Liu, Benjamin Van DurmeEMNLP 2022 · 被引用 4 次
- Understanding and Addressing the Under-Translation Problem from the Perspective of Decoding ObjectiveChenze Shao, Fandong Meng, Jiali Zeng, Jie ZhouACL 2024
- Towards Robust k-Nearest-Neighbor Machine TranslationHui Jiang, Ziyao Lu, Fandong Meng, Chulun Zhou 等EMNLP 2022 · 被引用 16 次
