Cross-lingual Intermediate Fine-tuning improves Dialogue State Tracking
Nikita Moghe, Mark Steedman, Alexandra Birch
Abstract
Recent progress in task-oriented neural dialogue systems is largely focused on a handful of languages, as annotation of training data is tedious and expensive. Machine translation has been used to make systems multilingual, but this can introduce a pipeline of errors. Another promising solution is using cross-lingual transfer learning through pretrained multilingual models. Existing methods train multilingual models with additional codemixed task data or refine the cross-lingual representations through parallel ontologies. In this work, we enhance the transfer learning process by intermediate fine-tuning of pretrained multilingual models, where the multilingual models are fine-tuned with different but related data and/or tasks. Specifically, we use parallel and conversational movie subtitles datasets to design cross-lingual intermediate tasks suitable for downstream dialogue tasks. We use only 200K lines of parallel data for intermediate fine-tuning which is already available for 1782 language pairs. We test our approach on the cross-lingual dialogue state tracking task for the parallel Mul-tiWoZ (English→Chinese, Chinese→English) and Multilingual WoZ (English→German, English→Italian) datasets. We achieve impressive improvements (> 20% on joint goal accuracy) on the parallel MultiWoZ dataset and the Multilingual WoZ dataset over the vanilla baseline with only 10% of the target language task data and zero-shot setup respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39933552-cd22-4bfb-88d1-883203676d86Cited by top-tier papers1
Ask how each one uses itBuilds on7
- XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual GeneralisationJunjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig et al.ICML 2020 · 1,132 citations
- Muppet: Massive Multi-task Representations with Pre-FinetuningArmen Aghajanyan, Anchit Gupta, Akshat Shrivastava, Xilun Chen et al.EMNLP 2021 · 176 citations
- Intermediate-Task Transfer Learning with Pretrained Language Models: When and Why Does It Work?Yada Pruksachatkun, Jason Phang, Haokun Liu, Phu Mon Htut et al.ACL 2020 · 168 citations
- ParaCrawl: Web-Scale Acquisition of Parallel CorporaMarta Bañón, Pinzhen Chen, Barry Haddow, Kenneth Heafield et al.ACL 2020 · 132 citations
- Attention-Informed Mixed-Language Training for Zero-Shot Cross-Lingual Task-Oriented Dialogue SystemsZihan Liu, Genta Indra Winata, Zhaojiang Lin, Peng Xu et al.AAAI 2020 · 105 citations
Related papers
- Zero-Shot Transfer Learning with Synthesized Data for Multi-Domain Dialogue State TrackingGiovanni Campagna, Agata Foryciarz, Mehrad Moradshahi, Monica S. LamACL 2020 · 5 citations
- GlobalWoZ: Globalizing MultiWoZ to Develop Multilingual Task-Oriented Dialogue SystemsBosheng Ding, Junjie Hu, Lidong Bing, Sharifah Mahani Aljunied et al.ACL 2022
- NeuralWOZ: Learning to Collect Task-Oriented Dialogue via Model-Based SimulationSungdong Kim, Minsuk Chang, Sang-Woo LeeACL 2021
- Code-switched inspired losses for spoken dialog representationsPierre Colombo, Emile Chapuis, Matthieu Labeau, Chloé ClavelEMNLP 2021 · 6 citations
- Zero-Shot Dialogue State Tracking via Cross-Task TransferZhaojiang Lin, Bing Liu, Andrea Madotto, Seungwhan Moon et al.EMNLP 2021
