Generic resources are what you need: Style transfer tasks without task-specific parallel training data
Huiyuan Lai, Antonio Toral, Malvina Nissim
摘要
Style transfer aims to rewrite a source text in a different target style while preserving its content. We propose a novel approach to this task that leverages generic resources, and without using any task-specific parallel (source-target) data outperforms existing unsupervised approaches on the two most popular style transfer tasks: formality transfer and polarity swap. In practice, we adopt a multistep procedure which builds on a generic pretrained sequence-to-sequence model (BART). First, we strengthen the model's ability to rewrite by further pre-training BART on both an existing collection of generic paraphrases, as well as on synthetic pairs created using a general-purpose lexical resource. Second, through an iterative back-translation approach, we train two models, each in a transfer direction, so that they can provide each other with synthetically generated pairs, dynamically in the training process. Lastly, we let our best resulting model generate static synthetic pairs to be used in a supervised training regime. Besides methodology and state-of-the-art results, a core contribution of this work is a reflection on the nature of the two tasks we address, and how their differences are highlighted by their response to our approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- MSSRNet: Manipulating Sequential Style Representation for Unsupervised Text Style TransferYazheng Yang, Zhou Zhao, Qi LiuKDD 2023 · 被引用 2 次
- Latent Constraints on Unsupervised Text-Graph Alignment with Information AsymmetryJidong Tian, Wenqing Chen, Yitian Li, Caoyun Fan 等AAAI 2023
- Multi-perspective Alignment for Increasing Naturalness in Neural Machine TranslationHuiyuan Lai, Esther Ploeger, Rik van Noord, Antonio ToralACL 2025
它引用的顶会 Paper5
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Unsupervised Paraphrasing by Simulated AnnealingXianggen Liu, Lili Mou, Fandong Meng, Hao Zhou 等ACL 2020 · 被引用 74 次
- BLEURT: Learning Robust Metrics for Text GenerationThibault Sellam, Dipanjan Das, Ankur P. ParikhACL 2020 · 被引用 40 次
- Exploring Contextual Word-level Style Relevance for Unsupervised Style TransferChulun Zhou, Liangyu Chen, Jiachen Liu, Xinyan Xiao 等ACL 2020 · 被引用 34 次
- COMET: A Neural Framework for MT EvaluationRicardo Rei, Craig Stewart, Ana C. Farinha, Alon LavieEMNLP 2020 · 被引用 6 次
相关 Paper
- Learning to Selectively Learn for Weakly-supervised Paraphrase GenerationKaize Ding, Dingcheng Li, Alexander Hanbo Li, Xing Fan 等EMNLP 2021 · 被引用 4 次
- Monolingual Transfer Learning via Bilingual Translators for Style-Sensitive Paraphrase GenerationTomoyuki Kajiwara, Biwa Miura, Yuki AraseAAAI 2020 · 被引用 8 次
- Reformulating Unsupervised Style Transfer as Paraphrase GenerationKalpesh Krishna, John Wieting, Mohit IyyerEMNLP 2020 · 被引用 9 次
- TextSETTR: Few-Shot Text Style Extraction and Tunable Targeted RestylingParker Riley, Noah Constant, Mandy Guo, Girish Kumar 等ACL 2021
- Masked Based Unsupervised Content TransferRon Mokady, Sagie Benaim, Lior Wolf, Amit BermanoICLR 2020
