Rethinking Style Transformer with Energy-based Interpretation: Adversarial Unsupervised Style Transfer using a Pretrained Model
Hojun Cho, Dohee Kim, Seungwoo Ryu, ChaeHun Park, Hyungjong Noh, Jeong-In Hwang, Minseok Choi, Edward Choi, Jaegul Choo
摘要
Style control, content preservation, and fluency determine the quality of text style transfer models. To train on a nonparallel corpus, several existing approaches aim to deceive the style discriminator with an adversarial loss. However, adversarial training significantly degrades fluency compared to the other two metrics. In this work, we explain this phenomenon using energy-based interpretation, and leverage a pretrained language model to improve fluency. Specifically, we propose a novel approach which applies the pretrained language model to the text style transfer framework by restructuring the discriminator and the model itself, allowing the generator and the discriminator to also take advantage of the power of the pretrained model. We evaluated our model on three public benchmarks GYAFC, Amazon, and Yelp and achieved state-of-the-art performance on the overall metrics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Residual Energy-Based Models for Text GenerationYuntian Deng, Anton Bakhtin, Myle Ott, Arthur Szlam 等ICLR 2020 · 被引用 147 次
- Your GAN is Secretly an Energy-based Model and You Should Use Discriminator Driven Latent SamplingTong Che, Ruixiang Zhang, Jascha Sohl-Dickstein, Hugo Larochelle 等NeurIPS 2020 · 被引用 128 次
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo 等ACL 2020 · 被引用 93 次
- Adapting Language Models for Non-Parallel Author-Stylized RewritingBakhtiyar Syed, Gaurav Verma, Balaji Vasan Srinivasan, Anandhavelu Natarajan 等AAAI 2020 · 被引用 53 次
相关 Paper
- Prompt-and-Rerank: A Method for Zero-Shot and Few-Shot Arbitrary Textual Style Transfer with Small Language ModelsMirac Suzgun, Luke Melas-Kyriazi, Dan JurafskyEMNLP 2022 · 被引用 34 次
- Exploring Contextual Word-level Style Relevance for Unsupervised Style TransferChulun Zhou, Liangyu Chen, Jiachen Liu, Xinyan Xiao 等ACL 2020 · 被引用 34 次
- Reformulating Unsupervised Style Transfer as Paraphrase GenerationKalpesh Krishna, John Wieting, Mohit IyyerEMNLP 2020 · 被引用 9 次
- Non-Parallel Text Style Transfer with Self-Parallel SupervisionRuibo Liu, Chongyang Gao, Chenyan Jia, Guangxuan Xu 等ICLR 2022 · 被引用 19 次
- Enhancing Content Preservation in Text Style Transfer Using Reverse Attention and Conditional Layer NormalizationDongkyu Lee, Zhiliang Tian, Lanqing Xue, Nevin L. ZhangACL 2021
