Educating Text Autoencoders: Latent Representation Guidance via Denoising
Tianxiao Shen, Jonas Mueller, Regina Barzilay, Tommi S. Jaakkola
摘要
Generative autoencoders offer a promising approach for controllable text generation by leveraging their latent sentence representations. However, current models struggle to maintain coherent latent spaces required to perform meaningful text manipulations via latent vector operations. Specifically, we demonstrate by example that neural encoders do not necessarily map similar sentences to nearby latent vectors. A theoretical explanation for this phenomenon establishes that highcapacity autoencoders can learn an arbitrary mapping between sequences and associated latent representations. To remedy this issue, we augment adversarial autoencoders with a denoising objective where original sentences are reconstructed from perturbed versions (referred to as DAAE). We prove that this simple modification guides the latent space geometry of the resulting model by encouraging the encoder to map similar texts to similar latent representations. In empirical comparisons with various types of autoencoders, our model provides the best trade-off between generation quality and reconstruction capacity. Moreover, the improved geometry of the DAAE latent space enables zero-shot text style transfer via simple latent vector arithmetic. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Data Augmentation for Cross-Domain Named Entity RecognitionShuguang Chen, Gustavo Aguilar, Leonardo Neves, Thamar SolorioEMNLP 2021 · 被引用 39 次
- Plug and Play Autoencoders for Conditional Text GenerationFlorian Mai, Nikolaos Pappas, Ivan Montero, Noah A. Smith 等EMNLP 2020 · 被引用 24 次
- Composable Text Controls in Latent Space with ODEsGuangyi Liu, Zeyu Feng, Yuan Gao, Zichao Yang 等EMNLP 2023 · 被引用 12 次
- Unified Generation, Reconstruction, and Representation: Generalized Diffusion with Adaptive Latent Encoding-DecodingGuangyi Liu, Yu Wang, Zeyu Feng, Qiyu Wu 等ICML 2024 · 被引用 9 次
- SALSA: Semantically-Aware Latent Space AutoencoderKathryn E. Kirchoff, Travis Maxfield, Alexander Tropsha, Shawn M. GomezAAAI 2024 · 被引用 3 次
相关 Paper
- On Variational Learning of Controllable Representations for Text without SupervisionPeng Xu, Jackie Chi Kit Cheung, Yanshuai CaoICML 2020 · 被引用 69 次
- So Different Yet So Alike! Constrained Unsupervised Text Style TransferAbhinav Ramesh Kashyap, Devamanyu Hazarika, Min-Yen Kan, Roger Zimmermann 等ACL 2022
- Revision in Continuous Space: Unsupervised Text Style Transfer without Adversarial LearningDayiheng Liu, Jie Fu, Yidan Zhang, Chris Pal 等AAAI 2020 · 被引用 53 次
- Adapting Language Models for Non-Parallel Author-Stylized RewritingBakhtiyar Syed, Gaurav Verma, Balaji Vasan Srinivasan, Anandhavelu Natarajan 等AAAI 2020 · 被引用 53 次
- Adversarial Latent AutoencodersStanislav Pidhorskyi, Donald A. Adjeroh, Gianfranco DorettoCVPR 2020
