Adversarial Mutual Information for Text Generation
Boyuan Pan, Yazheng Yang, Kaizhao Liang, Bhavya Kailkhura, Zhongming Jin, Xian-Sheng Hua, Deng Cai, Bo Li
Abstract
Recent advances in maximizing mutual information (MI) between the source and target have demonstrated its effectiveness in text generation. However, previous works paid little attention to modeling the backward network of MI (i.e., dependency from the target to the source), which is crucial to the tightness of the variational information maximization lower bound. In this paper, we propose Adversarial Mutual Information (AMI): a text generation framework which is formed as a novel saddle point (min-max) optimization aiming to identify joint interactions between the source and target. Within this framework, the forward and backward networks are able to iteratively promote or demote each other's generated instances by comparing the real and synthetic data distributions. We also develop a latent noise sampling strategy that leverages random variations at the high-level semantic space to enhance the long term dependency in the generation process. Extensive experiments based on different text generation tasks demonstrate that the proposed AMI framework can significantly outperform several strong baselines, and we also show that AMI has potential to lead to a tighter lower bound of maximum mutual information for the variational information maximization problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 577a809e-5284-4fa3-b378-03f69f169dbfCited by top-tier papers3
- Improving Group Robustness on Spurious Correlation Requires Preciser Group InferenceYujin Han, Difan ZouICML 2024 · 13 citations
- MPII: Multi-Level Mutual Promotion for Inference and InterpretationYan Liu, Sanyuan Chen, Yazheng Yang, Qi DaiACL 2022 · 7 citations
- MSSRNet: Manipulating Sequential Style Representation for Unsupervised Text Style TransferYazheng Yang, Zhou Zhao, Qi LiuKDD 2023 · 2 citations
Related papers
- Improving Adversarial Robustness via Mutual Information EstimationDawei Zhou, Nannan Wang, Xinbo Gao, Bo Han et al.ICML 2022 · 23 citations
- Connecting Jensen-Shannon and Kullback-Leibler Divergences: A New Bound for Representation LearningReuben Dorent, Polina Golland, William (Sandy) WellsNeurIPS 2025 · 7 citations
- A Novel Estimator of Mutual Information for Learning to Disentangle Textual RepresentationsPierre Colombo, Pablo Piantanida, Chloé ClavelACL 2021
- Improving Variational Autoencoders with Density Gap-based RegularizationJianfei Zhang, Jun Bai, Chenghua Lin, Yanmeng Wang et al.NeurIPS 2022 · 11 citations
- Learning to Drop Out: An Adversarial Approach to Training Sequence VAEsDjordje Miladinovic, Kumar Shridhar, Kushal Jain, Max B. Paulus et al.NeurIPS 2022 · 5 citations
