GAPX: Generalized Autoregressive Paraphrase-Identification X
Yifei Zhou, Renyu Li, Hayden Housen, Ser Nam Lim
Abstract
Paraphrase Identification is a fundamental task in Natural Language Processing. While much progress has been made in the field, the performance of many state-ofthe-art models often suffer from distribution shift during inference time. We verify that a major source of this performance drop comes from biases introduced by negative examples. To overcome these biases, we propose in this paper to train two separate models, one that only utilizes the positive pairs and the other the negative pairs. This enables us the option of deciding how much to utilize the negative model, for which we introduce a perplexity based out-of-distribution metric that we show can effectively and automatically determine how much weight it should be given during inference. We support our findings with strong empirical results. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e532803-9bca-4c1e-b2b8-7ce9c28ae0b6Builds on16
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
Related papers
- Interventional Training for Out-Of-Distribution Natural Language UnderstandingSicheng Yu, Jing Jiang, Hao Zhang, Yulei Niu et al.EMNLP 2022 · 3 citations
- Principled Paraphrase Generation with Parallel CorporaAitor Ormazabal, Mikel Artetxe, Aitor Soroa, Gorka Labaka et al.ACL 2022 · 12 citations
- Towards Better Characterization of ParaphrasesTimothy Liu, De Wen SohACL 2022 · 9 citations
- On the Evaluation Metrics for Paraphrase GenerationLingfeng Shen, Lemao Liu, Haiyun Jiang, Shuming ShiEMNLP 2022 · 29 citations
- Omitted Variable Bias in Language Models Under Distribution ShiftVictoria Lin, Louis-Philippe Morency, Eli Ben-MichaelICML 2026 · 1 citation
