Contrastive Novelty-Augmented Learning: Anticipating Outliers with Large Language Models
Albert Xu, Xiang Ren, Robin Jia
摘要
In many task settings, text classification models are likely to encounter examples from novel classes on which they cannot predict correctly. Selective prediction, in which models abstain on low-confidence examples, provides a possible solution, but existing models are often overly confident on unseen classes. To remedy this overconfidence, we introduce Contrastive Novelty-Augmented Learning (CoNAL), a twostep method that generates OOD examples representative of novel classes, then trains to decrease confidence on them. First, we generate OOD examples by prompting a large language model twice: we prompt it to enumerate relevant novel classes, then generate examples from each novel class matching the task format. Second, we train a classifier with a novel contrastive objective that encourages lower confidence on generated OOD examples than training examples. When trained with CoNAL, classifiers improve in their ability to detect and abstain on novel class examples over prior methods by an average of 2.3% in terms of accuracy under the accuracy-coverage curve (AUAC) and 5.5% AUROC across 4 NLP datasets, with no cost to in-distribution accuracy. 1 Novelty Prompting CCL Training Selective Prediction ✅ Yankees Unlikely to Get D-Rays… A new report from the Canadian Food… Sports World Business Sports World Business Kyrie Irving returns to play in… The Cartwheel Galaxy Is the… " Generate a diverse list of news categories: world, sports, business, ID Labels science, crime, travel, auto, food OOD Labels Given a label, generate a corresponding example: world Bush, Kerry Trade Barbs Following… sports Yankees Unlikely to Get D-Rays… business ECB sees gradual recovery in eurozone… Example Generation Generation Label Generation
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach 等ICLR 2022 · 被引用 1,976 次
- Out-of-Distribution Detection with Deep Nearest NeighborsYiyou Sun, Yifei Ming, Xiaojin Zhu, Yixuan LiICML 2022 · 被引用 789 次
- VOS: Learning What You Don't Know by Virtual Outlier SynthesisXuefeng Du, Zhaoning Wang, Mu Cai, Yixuan LiICLR 2022 · 被引用 417 次
- SSD: A Unified Framework for Self-Supervised Outlier DetectionVikash Sehwag, Mung Chiang, Prateek MittalICLR 2021 · 被引用 410 次
相关 Paper
- Learning Transferable Negative Prompts for Out-of-Distribution DetectionTianqi Li, Guansong Pang, Xiao Bai, Wenjun Miao 等CVPR 2024
- Beyond the Known: An Unknown-Aware Large Language Model for Open-Set Text ClassificationXi Chen, Chuan Qin, Ziqi Wang, Shasha Hu 等ICLR 2026
- Confidence-aware Contrastive Learning for Selective ClassificationYu-Chang Wu, Shen-Huan Lyu, Haopu Shang, Xiangyu Wang 等ICML 2024 · 被引用 9 次
- EAT: Towards Long-Tailed Out-of-Distribution DetectionTong Wei, Bo-Lin Wang, Min-Ling ZhangAAAI 2024 · 被引用 21 次
- Out-of-Distribution Detection and Selective Generation for Conditional Language ModelsJie Ren, Jiaming Luo, Yao Zhao, Kundan Krishna 等ICLR 2023 · 被引用 12 次
