Likelihood Ratios and Generative Classifiers for Unsupervised Out-of-Domain Detection in Task Oriented Dialog
Varun Gangal, Abhinav Arora, Arash Einolghozati, Sonal Gupta
Abstract
The task of identifying out-of-domain (OOD) input examples directly at test-time has seen renewed interest recently due to increased real world deployment of models. In this work, we focus on OOD detection for natural language sentence inputs to task-based dialog systems. Our findings are three-fold: First, we curate and release ROSTD (Real Out-of -Domain Sentences From Task-oriented Dialog) -a dataset of 4K OOD examples for the publicly available dataset from (Schuster et al. 2019) . In contrast to existing settings which synthesize OOD examples by holding out a subset of classes, our examples were authored by annotators with apriori instructions to be out-of-domain with respect to the sentences in an existing dataset. Second, we explore likelihood ratio based approaches as an alternative to currently prevalent paradigms. Specifically, we reformulate and apply these approaches to natural language inputs. We find that they match or outperform the latter on all datasets, with larger improvements on non-artificial OOD benchmarks such as our dataset. Our ablations validate that specifically using likelihood ratios rather than plain likelihood is necessary to discriminate well between OOD and indomain data. Third, we propose learning a generative classifier and computing a marginal likelihood (ratio) for OOD detection. This allows us to use a principled likelihood while at the same time exploiting training-time labels. We find that this approach outperforms both simple likelihood (ratio) based and other prior approaches. We are hitherto the first to investigate the use of generative classifiers for OOD detection at test-time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2bef725b-dfb4-4f6f-84b4-1a1afcb771daCited by top-tier papers11
- Heterogeneous Customizable Personalized Federated Fine-Tuning Approach for Large Language Modelsxin tong, Baojiang cuiICML 2026 · 199 citations
- Deep Open Intent Classification with Adaptive Decision BoundaryHanlei Zhang, Hua Xu, Ting-En LinAAAI 2021 · 127 citations
- Revisiting Mahalanobis Distance for Transformer-Based Out-of-Domain DetectionAlexander Podolskiy, Dmitry Lipin, Andrey Bout, Ekaterina Artemova et al.AAAI 2021 · 100 citations
- Nonparametric Uncertainty Quantification for Single Deterministic Neural NetworkNikita Kotelevskii, Aleksandr Artemenkov, Kirill Fedyanin, Fedor Noskov et al.NeurIPS 2022 · 52 citations
- Multimodal Dialog System: Relational Graph-based Context-aware Question UnderstandingHaoyu Zhang, Meng Liu, Zan Gao, Xiaoqiang Lei et al.ACM MM 2021 · 31 citations
Related papers
- Out-of-Scope Intent Detection with Self-Supervision and Discriminative TrainingLi-Ming Zhan, Haowen Liang, Bo Liu, Lu Fan et al.ACL 2021
- Novel Slot Detection: A Benchmark for Discovering Unknown Slot Types in the Task-Oriented Dialogue SystemYanan Wu, Zhiyuan Zeng, Keqing He, Hong Xu et al.ACL 2021
- Types of Out-of-Distribution Texts and How to Detect ThemUdit Arora, William Huang, He HeEMNLP 2021
- Concept Matching with Agent for Out-of-Distribution DetectionYuxiao Lee, Xiaofeng Cao, Jingcai Guo, Wei Ye et al.AAAI 2025 · 8 citations
- In or Out? Fixing ImageNet Out-of-Distribution Detection EvaluationJulian Bitterwolf, Maximilian Müller, Matthias HeinICML 2023 · 154 citations
