Predictive Querying for Autoregressive Neural Sequence Models
Alex Boyd, Samuel Showalter, Stephan Mandt, Padhraic Smyth
Abstract
In reasoning about sequential events it is natural to pose probabilistic queries such as "when will event A occur next" or "what is the probability of A occurring before B", with applications in areas such as user modeling, medicine, and finance. However, with machine learning shifting towards neural autoregressive models such as RNNs and transformers, probabilistic querying has been largely restricted to simple cases such as next-event prediction. This is in part due to the fact that future querying involves marginalization over large path spaces, which is not straightforward to do efficiently in such models. In this paper we introduce a general typology for predictive queries in neural autoregressive sequence models and show that such queries can be systematically represented by sets of elementary building blocks. We leverage this typology to develop new query estimation methods based on beam search, importance sampling, and hybrids. Across four large-scale sequence datasets from different application domains, as well as for the GPT-2 language model, we demonstrate the ability to make query answering tractable for arbitrary queries in exponentially-large predictive path-spaces, and find clear differences in cost-accuracy tradeoffs between search and sampling methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 38bcc666-7cc0-4890-9b40-378bfcb7f172Cited by top-tier papers3
- Tractable Control for Autoregressive Language GenerationHonghua Zhang, Meihua Dang, Nanyun Peng, Guy Van den BroeckICML 2023 · 63 citations
- Pairwise Causality Guided Transformers for Event SequencesXiao Shou, Debarun Bhattacharjya, Tian Gao, Dharmashankar Subramanian et al.NeurIPS 2023 · 6 citations
- Estimating Tail Risks in Language Model Output DistributionsRico Angell, Raghav Singhal, Zachary Horvitz, Zhou Yu et al.ICML 2026 · 3 citations
Builds on5
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- Incremental Sampling Without Replacement for Sequence ModelsKensen Shi, David Bieber, Charles SuttonICML 2020 · 29 citations
- A generative nonparametric Bayesian model for whole genomesAlan Nawzad Amin, Eli N. Weinstein, Debora S. MarksNeurIPS 2021 · 9 citations
- Conditional Poisson Stochastic BeamsClara Meister, Afra Amini, Tim Vieira, Ryan CotterellEMNLP 2021
- Language Model Evaluation Beyond PerplexityClara Meister, Ryan CotterellACL 2021
Related papers
- A Unified Deep Model of Learning from both Data and Queries for Cardinality EstimationPeizhi Wu, Gao CongSIGMOD 2021 · 73 citations
- Probabilistic Attention-to-Influence Neural Models for Event SequencesXiao Shou, Debarun Bhattacharjya, Tian Gao, Dharmashankar Subramanian et al.ICML 2023 · 4 citations
- Parallel Sampling via CountingNima Anari, Ruiquan Gao, Aviad RubinsteinSTOC 2024 · 2 citations
- How Transformers Learn Causal Structures In-Context: Explainable Mechanism Meets Theoretical GuaranteeJianzhe Wei, Siyu Chen, Jianliang He, Zhuoran YangICLR 2026
- How reinforcement learning after next-token prediction facilitates learningNikolaos Tsilivis, Eran Malach, Karen Ullrich, Julia KempeICLR 2026 · 9 citations
