LLM Processes: Numerical Predictive Distributions Conditioned on Natural Language
James Requeima, John Bronskill, Dami Choi, Richard E. Turner, David Kristjanson Duvenaud
Abstract
Machine learning practitioners often face significant challenges in formally integrating their prior knowledge and beliefs into predictive models, limiting the potential for nuanced and context-aware analyses. Moreover, the expertise needed to integrate this prior knowledge into probabilistic modeling typically limits the application of these models to specialists. Our goal is to build a regression model that can process numerical data and make probabilistic predictions at arbitrary locations, guided by natural language text which describes a user's prior knowledge. Large Language Models (LLMs) provide a useful starting point for designing such a tool since they 1) provide an interface where users can incorporate expert insights in natural language and 2) provide an opportunity for leveraging latent problem-relevant knowledge encoded in LLMs that users may not have themselves. We start by exploring strategies for eliciting explicit, coherent numerical predictive distributions from LLMs. We examine these joint predictive distributions, which we call LLM Processes, over arbitrarily-many quantities in settings such as forecasting, multi-dimensional regression, black-box optimization, and image modeling. We investigate the practical details of prompting to elicit coherent predictive distributions, and demonstrate their effectiveness at regression. Finally, we demonstrate the ability to usefully incorporate text into numerical predictions, improving predictive performance and giving quantitative structure that reflects qualitative descriptions. This lets us begin to explore the rich, grounded hypothesis space that LLMs implicitly encode.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c0a33d41-1e47-4d66-831c-a57da22badb7Cited by top-tier papers16
- Variational Uncertainty Decomposition for In-Context LearningI. Shavindra Jayasekera, Jacob Si, Filippo Valdettaro, Wenlong Chen et al.NeurIPS 2025 · 7 citations
- Adaptive Acquisition Selection for Bayesian Optimization with Large Language ModelsGiang Ngo, Dat Phan Trong, Dang Nguyen, Sunil Gupta et al.ICLR 2026 · 6 citations
- LLMs as World Models: Data-Driven and Human-Centered Pre-Event Simulation for Disaster Impact AssessmentLingyao Li, Dawei Li, Zhenhui Ou, Xiaoran Xu et al.EMNLP 2025 · 5 citations
- LILO: Bayesian Optimization with Natural Language FeedbackKatarzyna Kobalczyk, Zhiyuan Lin, Benjamin Letham, Zhuokai Zhao et al.ICML 2026 · 2 citations
- Eliciting Numerical Predictive Distributions of LLMs Without Auto-RegressionJulianna Piskorz, Kasia Kobalczyk, Mihaela van der SchaarICLR 2026 · 2 citations
Builds on11
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- An Explanation of In-context Learning as Implicit Bayesian InferenceSang Michael Xie, Aditi Raghunathan, Percy Liang, Tengyu MaICLR 2022 · 1,030 citations
- Large Language Models Are Zero-Shot Time Series ForecastersNate Gruver, Marc Finzi, Shikai Qiu, Andrew Gordon WilsonNeurIPS 2023 · 898 citations
- What Can Transformers Learn In-Context? A Case Study of Simple Function ClassesShivam Garg, Dimitris Tsipras, Percy Liang, Gregory ValiantNeurIPS 2022 · 883 citations
- A decoder-only foundation model for time-series forecastingAbhimanyu Das, Weihao Kong, Rajat Sen, Yichen ZhouICML 2024 · 601 citations
Related papers
- AutoElicit: Using Large Language Models for Expert Prior Elicitation in Predictive ModellingAlexander Capstick, Rahul G. Krishnan, Payam M. BarnaghiICML 2025
- BayesAgent: Bayesian Agentic Reasoning Under Uncertainty via Verbalized Probabilistic Graphical ModelingHengguan Huang, Xing Shen, Guang-Yuan Hao, Songtao Wang et al.AAAI 2026 · 2 citations
- Textual Bayes: Quantifying Prompt Uncertainty in LLM-Based SystemsBrendan Leigh Ross, Noël Vouitsis, Atiyeh Ashari Ghomi, Rasa Hosseinzadeh et al.ICLR 2026 · 8 citations
- Context is Key: A Benchmark for Forecasting with Essential Textual InformationAndrew Robert Williams, Arjun Ashok, Étienne Marcotte, Valentina Zantedeschi et al.ICML 2025
- LICO: Large Language Models for In-Context Molecular OptimizationTung Nguyen, Aditya GroverICLR 2025
