An Empirical Investigation of Contextualized Number Prediction
Taylor Berg-Kirkpatrick, Daniel Spokoyny
Abstract
We conduct a large scale empirical investigation of contextualized number prediction in running text. Specifically, we consider two tasks: (1) masked number prediction -predicting a missing numerical value within a sentence, and (2) numerical anomaly detectiondetecting an errorful numeric value within a sentence. We experiment with novel combinations of contextual encoders and output distributions over the real number line. Specifically, we introduce a suite of output distribution parameterizations that incorporate latent variables to add expressivity and better fit the natural distribution of numeric values in running text, and combine them with both recurrent and transformer-based encoder architectures. We evaluate these models on two numeric datasets in the financial and scientific domain. Our findings show that output distributions that incorporate discrete latent variables and allow for multiple modes outperform simple flow-based counterparts on all datasets, yielding more accurate numerical prediction and anomaly detection. We also show that our models effectively utilize textual context and benefit from general-purpose unsupervised pretraining. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 471aea56-6233-45b5-889d-5e3040d85cf0Cited by top-tier papers9
- MultiHiertt: Numerical Reasoning over Multi Hierarchical Tabular and Textual DataYilun Zhao, Yunxiang Li, Chenying Li, Rui ZhangACL 2022 · 168 citations
- A Survey of Deep Learning for Mathematical ReasoningPan Lu, Liang Qiu, Wenhao Yu, Sean Welleck et al.ACL 2023 · 43 citations
- NextQuill: Causal Preference Modeling for Enhancing LLM PersonalizationXiaoyan Zhao, Juntao You, Yang Zhang, Wenjie Wang et al.ICLR 2026 · 38 citations
- A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language ModelsAlessandro Stolfo, Zhijing Jin, Kumar Shridhar, Bernhard Schölkopf et al.ACL 2023 · 15 citations
- W2W: Language-Model-Based Trajectory Prediction with Reinforcement LearningZirui Xu, Biao Yang, Rongrong Ni, Zhongkai Zhou et al.CVPR 2026 · 2 citations
Builds on6
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- Deep Learning For Symbolic MathematicsGuillaume Lample, François ChartonICLR 2020 · 477 citations
- S2ORC: The Semantic Scholar Open Research CorpusKyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney et al.ACL 2020 · 424 citations
- Neural Arithmetic UnitsAndreas Madsen, Alexander Rosenberg JohansenICLR 2020 · 53 citations
- Mathematical Reasoning in Latent SpaceDennis Lee, Christian Szegedy, Markus N. Rabe, Sarah M. Loos et al.ICLR 2020 · 36 citations
Related papers
- CONE: Embeddings for Complex Numerical Data Preserving Unit and Variable SemanticsGyanendra Shrestha, Anna Pyayt, Michael N. GubanovSIGMOD 2026
- Eliciting Numerical Predictive Distributions of LLMs Without Auto-RegressionJulianna Piskorz, Kasia Kobalczyk, Mihaela van der SchaarICLR 2026 · 2 citations
- Regress, Don't Guess: A Regression-like Loss on Number Tokens for Language ModelsJonas Zausinger, Lars Pennig, Anamarija Kozina, Sean Sdahl et al.ICML 2025
- Are Language Models Any Good at Density Modeling?Sriram Ranga, Sai Shashank Bedampeta, Rui Mao, Anupam ChattopadhyayAAAI 2026
- Methods for Numeracy-Preserving Word EmbeddingsDhanasekar Sundararaman, Shijing Si, Vivek Subramanian, Guoyin Wang et al.EMNLP 2020 · 28 citations
