Context is Key: A Benchmark for Forecasting with Essential Textual Information
Andrew Robert Williams, Arjun Ashok, Étienne Marcotte, Valentina Zantedeschi, Jithendaraa Subramanian, Roland Riachi, James Requeima, Alexandre Lacoste, Irina Rish, Nicolas Chapados, Alexandre Drouin
摘要
Forecasting is a critical task in decision making across various domains. While numerical data provides a foundation, it often lacks crucial context necessary for accurate predictions. Human forecasters frequently rely on additional information, such as background knowledge or constraints, which can be efficiently communicated through natural language. However, the ability of existing forecasting models to effectively integrate this textual information remains an open question. To address this, we introduce "Context is Key" (CiK), a time series forecasting benchmark that pairs numerical data with diverse types of carefully crafted textual context, requiring models to integrate both modalities. We evaluate a range of approaches, including statistical models, time series foundation models, and LLMbased forecasters, and propose a simple yet effective LLM prompting method that outperforms all other tested methods on our benchmark. Our experiments highlight the importance of incorporating contextual information, demonstrate surprising performance when using LLM-based forecasting models, and also reveal some of their critical shortcomings. By presenting this benchmark, we aim to advance multimodal forecasting, promoting models that are both accurate and accessible to decision-makers with varied technical expertise. The benchmark can be visualized at https://servicenow.github.io/context-is-key-forecasting/v0/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Language in the Flow of Time: Time-Series-Paired Texts Weaved into a Unified Temporal NarrativeZihao Li, Xiao Lin, Zhining Liu, Jiaru Zou 等ICLR 2026 · 被引用 41 次
- True Zero-Shot Inference of Dynamical Systems Preserving Long-Term StatisticsChristoph Jürgen Hemmer, Daniel DurstewitzNeurIPS 2025 · 被引用 25 次
- Improving Time Series Forecasting via Instance-aware Post-hoc RevisionZhiding Liu, Mingyue Cheng, Guanhao Zhao, Jiqian Yang 等NeurIPS 2025 · 被引用 16 次
- Inferring Events from Time Series using Language ModelsMingtian Tan, Mike A. Merrill, Zachary Gottesman, Tim Althoff 等ACL 2026 · 被引用 7 次
- TimeRecipe: A Time-Series Forecasting Recipe via Benchmarking Module Level EffectivenessZhiyuan Zhao, Juntong Ni, Shangqing Xu, Haoxin Liu 等ICLR 2026 · 被引用 7 次
它引用的顶会 Paper9
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu 等ICLR 2022 · 被引用 4,966 次
- Time-LLM: Time Series Forecasting by Reprogramming Large Language ModelsMing Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu 等ICLR 2024 · 被引用 915 次
- Large Language Models Are Zero-Shot Time Series ForecastersNate Gruver, Marc Finzi, Shikai Qiu, Andrew Gordon WilsonNeurIPS 2023 · 被引用 898 次
- Unified Training of Universal Time Series Forecasting TransformersGerald Woo, Chenghao Liu, Akshat Kumar, Caiming Xiong 等ICML 2024 · 被引用 513 次
相关 Paper
- Rethinking Multimodal Time-Series Forecasting EvaluationHaoxin Liu, Yichen Zhou, Rajat Sen, B. Aditya Prakash 等ICML 2026 · 被引用 3 次
- M3Time: LLM-Enhanced Multi-Modal, Multi-Scale, and Multi-Frequency Multivariate Time Series ForecastingShuning Jia, Baijun Song, Canming Ye, Chun YuanAAAI 2026 · 被引用 1 次
- Temporal Knowledge Graph Forecasting Without Knowledge Using In-Context LearningDong-Ho Lee, Kian Ahrabian, Woojeong Jin, Fred Morstatter 等EMNLP 2023 · 被引用 30 次
- TFRBench: A Reasoning Benchmark for Evaluating Forecasting SystemsMd Atik Ahamed, Mihir Parmar, Palash Goyal, Yiwen Song 等ICML 2026
- ExoTimer: Leveraging Large Language Models for Time Series Forecasting with Exogenous VariablesLan Wu, Xuebin Wang, Chenglong Ge, Ruijuan Chu 等AAAI 2026
