This Time is Different: An Observability Perspective on Time Series Foundation Models
Ben Cohen, Emaad Khwaja, Youssef Doubli, Salahidine Lemaachi, Chris Lettieri, Charles Masson, Hugo Miccinilli, Elise Ramé, Qiqi Ren, Afshin Rostamizadeh, Jean Ogier du Terrail, Anna-Monica Toon
摘要
We introduce Toto, a time series forecasting foundation model with 151 million parameters. Toto uses a modern decoder-only architecture coupled with architectural innovations designed to account for specific challenges found in multivariate observability time series data. Toto's pre-training corpus is a mixture of observability data, open datasets, and synthetic data, and is 4-10 larger than those of leading time series foundation models. Additionally, we introduce BOOM, a large-scale benchmark consisting of 350 million observations across 2,807 real-world time series. For both Toto and BOOM, we source observability data exclusively from Datadog's own telemetry and internal observability metrics. Extensive evaluations demonstrate that Toto achieves state-of-the-art performance on both BOOM and on established general purpose time series forecasting benchmarks. Toto's model weights, inference code, and evaluation scripts, as well as BOOM's data and evaluation code, are all available as open source under the Apache 2.0 License available at https://huggingface.co/Datadog/Toto-Open-Base-1.0 and https://github.com/DataDog/toto.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context LearningAndreas Auer, Patrick Podest, Daniel Klotz, Sebastian Böck 等NeurIPS 2025 · 被引用 126 次
- SleepLM: Natural-Language Intelligence for Human SleepZongzhe Xu, Zitao Shuai, Eideen Mozaffari, Ravi Aysola 等ICML 2026 · 被引用 10 次
- It's TIME: Towards the Next Generation of Time Series Forecasting BenchmarksZhongzheng Qiao, SHENG PAN, Anni Wang, Viktoriya Zhukova 等ICML 2026 · 被引用 9 次
- OSF: On Pre-training and Scaling of Sleep Foundation ModelsZitao Shuai, Zongzhe Xu, David Yang, Wei Wang 等ICML 2026 · 被引用 8 次
- TelecomTS: A Multi-Modal Observability Dataset for Time Series and Language AnalysisAustin Feng, Andreas Varvarigos, Ioannis Panitsas, Daniela Fernandez 等ICML 2026 · 被引用 6 次
它引用的顶会 Paper30
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series ForecastingHaixu Wu, Jiehui Xu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 被引用 5,824 次
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun 等ICCV 2021 · 被引用 2,947 次
- iTransformer: Inverted Transformers Are Effective for Time Series ForecastingYong Liu, Tengge Hu, Haoran Zhang, Haixu Wu 等ICLR 2024 · 被引用 1,703 次
相关 Paper
- Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of ExpertsXiaoming Shi, Shiyu Wang, Yuqi Nie, Dianqi Li 等ICLR 2025
- MOMENT: A Family of Open Time-series Foundation ModelsMononito Goswami, Konrad Szafer, Arjun Choudhry, Yifu Cai 等ICML 2024 · 被引用 442 次
- Sundial: A Family of Highly Capable Time Series Foundation ModelsYong Liu, Guo Qin, Zhiyuan Shi, Zhi Chen 等ICML 2025
- Adapt Data to Model: Adaptive Transformation Optimization for Domain-shared Time Series Foundation ModelsYunzhong Qiu, Zhiyao Cen, Zhongyi Pei, Chen Wang 等ICLR 2026 · 被引用 1 次
- AutoTimes: Autoregressive Time Series Forecasters via Large Language ModelsYong Liu, Guo Qin, Xiangdong Huang, Jianmin Wang 等NeurIPS 2024 · 被引用 138 次
