Forging Time Series with Language: A Large Language Model Approach to Synthetic Data Generation
Cécile Rousseau, Tobia Boschi, Giandomenico Cornacchia, Dhaval Salwala, Alessandra Pascale, Juan Bernabé-Moreno
摘要
SDForger is a flexible and efficient framework for generating high-quality multivariate time series using LLMs. Leveraging a compact data representation, SDForger provides synthetic time series generation from a few samples and low-computation fine-tuning of any autoregressive LLM. Specifically, the framework transforms univariate and multivariate signals into tabular embeddings, which are then encoded into text and used to fine-tune the LLM. At inference, new textual embeddings are sampled and decoded into synthetic time series that retain the original data's statistical properties and temporal dynamics. Across a diverse range of datasets, SDForger outperforms existing generative models in many scenarios, both in similarity-based evaluations and downstream forecasting tasks. By enabling textual conditioning in the generation process, SDForger paves the way for multimodal modeling and the streamlined integration of time series with textual information. The model is open-sourced at https://github.com/IBM/fms-dgt/tree/main/fms_dgt/ public/databuilders/time_series. * [answer]." ( ) *+ , ) ∈ , + , 6 "Input: value_1 is [blank], value_2 is [blank] … value_6 is [blank]. [sep] Target: , ( ,) ) [answer] , ( , ) [answer] … , ( ,+ * [answer].
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- One Fits All: Power General Time Series Analysis by Pretrained LMTian Zhou, Peisong Niu, Xue Wang, Liang Sun 等NeurIPS 2023 · 被引用 1,178 次
- Time-LLM: Time Series Forecasting by Reprogramming Large Language ModelsMing Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu 等ICLR 2024 · 被引用 915 次
- Large Language Models Are Zero-Shot Time Series ForecastersNate Gruver, Marc Finzi, Shikai Qiu, Andrew Gordon WilsonNeurIPS 2023 · 被引用 898 次
- Neural SDEs as Infinite-Dimensional GANsPatrick Kidger, James Foster, Xuechen Li, Terry J. LyonsICML 2021 · 被引用 214 次
- Tiny Time Mixers (TTMs): Fast Pre-trained Models for Enhanced Zero/Few-Shot Forecasting of Multivariate Time SeriesVijay Ekambaram, Arindam Jati, Pankaj Dayama, Sumanta Mukherjee 等NeurIPS 2024 · 被引用 207 次
相关 Paper
- Latent Diffusion Transformer for Probabilistic Time Series ForecastingShibo Feng, Chunyan Miao, Zhong Zhang, Peilin ZhaoAAAI 2024 · 被引用 60 次
- Latent-to-Data Cascaded Diffusion Models for Unconditional Time Series GenerationLifeng Shen, Kai Syun Hou, Weiyu Chen, James T. KwokICLR 2026 · 被引用 15 次
- SDformer: Similarity-driven Discrete Transformer For Time Series GenerationZhicheng Chen, Shibo Feng, Zhong Zhang, Xi Xiao 等NeurIPS 2024 · 被引用 28 次
- M3Time: LLM-Enhanced Multi-Modal, Multi-Scale, and Multi-Frequency Multivariate Time Series ForecastingShuning Jia, Baijun Song, Canming Ye, Chun YuanAAAI 2026 · 被引用 1 次
- TimeCMA: Towards LLM-Empowered Multivariate Time Series Forecasting via Cross-Modality AlignmentChenxi Liu, Qianxiong Xu, Hao Miao, Sun Yang 等AAAI 2025 · 被引用 141 次
