Time-Gated Multi-Scale Flow Matching for Time-Series Imputation
Hangtian Wang, Mahito Sugiyama
Abstract
We propose to perform multivariate time-series imputation by learning the velocity field of a data-conditioned ordinary differential equation (ODE) via flow matching (FM). Our method, called Time-Gated Multi-Scale Flow Matching (TG-MSFM), conditions the flow on a structured endpoint comprising observed values, a per-time visibility mask, and short left/right context, processed by a time-aware Transformer whose self-attention is masked to aggregate only from observed timestamps. To reconcile global trends with local details along the trajectory, we introduce time-gated multi-scale velocity heads on a fixed 1D pyramid and blend them through a time-dependent gate; a mild anti-aliasing filter stabilizes the finest branch. At inference, we use a second-order Heun integrator with a per-step data-consistency (DC) projection that keeps observed coordinates exactly on the straight path from the initial noise to the endpoint, reducing boundary artifacts and drift. Training adopts gap-only supervision of the velocity on missing data coordinates, with small optional regularizers for numerical stability. Across standard benchmarks, TG-MSFM attains competitive or improved performance with favorable speed-quality trade-offs, and ablations demonstrate the isolated contributions of the time-gated multi-scale heads, masked attention, and the data-consistent ODE integration. RELATED WORK Time-series imputation with deep models. Early neural approaches address missingness by designing recurrent architectures and decay mechanisms to handle irregular observations (e.g., GRU-D and BRITS). More recent encoder-decoder designs rely on self-attention to aggregate long-range temporal context. Representative non-generative baselines in our study include Transformer variants and strong forecasting models that are often adapted to imputation, such as DLinear, TimesNet,
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51c154be-1a31-4b3d-9140-4bdd07fdfed9Builds on20
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 3,619 citations
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 1,720 citations
- iTransformer: Inverted Transformers Are Effective for Time Series ForecastingYong Liu, Tengge Hu, Haoran Zhang, Haixu Wu et al.ICLR 2024 · 1,703 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- CSDI: Conditional Score-based Diffusion Models for Probabilistic Time Series ImputationYusuke Tashiro, Jiaming Song, Yang Song, Stefano ErmonNeurIPS 2021 · 1,245 citations
Related papers
- Missing Value Imputation on Multidimensional Time SeriesParikshit Bansal, Prathamesh Deshpande, Sunita SarawagiVLDB 2021 · 90 citations
- T1: One-to-One Channel-Head Binding for Multivariate Time-Series ImputationDongik Park, Hyunwoo Ryu, Suahn Bae, Keondo Park et al.ICLR 2026 · 3 citations
- Spatiotemporal Imputation with Graph-Informed Flow MatchingZepeng Zhang, Aref Einizade, Jhony H. Giraldo, Olga FinkICML 2026
- Missing Data Imputation by Reducing Mutual Information with Rectified FlowsJiahao Yu, Qizhen Ying, Leyang Wang, Ziyue Jiang et al.NeurIPS 2025 · 9 citations
- A Transformer-based Framework for Multivariate Time Series Representation LearningGeorge Zerveas, Srideepika Jayaraman, Dhaval Patel, Anuradha Bhamidipaty et al.KDD 2021 · 66 citations
