Information Geometry Loss for Time Series Forecasting
Jiayu Fang, Xuande Liu, Sangsha Fang, Zhen Tian, Hongwei Ma, Zhiqi Shao, Junbin Gao
摘要
Time series forecasting fundamentally involves learning probability distributions over future observations. However, existing loss functions rely on point-wise Euclidean metrics, neglecting the intrinsic geometric structure of probability distributions. This leads to suboptimal alignment between predicted and true distributions, particularly for uncertainty quantification. We propose InfoGeo Loss, a principled loss function grounded in information geometry that measures distributional discrepancies on statistical manifolds. Our approach comprises three key components: (1) a distribution parameterization module that models predictions with learnable sufficient statistics, (2) a Fisher information metric that quantifies intrinsic distributional distance, and (3) a Bregman divergence component that captures asymmetric prediction errors. We further introduce a natural gradient weighting strategy for efficient optimization on statistical manifolds. Theoretically, we prove statistical consistency and establish convergence guarantees. Extensive experiments on seven datasets with five architectures show that InfoGeo Loss consistently outperforms existing losses, achieving average improvements of 6.8% in MSE and 5.3% in MAE. The code is available at https://github.com/fangjiayu98/ infoGeo.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang 等ICML 2022 · 被引用 2,912 次
- One Fits All: Power General Time Series Analysis by Pretrained LMTian Zhou, Peisong Niu, Xue Wang, Liang Sun 等NeurIPS 2023 · 被引用 1,178 次
- Time-LLM: Time Series Forecasting by Reprogramming Large Language ModelsMing Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu 等ICLR 2024 · 被引用 915 次
相关 Paper
- Categorical Flow Matching on Statistical ManifoldsChaoran Cheng, Jiahan Li, Jian Peng, Ge LiuNeurIPS 2024 · 被引用 48 次
- Beyond MSE: Ordinal Cross-Entropy for Probabilistic Time Series ForecastingJieting Wang, Huimei Shi, Feijiang Li, Xiaolei ShangAAAI 2026
- InfoGlobe: Local-and-Global Information-Preserving Statistical Manifold Learning for Single-Cell TranscriptomicsCheng Wang, Jinpu Cai, Chongxiao Mao, Yuxuan Wang 等ICML 2026
- Fisher SAM: Information Geometry and Sharpness Aware MinimisationMinyoung Kim, Da Li, Shell Xu Hu, Timothy M. HospedalesICML 2022 · 被引用 96 次
- DistDF: Time-series Forecasting Needs Joint-distribution Wasserstein AlignmentEric Wang, Licheng Pan, Yuan Lu, Zhixuan Chu 等ICLR 2026 · 被引用 19 次
