Learning Representations for Incomplete Time Series Clustering
Qianli Ma, Chuxin Chen, Sen Li, Garrison W. Cottrell
摘要
Time-series clustering is an essential unsupervised technique for data analysis, applied to many real-world fields, such as medical analysis and DNA microarray. Existing clustering methods are usually based on the assumption that the data is complete. However, time series in real-world applications often contain missing values. Traditional strategy (imputing first and then clustering) does not optimize the imputation and clustering process as a whole, which not only makes per- formance dependent on the combination of imputation and clustering methods but also fails to achieve satisfactory re- sults. How to best improve the clustering performance on incomplete time series remains a challenge. This paper pro- poses a novel unsupervised temporal representation learning model, named Clustering Representation Learning on Incom- plete time-series data (CRLI). CRLI jointly optimizes the im- putation and clustering process to impute more discrimina- tive values for clustering and make the learned representa- tions possessed good clustering property. Also, to reduce the error propagation from imputation to clustering, we introduce a discriminator to make the distribution of imputation values close to the true one and train CRLI in an alternating train- ing manner. An experiment conducted on eight real-world in- complete time-series datasets shows that CRLI outperforms existing methods. We demonstrates the effectiveness of the learned representations and the convergence of the model through visualization analysis. Moreover, we reveal that the joint training strategy can impute values close to the true ones in those important sub-sequences, and impute more discrim- inative values in those less important sub-sequences at the same time, making the imputed sequence cluster-friendly.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Representative Time Series Discovery for Data ExplorationGe Lee, Shixun Huang, Zhifeng Bao, Yanchang ZhaoVLDB 2025 · 被引用 2 次
- Rethinking Time-Series Imputation as Conditional Inference along Temporal EvolutionYu Fan, Yang Yang, guo yufan, Huazhong Yang 等ICML 2026
相关 Paper
- Self-Representation Subspace Clustering for Incomplete Multi-view DataJiyuan Liu, Xinwang Liu, Yi Zhang, Pei Zhang 等ACM MM 2021 · 被引用 98 次
- Integrating Sequence and Image Modeling in Irregular Medical Time Series Through Self-Supervised LearningLiuqing Chen, Shuhong Xiao, Shixian Ding, Shanhai Hu 等AAAI 2025 · 被引用 3 次
- Attribute-Missing Graph Clustering NetworkWenxuan Tu, Renxiang Guan, Sihang Zhou, Chuan Ma 等AAAI 2024 · 被引用 51 次
- COMPLETER: Incomplete Multi-View Clustering via Contrastive PredictionYijie Lin, Yuanbiao Gou, Zitao Liu, Boyun Li 等CVPR 2021
- Generative Semi-supervised Learning for Multivariate Time Series ImputationXiaoye Miao, Yangyang Wu, Jun Wang, Yunjun Gao 等AAAI 2021 · 被引用 212 次
