Time Series Representation for Visualization in Apache IoTDB
Lei Rui, Xiangdong Huang, Shaoxu Song, Yuyuan Kang, Chen Wang, Jianmin Wang
摘要
When analyzing time series, often interactively, the analysts frequently demand to visualize instantly large-scale data stored in databases. M4 visualization selects the first, last, bottom and top data points in each pixel column to ensure pixel-perfectness of the two-color line chart visualization. While M4 already shows its preciseness of encasing time series in different scales into a fixed size of pixels, how to efficiently support M4 representation in a time series native database is still absent. It is worth noting that, to enable fast writes, the commodity time series database systems, such as Apache IoTDB or InfluxDB, employ LSM-Tree based storage. That is, a time series is segmented and stored in a number of chunks, with possibly out-of-order arrivals, i.e., disordered on timestamps. To implement M4, a natural idea is to merge online the chunks as a whole series, with costly merge sort on timestamps, and then perform M4 representation as in relational databases. In this study, we propose a novel chunk merge free approach called M4-LSM to accelerate M4 representation and visualization. In particular, we utilize the metadata of chunks to prune and avoid the costly merging of any chunk. Moreover, intra-chunk indexing and pruning are enabled for efficiently accessing the representation points, referring to the special properties of time series. Remarkably, the time series database native operator M4-LSM has been implemented in Apache IoTDB, an open-source time series database, and deployed in companies across various industries. In the experiments over real-world datasets, the proposed M4-LSM operator demonstrates high efficiency without sacrificing preciseness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- KD-Box: Line-segment-based KD-tree for Interactive Exploration of Large-scale Time-Series DataYue Zhao, Yunhai Wang, Jian Zhang, Chi-Wing Fu 等IEEE VIS 2021 · 被引用 45 次
- Time Series Data Encoding for Efficient Storage: A Comparative Analysis in Apache IoTDBJinzhao Xiao, Yuxiang Huang, Changyu Hu, Shaoxu Song 等VLDB 2022 · 被引用 37 次
- On Repairing Timestamps for Regular Interval Time SeriesChenguang Fang, Shaoxu Song, Yinan MeiVLDB 2022 · 被引用 18 次
相关 Paper
- In-Database Time Series ClusteringYunxiang Su, Kenny Ye Liang, Shaoxu SongSIGMOD 2025 · 被引用 3 次
- Learning Autoregressive Model in LSM-Tree based StoreYunxiang Su, Wenxuan Ma, Shaoxu SongKDD 2023 · 被引用 4 次
- On Reducing Space Amplification with Multi-Column Compaction in Apache IoTDBChenguang Fang, Zijie Chen, Shaoxu Song, Xiangdong Huang 等VLDB 2024 · 被引用 1 次
- Distance-based Outlier Query Optimization in Apache IoTDBYunxiang Su, Shaoxu Song, Xiangdong Huang, Chen Wang 等VLDB 2024 · 被引用 2 次
- OM3: An Ordered Multi-level Min-Max Representation for Interactive Progressive Visualization of Time SeriesYunhai Wang, Yuchun Wang, Xin Chen, Yue Zhao 等SIGMOD 2023 · 被引用 8 次
