TraSS: Efficient Trajectory Similarity Search Based on Key-Value Data Stores
Huajun He, Ruiyuan Li, Sijie Ruan, Tianfu He, Jie Bao, Tianrui Li, Yu Zheng
Abstract
Similarity search has recently become an integral part of many trajectory data analysis tasks. As the number of trajectories increases, we must find similar trajectories among massive trajectories, necessitating a scalable and efficient frame-work. Typically, massive trajectory data can be managed by key-value data stores. However, existing works with key-value data stores use a coarse representation to store trajectory data. Besides, they do not provide efficient query processing to search similar trajectories. Thus, this paper proposes TraSS, an efficient framework for trajectory similarity search in key-value data stores. We propose a novel spatial index, XZ*, which utilizes fine-grained index spaces with irregular shapes and sizes to represent trajectories elaborately. Further, we devise a bijective function from the index spaces of XZ* to continuous integers, which is simple but effective for query processing. To improve the efficiency of similarity search, we employ two steps to prune dissimilar trajectories: (1) global pruning. It leverages the XZ* index to prune index spaces with no trajectories similar to the query trajectory. Our global pruning can only pick out index spaces with similar sizes and shapes to the query trajectory. Compared to the state-of-the-art index, our global pruning reduces I/O overhead up to 66.4 % during query processing; (2) local filtering. It filters dissimilar trajectories in a way with low complexity. We use a few representative features extracted from a trajectory by the Douglas-Peucker algorithm to accelerate the local filtering. We implement an open-source toolkit (TraSS) on a popular key-value data store. Extensive experiments show that TraSS outperforms state-of-the-art solutions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Elf: Erasing-based Lossless Floating-Point CompressionRuiyuan Li, Zheng Li, Yi Wu, Chao Chen et al.VLDB 2023 · 44 citations
- Trajectory Similarity Measurement: An Efficiency PerspectiveYanchuan Chang, Egemen Tanin, Gao Cong, Christian S. Jensen et al.VLDB 2024 · 28 citations
Builds on2
Related papers
- TMan: A High-Performance Trajectory Data Management System Based on Key-Value StoresHuajun He, Zihang Xu, Ruiyuan Li, Jie Bao et al.ICDE 2024 · 12 citations
- Contrastive Trajectory Similarity Learning with Dual-Feature AttentionYanchuan Chang, Jianzhong Qi, Yuxuan Liang, Egemen TaninICDE 2023 · 77 citations
- Learning to Hash for Trajectory Similarity Computation and SearchLiwei Deng, Yan Zhao, Jin Chen, Shuncheng Liu et al.ICDE 2024 · 18 citations
- PPQ-Trajectory: Spatio-temporal Quantization for Querying in Large Trajectory RepositoriesShuang Wang, Hakan FerhatosmanogluVLDB 2021 · 12 citations
- Efficient Learning-based Top-k Representative Similar Subtrajectory QueryKunming Wang, Shiyu Yang, Jiabao Jin, Peng Cheng et al.ICDE 2024 · 2 citations
