Turbocharging Geospatial Visualization Dashboards via a Materialized Sampling Cube Approach
Jia Yu, Mohamed Sarwat
Abstract
In this paper, we present a middleware framework that runs on top of a SQL data system with the purpose of increasing the interactivity of geospatial visualization dashboards. The proposed system adopts a sampling cube approach that stores pre-materialized spatial samples and allows users to define their own accuracy loss function such that the produced samples can be used for various user-defined visualization tasks. The system ensures that the difference between the sample fed into the visualization dashboard and the raw query answer never exceeds the user-specified loss threshold. To reduce the number of cells in the sampling cube and hence mitigate the initialization time and memory utilization, the system employs two main strategies: (1) a partially materialized cube to only materialize local samples of those queries for which the global sample (the sample drawn from the entire dataset) exceeds the required accuracy loss threshold. (2) a sample selection technique that finds similarities between different local samples and only persists a few representative samples. Based on the extensive experimental evaluation, Tabula can bring down the total data-to-visualization time (including both data-system and visualization times) of a heat map generated over 700 million taxi rides to 600 milliseconds with 250 meters user-defined accuracy loss. Besides, Tabula costs up to two orders of magnitude less memory footprint (e.g., only 800 MB for the running example) and one order of magnitude less initialization time than the fully materialized sampling cube.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itRelated papers
- Marviq: Quality-Aware Geospatial Visualization of Range-Selection Queries Using MaterializationLiming Dong, Qiushi Bai, Taewoo Kim, Taiji Chen et al.SIGMOD 2020 · 3 citations
- Large-Scale Spatiotemporal Kernel Density VisualizationTsz Nam Chan, Pak Lon Ip, Bojian Zhu, Leong Hou U et al.ICDE 2025 · 6 citations
- An Adaptive Benchmark for Modeling User Exploration of Large DatasetsJoanna Purich, Anthony Wise, Leilani BattleSIGMOD 2025 · 1 citation
- Visualization-aware Time Series Min-Max Caching with Error Bound GuaranteesStavros Maroulis, Vassilis Stamatopoulos, George Papastefanatos, Manolis TerrovitisVLDB 2024 · 8 citations
- PilotDB: Database-Agnostic Online Approximate Query Processing with A Priori Error GuaranteesYuxuan Zhu, Tengjun Jin, Stefanos Baziotis, Chengsong Zhang et al.SIGMOD 2025 · 3 citations
