Database Benchmarking for Supporting Real-Time Interactive Querying of Large Data
Leilani Battle, Philipp Eichmann, Marco Angelini, Tiziana Catarci, Giuseppe Santucci, Yukun Zheng, Carsten Binnig, Jean-Daniel Fekete, Dominik Moritz
Abstract
In this paper, we present a new benchmark to validate the suitability of database systems for interactive visualization workloads. While there exist proposals for evaluating database systems on interactive data exploration workloads, none rely on real user traces for database benchmarking. To this end, our long term goal is to collect user traces that represent workloads with different exploration characteristics. In this paper, we present an initial benchmark that focuses on "crossfilter"-style applications, which are a popular interaction type for data exploration and a particularly demanding scenario for testing database system performance. We make our benchmark materials, including input datasets, interaction sequences, corresponding SQL queries, and analysis code, freely available as a community resource, to foster further research in this area: https://osf.io/9xerb/?view_only= 81de1a3f99d04529b6b173a3bd5b4d23.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1eb6ad58-3c66-4aa1-9f9c-667205cf17daCited by top-tier papers9
- Synthesizing Natural Language to Visualization (NL2VIS) Benchmarks from NL2SQL BenchmarksYuyu Luo, Nan Tang, Guoliang Li, Chengliang Chai et al.SIGMOD 2021 · 90 citations
- An Evaluation-Focused Framework for Visualization Recommendation AlgorithmsZehua Zeng, Phoebe Moh, Fan Du, Jane Hoffswell et al.IEEE VIS 2021 · 35 citations
- LearnedSQLGen: Constraint-aware SQL Generation using Reinforcement LearningLixi Zhang, Chengliang Chai, Xuanhe Zhou, Guoliang LiSIGMOD 2022 · 26 citations
- Mosaic: An Architecture for Scalable & Interoperable Data ViewsJeffrey Heer, Dominik MoritzIEEE VIS 2023 · 23 citations
- Continuous Prefetch for Interactive Data ApplicationsHaneen Mohammed, Ziyun Wei, Ravi Netravali, Eugene WuVLDB 2020 · 16 citations
Builds on1
Related papers
- An Adaptive Benchmark for Modeling User Exploration of Large DatasetsJoanna Purich, Anthony Wise, Leilani BattleSIGMOD 2025 · 1 citation
- DSB: A Decision Support Benchmark for Workload-Driven and Traditional Database SystemsBailu Ding, Surajit Chaudhuri, Johannes Gehrke, Vivek R. NarasayyaVLDB 2021 · 62 citations
- HyBench: A New Benchmark for HTAP DatabasesChao Zhang, Guoliang Li, Tao LvVLDB 2024 · 28 citations
- Redbench: Workload Synthesis From Cloud TracesJohannes Wehrstein, Roman Heinrich, Mihail Stoian, Skander Krid et al.VLDB 2026 · 7 citations
- PBench: Workload Synthesizer with Real Statistics for Cloud Analytics BenchmarkingYan Zhou, Chunwei Liu, Bhuvan Urgaonkar, Zhengle Wang et al.VLDB 2025 · 4 citations
