High-dimensional Data Cubes
Sachin Basil John, Christoph Koch
Abstract
This paper introduces an approach to supporting high-dimensional data cubes at interactive query speeds and moderate storage cost. The approach is based on binary(-domain) data cubes that are judiciously partially materialized; the missing information can be quickly reconstructed using statistical or linear programming techniques. This enables new applications such as exploratory data analysis for feature engineering and other fields of data science. Moreover, it removes the need to compromise when building a data cube -all columns that we might ever wish to use can be included as dimensions. Our approach also speeds up certain dice, roll-up, and drill-down operations on data cubes with hierarchical dimensions compared to traditional data cubes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itRelated papers
- Index Intersection for High-Dimensional Range QueriesMaximilian Berens, Jens TeubnerVLDB 2026
- HDPView: Differentially Private Materialized View for Exploring High Dimensional Relational DataFumiyuki Kato, Tsubasa Takahashi, Shun Takagi, Yang Cao et al.VLDB 2022 · 8 citations
- Answering Multi-Dimensional Range Queries under Local Differential PrivacyJianyu Yang, Tianhao Wang, Ninghui Li, Xiang Cheng et al.VLDB 2021 · 46 citations
- Selecting Tangible Media for Immersive Exploration of Volumetric Scientific DataZhouhao Wu, Huiting Kong, Mingming Zhou, Qichen Liu et al.CHI 2026
- Multivariate Probabilistic Range Queries for Scalable Interactive 3D VisualizationAmani Ageeli, Alberto Jaspe-Villanueva, Ronell Sicat, Florian Mannuss et al.IEEE VIS 2022 · 2 citations
