Reinforced Approximate Exploratory Data Analysis
Shaddy Garg, Subrata Mitra, Tong Yu, Yash Gadhia, Arjun Kashettiwar
摘要
Exploratory data analytics (EDA) is a sequential decision making process where analysts choose subsequent queries that might lead to some interesting insights based on the previous queries and corresponding results. Data processing systems often execute the queries on samples to produce results with low latency. Different downsampling strategy preserves different statistics of the data and have different magnitude of latency reductions. The optimum choice of sampling strategy often depends on the particular context of the analysis flow and the hidden intent of the analyst. In this paper, we are the first to consider the impact of sampling in interactive data exploration settings as they introduce approximation errors. We propose a Deep Reinforcement Learning (DRL) based framework which can optimize the sample selection in order to keep the analysis and insight generation flow intact. Evaluations with 3 real datasets show that our technique can preserve the original insight generation flow while improving the interaction latency, compared to baseline methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Approximate Caching for Efficiently Serving Text-to-Image Diffusion ModelsShubham Agarwal, Subrata Mitra, Sarthak Chakraborty, Srikrishna Karanam 等NSDI 2024 · 被引用 44 次
- SEIDEN: Revisiting Query Processing in Video Database SystemsJaeho Bang, Gaurav Tarlok Kakkar, Pramod Chunduri, Subrata Mitra 等VLDB 2023 · 被引用 24 次
- PairwiseHist: Fast, Accurate, and Space-Efficient Approximate Query Processing with Data CompressionAaron Hurst, Daniel E. Lucani, Qi ZhangVLDB 2024 · 被引用 5 次
它引用的顶会 Paper6
- Conversational Contextual Bandit: Algorithm and ApplicationXiaoying Zhang, Hong Xie, Hang Li, John C. S. LuiWWW 2020 · 被引用 97 次
- Multi-Agent Task-Oriented Dialog Policy Learning with Role-Aware Reward DecompositionRyuichi Takanobu, Runze Liang, Minlie HuangACL 2020 · 被引用 47 次
- Reinforcement Learning from Reformulations in Conversational Question Answering over Knowledge GraphsMagdalena Kaiser, Rishiraj Saha Roy, Gerhard WeikumSIGIR 2021 · 被引用 45 次
- Few-Shot Complex Knowledge Base Question Answering via Meta Reinforcement LearningYuncheng Hua, Yuan-Fang Li, Gholamreza Haffari, Guilin Qi 等EMNLP 2020 · 被引用 36 次
- Conditional Generative Model Based Predicate-Aware Query ApproximationNikhil Sheoran, Subrata Mitra, Vibhor Porwal, Siddharth Ghetia 等AAAI 2022 · 被引用 14 次
相关 Paper
- Guided Exploration of User GroupsMariia Seleznova, Behrooz Omidvar-Tehrani, Sihem Amer-Yahia, Eric SimonVLDB 2020 · 被引用 24 次
- Guided Exploration of Data SummariesBrit Youngmann, Sihem Amer-Yahia, Aurélien PersonnazVLDB 2022 · 被引用 22 次
- Supporting Guided Exploratory Visual Analysis on Time Series Data with Reinforcement LearningYang Shi, Bingchang Chen, Ying Chen, Zhuochen Jin 等IEEE VIS 2023 · 被引用 8 次
- Interactive Search with Reinforcement LearningWeicheng Wang, Victor Junqiu Wei, Min Xie, Di Jiang 等ICDE 2025 · 被引用 1 次
- Holistic query Approximation via RL ModelingSusan B. Davidson, Tova Milo, Kathy Razmadze, Gal ZeeviVLDB 2025
