Interactive Cleaning for Progressive Visualization through Composite Questions
Yuyu Luo, Chengliang Chai, Xuedi Qin, Nan Tang, Guoliang Li
摘要
In this paper, we study the problem of interactive cleaning for progressive visualization (ICPV): Given a bad visualization V , it is to obtain a "cleaned" visualization V whose distance is far from V , under a given (small) budget w.r.t. human cost. In ICPV, a system interacts with a user iteratively. During each iteration, it asks the user a data cleaning question such as "how to clean detected errors x?", and takes value updates from the user to clean V . Conventional wisdom typically picks a single question (e.g., "Are SIGMOD conference and SIGMOD the same?") with the maximum expected benefit in each iteration. We propose to use a composite questioni.e., a group of single questions to be treated as one question -in each iteration (for example, Are SIGMOD conference in t1 and SIGMOD in t2 the same value, and are t1 and t2 duplicates?). A composite question is presented to the user as a small connected graph through a novel GUI that the user can directly operate on. We propose algorithms to select the best composite question in each iteration. Experiments on real-world datasets verify that composite questions are more effective than asking single questions in isolation w.r.t. the human cost.
However, a visualization is not necessarily dirty, even if the data is dirty. Consider another example. Example 2: [A Correct Pie Chart] Figure 1(b) shows the proportion of the #-publications by Year and the result is
Id Year Title (abbr.) Venue Affiliation Citations t1 2013 NADEEF ACM SIGMOD QCRI 174.0 t2 2013 NADEEF SIGMOD Conf.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Natural Language to Visualization by Neural Machine TranslationYuyu Luo, Nan Tang, Guoliang Li, Jiawei Tang 等IEEE VIS 2021 · 被引用 145 次
- Synthesizing Natural Language to Visualization (NL2VIS) Benchmarks from NL2SQL BenchmarksYuyu Luo, Nan Tang, Guoliang Li, Chengliang Chai 等SIGMOD 2021 · 被引用 90 次
- Selective Data Acquisition in the Wild for Model ChargingChengliang Chai, Jiabin Liu, Nan Tang, Guoliang Li 等VLDB 2022 · 被引用 62 次
- Human-in-the-loop Outlier DetectionChengliang Chai, Lei Cao, Guoliang Li, Jian Li 等SIGMOD 2020 · 被引用 57 次
- HAIChart: Human and AI Paired Visualization SystemYupeng Xie, Yuyu Luo, Guoliang Li, Nan TangVLDB 2024 · 被引用 47 次
相关 Paper
- reVISit: Looking Under the Hood of Interactive Visualization StudiesCarolina Nobre, Dylan Wootton, Zach Cutler, Lane Harrison 等CHI 2021 · 被引用 20 次
- Composition and Configuration Patterns in Multiple-View VisualizationsXi Chen, Wei Zeng, Yanna Lin, Hayder Mahdi Al-Maneea 等IEEE VIS 2020 · 被引用 138 次
- CompositingVis: Exploring Interactions for Creating Composite Visualizations in Immersive EnvironmentsQian Zhu, Tao Lu, Shunan Guo, Xiaojuan Ma 等IEEE VIS 2024 · 被引用 11 次
- Competing Models: Inferring Exploration Patterns and Information Relevance via Bayesian Model SelectionShayan Monadjemi, Roman Garnett, Alvitta OttleyIEEE VIS 2020 · 被引用 21 次
- Does Interaction Improve Bayesian Reasoning with Visualization?Abigail Mosca, Alvitta Ottley, Remco ChangCHI 2021 · 被引用 14 次
