CAVA: A Visual Analytics System for Exploratory Columnar Data Augmentation Using Knowledge Graphs
Dylan Cashman, Shenyu Xu, Subhajit Das, Florian Heimerl, Cong Liu, Shah Rukh Humayoun, Michael Gleicher, Alex Endert, Remco Chang
摘要
Most visual analytics systems assume that all foraging for data happens before the analytics process; once analysis begins, the set of data attributes considered is fixed. Such separation of data construction from analysis precludes iteration that can enable foraging informed by the needs that arise in-situ during the analysis. The separation of the foraging loop from the data analysis tasks can limit the pace and scope of analysis. In this paper, we present CAVA, a system that integrates data curation and data augmentation with the traditional data exploration and analysis tasks, enabling information foraging in-situ during analysis. Identifying attributes to add to the dataset is difficult because it requires human knowledge to determine which available attributes will be helpful for the ensuing analytical tasks. CAVA crawls knowledge graphs to provide users with a a broad set of attributes drawn from external data to choose from. Users can then specify complex operations on knowledge graphs to construct additional attributes. CAVA shows how visual analytics can help users forage for attributes by letting users visually explore the set of available data, and by serving as an interface for query construction. It also provides visualizations of the knowledge graph itself to help users understand complex joins such as multi-hop aggregations. We assess the ability of our system to enable users to perform complex data combinations without programming in a user study over two datasets. We then demonstrate the generalizability of CAVA through two additional usage scenarios. The results of the evaluation confirm that CAVA is effective in helping the user perform data foraging that leads to improved analysis outcomes, and offer evidence in support of integrating data augmentation as a part of the visual analytics pipeline.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- fAIlureNotes: Supporting Designers in Understanding the Limits of AI Models for Computer Vision TasksSteven Moore, Q. Vera Liao, Hariharan SubramonyamCHI 2023 · 被引用 36 次
- Knowledge Graphs in Practice: Characterizing their Users, Challenges, and Visualization OpportunitiesHarry X. Li, Gabriel Appleby, Camelia Daniela Brumar, Remco Chang 等IEEE VIS 2023 · 被引用 35 次
- Incorporation of Human Knowledge into Data Embeddings to Improve Pattern Significance and InterpretabilityJie Li, Chun-qi ZhouIEEE VIS 2022 · 被引用 18 次
- Crowdsourced Think-Aloud StudiesZach Cutler, Lane Harrison, Carolina Nobre, Alexander LexCHI 2025 · 被引用 4 次
相关 Paper
- Envisage: Towards Expressive Visual Graph QueryingXiaolin Wen, Qishuang Fu, Shuangyue Han, Yichen Guo 等IEEE VIS 2025 · 被引用 6 次
- Boosting Visual Question Answering with Context-aware Knowledge AggregationGuohao Li, Xin Wang, Wenwu ZhuACM MM 2020 · 被引用 82 次
- KTabulator: Interactive Ad hoc Table Creation using Knowledge GraphsSiyuan Xia, Nafisa Anzum, Semih Salihoglu, Jian ZhaoCHI 2021 · 被引用 7 次
- PUREsuggest: Citation-Based Literature Search and Visual Exploration with Keyword-Controlled RankingsFabian BeckIEEE VIS 2024 · 被引用 4 次
- Characterizing Practices, Limitations, and Opportunities Related to Text Information Extraction Workflows: A Human-in-the-loop PerspectiveSajjadur Rahman, Eser KandoganCHI 2022 · 被引用 20 次
