: A Visual Analytics Approach for Interactive Video Programming
Jianben He, Xingbo Wang, Kamkwai Wong, Xijie Huang, Changjian Chen, Zixin Chen, Fengjie Wang, Min Zhu, Huamin Qu
摘要
Constructing supervised machine learning models for real-world video analysis require substantial labeled data, which is costly to acquire due to scarce domain expertise and laborious manual inspection. While data programming shows promise in generating labeled data at scale with user-defined labeling functions, the high dimensional and complex temporal information in videos poses additional challenges for effectively composing and evaluating labeling functions. In this paper, we propose VideoPro, a visual analytics approach to support flexible and scalable video data programming for model steering with reduced human effort. We first extract human-understandable events from videos using computer vision techniques and treat them as atomic components of labeling functions. We further propose a two-stage template mining algorithm that characterizes the sequential patterns of these events to serve as labeling function templates for efficient data labeling. The visual interface of VideoPro facilitates multifaceted exploration, examination, and application of the labeling templates, allowing for effective programming of video data at scale. Moreover, users can monitor the impact of programming on model performance and make informed adjustments during the iterative programming process. We demonstrate the efficiency and effectiveness of our approach with two case studies and expert interviews.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action LocalizationYuchen He, Jianbing Lv, Liqi Cheng, Lingyu Meng 等CHI 2025 · 被引用 3 次
- DALL: Data Labeling via Data Programming and Active Learning Enhanced by Large Language ModelsGuozheng Li, Ao Wang, Shaoxiang Wang, Yu Zhang 等CHI 2026
它引用的顶会 Paper21
- PromptMagician: Interactive Prompt Engineering for Text-to-Image CreationYingchaojie Feng, Xingbo Wang, Kamkwai Wong, Sijia Wang 等IEEE VIS 2023 · 被引用 127 次
- A Visual Analytics Approach to Facilitate the Proctoring of Online ExamsHaotian Li, Min Xu, Yong Wang, Huan Wei 等CHI 2021 · 被引用 71 次
- Augmenting Sports Videos with VisCommentatorChen Zhu-Tian, Shuainan Ye, Xiangtong Chu, Haijun Xia 等IEEE VIS 2021 · 被引用 60 次
- Towards Visual Explainable Active Learning for Zero-Shot ClassificationShichao Jia, Zeyu Li, Nuo Chen, Jiawan ZhangIEEE VIS 2021 · 被引用 36 次
- VoiceCoach: Interactive Evidence-based Training for Voice Modulation Skills in Public SpeakingXingbo Wang, Haipeng Zeng, Yong Wang, Aoyu Wu 等CHI 2020 · 被引用 34 次
相关 Paper
- Visual Concept Programming: A Visual Analytics Approach to Injecting Human Intelligence at ScaleMd. Naimul Hoque, Wenbin He, Arvind Kumar Shekar, Liang Gou 等IEEE VIS 2022 · 被引用 22 次
- Reflecting on Design Paradigms of Animated Data Video ToolsLeixian Shen, Haotian Li, Yun Wang, Huamin QuCHI 2025 · 被引用 10 次
- Witan: Unsupervised Labelling Function Generation for Assisted Data ProgrammingBenjamin Denham, Edmund M.-K. Lai, Roopak Sinha, Muhammad Asif NaeemVLDB 2022 · 被引用 12 次
- OneLabeler: A Flexible System for Building Data Labeling ToolsYu Zhang, Yun Wang, Haidong Zhang, Bin Zhu 等CHI 2022 · 被引用 27 次
- VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in VideosBaoyu Liang, Qile Su, Shoutai Zhu, Yuchen Liang 等AAAI 2025 · 被引用 5 次
