: A Visual Analytics Approach for Interactive Video Programming
Jianben He, Xingbo Wang, Kamkwai Wong, Xijie Huang, Changjian Chen, Zixin Chen, Fengjie Wang, Min Zhu, Huamin Qu
Abstract
Constructing supervised machine learning models for real-world video analysis require substantial labeled data, which is costly to acquire due to scarce domain expertise and laborious manual inspection. While data programming shows promise in generating labeled data at scale with user-defined labeling functions, the high dimensional and complex temporal information in videos poses additional challenges for effectively composing and evaluating labeling functions. In this paper, we propose VideoPro, a visual analytics approach to support flexible and scalable video data programming for model steering with reduced human effort. We first extract human-understandable events from videos using computer vision techniques and treat them as atomic components of labeling functions. We further propose a two-stage template mining algorithm that characterizes the sequential patterns of these events to serve as labeling function templates for efficient data labeling. The visual interface of VideoPro facilitates multifaceted exploration, examination, and application of the labeling templates, allowing for effective programming of video data at scale. Moreover, users can monitor the impact of programming on model performance and make informed adjustments during the iterative programming process. We demonstrate the efficiency and effectiveness of our approach with two case studies and expert interviews.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 250da140-fb9e-4643-806f-ad52bdb8eee1Cited by top-tier papers2
- ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action LocalizationYuchen He, Jianbing Lv, Liqi Cheng, Lingyu Meng et al.CHI 2025 · 3 citations
- DALL: Data Labeling via Data Programming and Active Learning Enhanced by Large Language ModelsGuozheng Li, Ao Wang, Shaoxiang Wang, Yu Zhang et al.CHI 2026
Builds on21
- PromptMagician: Interactive Prompt Engineering for Text-to-Image CreationYingchaojie Feng, Xingbo Wang, Kamkwai Wong, Sijia Wang et al.IEEE VIS 2023 · 127 citations
- A Visual Analytics Approach to Facilitate the Proctoring of Online ExamsHaotian Li, Min Xu, Yong Wang, Huan Wei et al.CHI 2021 · 71 citations
- Augmenting Sports Videos with VisCommentatorChen Zhu-Tian, Shuainan Ye, Xiangtong Chu, Haijun Xia et al.IEEE VIS 2021 · 60 citations
- Towards Visual Explainable Active Learning for Zero-Shot ClassificationShichao Jia, Zeyu Li, Nuo Chen, Jiawan ZhangIEEE VIS 2021 · 36 citations
- VoiceCoach: Interactive Evidence-based Training for Voice Modulation Skills in Public SpeakingXingbo Wang, Haipeng Zeng, Yong Wang, Aoyu Wu et al.CHI 2020 · 34 citations
Related papers
- Visual Concept Programming: A Visual Analytics Approach to Injecting Human Intelligence at ScaleMd. Naimul Hoque, Wenbin He, Arvind Kumar Shekar, Liang Gou et al.IEEE VIS 2022 · 22 citations
- Reflecting on Design Paradigms of Animated Data Video ToolsLeixian Shen, Haotian Li, Yun Wang, Huamin QuCHI 2025 · 10 citations
- Witan: Unsupervised Labelling Function Generation for Assisted Data ProgrammingBenjamin Denham, Edmund M.-K. Lai, Roopak Sinha, Muhammad Asif NaeemVLDB 2022 · 12 citations
- OneLabeler: A Flexible System for Building Data Labeling ToolsYu Zhang, Yun Wang, Haidong Zhang, Bin Zhu et al.CHI 2022 · 27 citations
- VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in VideosBaoyu Liang, Qile Su, Shoutai Zhu, Yuchen Liang et al.AAAI 2025 · 5 citations
