Visual Concept Programming: A Visual Analytics Approach to Injecting Human Intelligence at Scale
Md. Naimul Hoque, Wenbin He, Arvind Kumar Shekar, Liang Gou, Liu Ren
Abstract
Data-centric AI has emerged as a new research area to systematically engineer the data to land AI models for real-world applications. As a core method for data-centric AI, data programming helps experts inject domain knowledge into data and label data at scale using carefully designed labeling functions (e.g., heuristic rules, logistics). Though data programming has shown great success in the NLP domain, it is challenging to program image data because of a) the challenge to describe images using visual vocabulary without human annotations and b) lacking efficient tools for data programming of images. We present Visual Concept Programming, a first-of-its-kind visual analytics approach of using visual concepts to program image data at scale while requiring a few human efforts. Our approach is built upon three unique components. It first uses a self-supervised learning approach to learn visual representation at the pixel level and extract a dictionary of visual concepts from images without using any human annotations. The visual concepts serve as building blocks of labeling functions for experts to inject their domain knowledge. We then design interactive visualizations to explore and understand visual concepts and compose labeling functions with concepts without writing code. Finally, with the composed labeling functions, users can label the image data at scale and use the labeled data to refine the pixel-wise visual representation and concept quality. We evaluate the learned pixel-wise visual representation for the downstream task of semantic segmentation to show the effectiveness and usefulness of our approach. In addition, we demonstrate how our approach tackles real-world problems of image retrieval for autonomous driving.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 36790b57-a06c-4e5f-a39b-55914f405a91Cited by top-tier papers6
- LLM Comparator: Interactive Analysis of Side-by-Side Evaluation of Large Language ModelsMinsuk Kahng, Ian Tenney, Mahima Pushkarna, Michael Xieyang Liu et al.IEEE VIS 2024 · 23 citations
- : A Visual Analytics Approach for Interactive Video ProgrammingJianben He, Xingbo Wang, Kamkwai Wong, Xijie Huang et al.IEEE VIS 2023 · 17 citations
- ESCAPE: Countering Systematic Errors from Machine's Blind Spots via Interactive Visual AnalysisYongsu Ahn, Yu-Ru Lin, Panpan Xu, Zeng DaiCHI 2023 · 10 citations
- ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action LocalizationYuchen He, Jianbing Lv, Liqi Cheng, Lingyu Meng et al.CHI 2025 · 3 citations
- ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language ModelsHaoxuan Li, Zhen Wen, Qiqi Jiang, Chenxiao Li et al.IEEE VIS 2025 · 3 citations
Related papers
- Inspector Gadget: A Data Programming-based Labeling System for Industrial ImagesGeon Heo, Yuji Roh, Seonghyeon Hwang, Dayun Lee et al.VLDB 2021 · 9 citations
- Visual Programming: Compositional visual reasoning without trainingTanmay Gupta, Aniruddha KembhaviCVPR 2023
- A Vision Check-up for Language ModelsPratyusha Sharma, Tamar Rott Shaham, Manel Baradad, Adrián Rodríguez-Muñoz et al.CVPR 2024 · 10 citations
- Self-Training Large Language Models for Improved Visual Program Synthesis With Visual ReinforcementZaid Khan, Vijay Kumar B. G, Samuel Schulter, Yun Fu et al.CVPR 2024
- Bridging the gap to real-world language-grounded visual concept learningWhie Jung, Semin Kim, Junee Kim, Seunghoon HongNeurIPS 2025
