Towards Visual Explainable Active Learning for Zero-Shot Classification
Shichao Jia, Zeyu Li, Nuo Chen, Jiawan Zhang
Abstract
Zero-shot classification is a promising paradigm to solve an applicable problem when the training classes and test classes are disjoint. Achieving this usually needs experts to externalize their domain knowledge by manually specifying a class-attribute matrix to define which classes have which attributes. Designing a suitable class-attribute matrix is the key to the subsequent procedure, but this design process is tedious and trial-and-error with no guidance. This paper proposes a visual explainable active learning approach with its design and implementation called semantic navigator to solve the above problems. This approach promotes human-AI teaming with four actions (ask, explain, recommend, respond) in each interaction loop. The machine asks contrastive questions to guide humans in the thinking process of attributes. A novel visualization called semantic map explains the current status of the machine. Therefore analysts can better understand why the machine misclassifies objects. Moreover, the machine recommends the labels of classes for each attribute to ease the labeling burden. Finally, humans can steer the model by modifying the labels interactively, and the machine adjusts its recommendations. The visual explainable active learning approach improves humans' efficiency of building zero-shot classification models interactively, compared with the method without guidance. We justify our results with user studies using the standard benchmarks for zero-shot classification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ece74e61-dfee-4901-a9d0-c4f4a9571634Cited by top-tier papers5
- On Selective, Mutable and Dialogic XAI: a Review of What Users Say about Different Types of Interactive ExplanationsAstrid Bertrand, Tiphaine Viard, Rafik Belloum, James R. Eagan et al.CHI 2023 · 53 citations
- : A Visual Analytics Approach for Interactive Video ProgrammingJianben He, Xingbo Wang, Kamkwai Wong, Xijie Huang et al.IEEE VIS 2023 · 17 citations
- AdversaFlow: Visual Red Teaming for Large Language Models with Multi-Level Adversarial FlowDazhen Deng, Chuhan Zhang, Huawei Zheng, Yuwen Pu et al.IEEE VIS 2024 · 14 citations
- Are We Closing the Loop Yet? Gaps in the Generalizability of VIS4ML ResearchHariharan Subramonyam, Jessica HullmanIEEE VIS 2023 · 10 citations
- OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI CollaborationJunwen Duan, Wei Xue, Ziyao Kang, Shixia Liu et al.IEEE VIS 2025 · 1 citation
Builds on1
Related papers
- Field-Guide-Inspired Zero-Shot LearningUtkarsh Mall, Bharath Hariharan, Kavita BalaICCV 2021 · 10 citations
- VGSE: Visually-Grounded Semantic Embeddings for Zero-Shot LearningWenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele et al.CVPR 2022 · 61 citations
- Counterfactual-Driven Zero-Shot Classifier ExpansionXiangyu Wang, Yanze Gao, Changxin Rong, Lyuzhou Chen et al.AAAI 2026 · 1 citation
- ALICE: Active Learning with Contrastive Natural Language ExplanationsWeixin Liang, James Zou, Zhou YuEMNLP 2020 · 36 citations
- Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual KnowledgeFawaz Sammani, Nikos DeligiannisNeurIPS 2024 · 15 citations
