What Should You Know? A Human-In-the-Loop Approach to Unknown Unknowns Characterization in Image Recognition
Shahin Sharifi Noorian, Sihang Qiu, Ujwal Gadiraju, Jie Yang, Alessandro Bozzon
Abstract
Unknown unknowns represent a major challenge in reliable image recognition. Existing methods mainly focus on unknown unknowns identification, leveraging human intelligence to gather images that are potentially difficult for the machine. To drive a deeper understanding of unknown unknowns and more effective identification and treatment, this paper focuses on unknown unknowns characterization. We introduce a human-in-the-loop, semantic analysis framework for characterizing unknown unknowns at scale. We engage humans in two tasks that specify what a machine should know and describe what it really knows, respectively, both at the conceptual level, supported by information extraction and machine learning interpretability methods. Data partitioning and sampling techniques are employed to scale out human contributions in handling large data. Through extensive experimentation on scene recognition tasks, we show that our approach provides a rich, descriptive characterization of unknown unknowns and allows for more effective and cost-efficient detection than the state of the art.
• Computing methodologies → Machine learning; Knowledge representation and reasoning; • Human-centered computing → Human computer interaction (HCI).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0835d62f-1b8a-4b75-853a-0c912c85d8a7Cited by top-tier papers5
- Faulty or Ready? Handling Failures in Deep-Learning Computer Vision Models until Deployment: A Study of Practices, Challenges, and NeedsAgathe Balayn, Natasa Rikalo, Jie Yang, Alessandro BozzonCHI 2023 · 10 citations
- Hidden Indicators of Collective Intelligence in CrowdfundingEmoke-Ágnes Horvát, Henry Kudzanai Dambanemuya, Jayaram Uparna, Brian UzziWWW 2023 · 6 citations
- HybridEval: A Human-AI Collaborative Approach for Evaluating Design Ideas at ScaleSepideh Mesbah, Ines Arous, Jie Yang, Alessandro BozzonWWW 2023 · 5 citations
- RealBirdID: Benchmarking Bird Species Identification in the Era of MLLMsLogan Lawrence, Oindrila Saha, Rangel Daroya, Mustafa Chasmai et al.CVPR 2026
- Attribution Analysis-based Concept Alignment: A Human-in-the-loop Data Debugging FrameworkLei Chai, Lu Qi, Hailong Sun, Jing Zhang et al.AAAI 2026
Builds on4
- WinoGrande: An Adversarial Winograd Schema Challenge at ScaleKeisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, Yejin ChoiAAAI 2020 · 3,037 citations
- InfoGraph: Unsupervised and Semi-supervised Graph-Level Representation Learning via Mutual Information MaximizationFan-Yun Sun, Jordan Hoffmann, Vikas Verma, Jian TangICLR 2020 · 1,010 citations
- Towards Hybrid Human-AI Workflows for Unknown Unknown DetectionAnthony Z. Liu, Santiago Guerra, Isaac Fung, Gabriel Matute et al.WWW 2020 · 35 citations
- What do You Mean? Interpreting Image Classification with Crowdsourced Concept Extraction and AnalysisAgathe Balayn, Panagiotis Soilis, Christoph Lofi, Jie Yang et al.WWW 2021 · 31 citations
Related papers
- Are We Closing the Loop Yet? Gaps in the Generalizability of VIS4ML ResearchHariharan Subramonyam, Jessica HullmanIEEE VIS 2023 · 10 citations
- UMB: Understanding Model Behavior for Open-World Object DetectionXing Xi, Yangyang Huang, Zhijie Zhong, Ronghua LuoNeurIPS 2024 · 10 citations
- Learning to Characterize Matching ExpertsRoee Shraga, Ofra Amir, Avigdor GalICDE 2021 · 10 citations
- Detecting the Unexpected via Image ResynthesisKrzysztof Lis, Krishna Kanth Nakka, Pascal Fua, Mathieu SalzmannICCV 2019 · 217 citations
- Towards Professional Level Crowd Annotation of Expert Domain DataPei Wang, Nuno VasconcelosCVPR 2023
