A Machine Teaching Framework for Scalable Recognition
Pei Wang, Nuno Vasconcelos
Abstract
We consider the scalable recognition problem in the fine-grained expert domain where large-scale data collection is easy whereas annotation is difficult. Existing solutions are typically based on semi-supervised or self-supervised learning. We propose an alternative new framework, MEMORABLE, based on machine teaching and online crowd-sourcing platforms. A small amount of data is first labeled by experts and then used to teach online annotators for the classes of interest, who finally label the entire dataset. Preliminary studies show that the accuracy of classifiers trained on the final dataset is a function of the accuracy of the student annotators. A new machine teaching algorithm, CMaxGrad, is then proposed to enhance this accuracy by introducing explanations in a state-of-the-art machine teaching algorithm. For this, CMaxGrad leverages counterfactual explanations, which take into account student predictions, thereby proving feedback that is student-specific, explicitly addresses the causes of student confusion, and adapts to the level of competence of the student. Experiments show that both MEMORABLE and CMaxGrad outperform existing solutions to their respective problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 426cda85-439c-4ffe-879e-0ad86d60170dCited by top-tier papers7
- Nonparametric Iterative Machine TeachingChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang et al.ICML 2023 · 13 citations
- Nonparametric Teaching of Implicit Neural RepresentationsChen Zhang, Steven Tin Sui Luo, Jason Chun Lok Li, Yik-Chung Wu et al.ICML 2024 · 12 citations
- Nonparametric Teaching for Multiple LearnersChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang et al.NeurIPS 2023 · 8 citations
- ML-Based Teaching Systems: A Conceptual FrameworkPhilipp Spitzer, Niklas Kühl, Daniel Heinz, Gerhard SatzgerCSCW 2023 · 6 citations
- Nonparametric Teaching of Attention LearnersChen Zhang, Jianghui Wang, Bingyang Cheng, Zhongtao Chen et al.ICLR 2026 · 3 citations
Builds on7
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 1,326 citations
- S4L: Self-Supervised Semi-Supervised LearningLucas Beyer, Xiaohua Zhai, Avital Oliver, Alexander KolesnikovICCV 2019 · 854 citations
- What Should Not Be Contrastive in Contrastive LearningTete Xiao, Xiaolong Wang, Alexei A. Efros, Trevor DarrellICLR 2021 · 338 citations
Related papers
- Gradient-Based Algorithms for Machine TeachingPei Wang, Kabir Nagrecha, Nuno VasconcelosCVPR 2021
- Crowd Teaching with Imperfect LabelsYao Zhou, Arun Reddy Nelakurthi, Ross Maciejewski, Wei Fan et al.WWW 2020 · 12 citations
- SCOUT: Self-Aware Discriminant Counterfactual ExplanationsPei Wang, Nuno VasconcelosCVPR 2020
- Crowdsourcing Learning as Domain Adaptation: A Case Study on Named Entity RecognitionXin Zhang, Guangwei Xu, Yueheng Sun, Meishan Zhang et al.ACL 2021
- Towards Professional Level Crowd Annotation of Expert Domain DataPei Wang, Nuno VasconcelosCVPR 2023
