Agile Modeling: From Concept to Classifier in Minutes
Otilia Stretcu, Edward Vendrow, Kenji Hata, Krishnamurthy Viswanathan, Vittorio Ferrari, Sasan Tavakkol, Wenlei Zhou, Aditya Avinash, Enming Luo, Neil Gordon Alldrin, MohammadHossein Bateni, Gabriel Berger
Abstract
The application of computer vision to nuanced subjective use cases is growing. While crowdsourcing has served the vision community well for most objective tasks (such as labeling a "zebra"), it now falters on tasks where there is substantial subjectivity in the concept (such as identifying "gourmet tuna"). However, empowering any user to develop a classifier for their concept is technically difficult: users are neither machine learning experts nor have the patience to label thousands of examples. In reaction, we introduce the problem of Agile Modeling: the process of turning any subjective visual concept into a computer vision model through a real-time user-in-the-loop interactions. We instantiate an Agile Modeling prototype for image classification and show through a user study (N=14) that users can create classifiers with minimal effort under 30 minutes. We compare this user driven process with the traditional crowdsourcing paradigm and find that the crowd's notion often differs from that of the user's, especially as the concepts become more subjective. Finally, we scale our experiments with simulations of users training classifiers for ImageNet21k categories to further demonstrate the efficacy. * Equal contribution. Sandwiches are NOT gourmet. This sandwich looks elegant.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM PromptingLeijie Wang, Kathryn Yurechko, Pranati Dani, Quan Ze Chen et al.CHI 2025 · 7 citations
- Clarify: Improving Model Robustness With Natural Language CorrectionsYoonho Lee, Michelle S. Lam, Helena Vasconcelos, Michael S. Bernstein et al.UIST 2024 · 3 citations
- ConCon-Chi: Concept-Context Chimera Benchmark for Personalized Vision-Language TasksAndrea Rosasco, Stefano Berti, Giulia Pasquale, Damiano Malafronte et al.CVPR 2024 · 1 citation
- Agile Deliberation: Concept Deliberation for Subjective Visual ClassificationLeijie Wang, Otilia Stretcu, Wei Qiao, Thomas Denby et al.CVPR 2026
- Semantic and Expressive Variations in Image Captions Across LanguagesAndre Ye, Sebastin Santy, Jena D. Hwang, Amy X. Zhang et al.CVPR 2025
Builds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen et al.ICML 2021 · 5,401 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
Related papers
- Modeling Collaborator: Enabling Subjective Vision Classification with Minimal Human Effort via LLM Tool-UseImad Eddine Toubal, Aditya Avinash, Neil Gordon Alldrin, Jan Dlabal et al.CVPR 2024 · 9 citations
- Rapid Image Labeling via Neuro-Symbolic LearningYifeng Wang, Zhi Tu, Yiwen Xiang, Shiyuan Zhou et al.KDD 2023 · 3 citations
- Supporting Co-Adaptive Machine Teaching through Human Concept Learning and Cognitive TheoriesSimret Araya Gebreegziabher, Yukun Yang, Elena L. Glassman, Toby Jia-Jun LiCHI 2025 · 5 citations
- Usable and Fast Interactive Mental Face ReconstructionFlorian Strohm, Mihai Bâce, Andreas BullingUIST 2023 · 5 citations
- What do You Mean? Interpreting Image Classification with Crowdsourced Concept Extraction and AnalysisAgathe Balayn, Panagiotis Soilis, Christoph Lofi, Jie Yang et al.WWW 2021 · 31 citations
