Agile Modeling: From Concept to Classifier in Minutes
Otilia Stretcu, Edward Vendrow, Kenji Hata, Krishnamurthy Viswanathan, Vittorio Ferrari, Sasan Tavakkol, Wenlei Zhou, Aditya Avinash, Enming Luo, Neil Gordon Alldrin, MohammadHossein Bateni, Gabriel Berger
摘要
The application of computer vision to nuanced subjective use cases is growing. While crowdsourcing has served the vision community well for most objective tasks (such as labeling a "zebra"), it now falters on tasks where there is substantial subjectivity in the concept (such as identifying "gourmet tuna"). However, empowering any user to develop a classifier for their concept is technically difficult: users are neither machine learning experts nor have the patience to label thousands of examples. In reaction, we introduce the problem of Agile Modeling: the process of turning any subjective visual concept into a computer vision model through a real-time user-in-the-loop interactions. We instantiate an Agile Modeling prototype for image classification and show through a user study (N=14) that users can create classifiers with minimal effort under 30 minutes. We compare this user driven process with the traditional crowdsourcing paradigm and find that the crowd's notion often differs from that of the user's, especially as the concepts become more subjective. Finally, we scale our experiments with simulations of users training classifiers for ImageNet21k categories to further demonstrate the efficacy. * Equal contribution. Sandwiches are NOT gourmet. This sandwich looks elegant.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM PromptingLeijie Wang, Kathryn Yurechko, Pranati Dani, Quan Ze Chen 等CHI 2025 · 被引用 7 次
- Clarify: Improving Model Robustness With Natural Language CorrectionsYoonho Lee, Michelle S. Lam, Helena Vasconcelos, Michael S. Bernstein 等UIST 2024 · 被引用 3 次
- ConCon-Chi: Concept-Context Chimera Benchmark for Personalized Vision-Language TasksAndrea Rosasco, Stefano Berti, Giulia Pasquale, Damiano Malafronte 等CVPR 2024 · 被引用 1 次
- Agile Deliberation: Concept Deliberation for Subjective Visual ClassificationLeijie Wang, Otilia Stretcu, Wei Qiao, Thomas Denby 等CVPR 2026
- Semantic and Expressive Variations in Image Captions Across LanguagesAndre Ye, Sebastin Santy, Jena D. Hwang, Amy X. Zhang 等CVPR 2025
它引用的顶会 Paper16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen 等ICML 2021 · 被引用 5,401 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
相关 Paper
- Modeling Collaborator: Enabling Subjective Vision Classification with Minimal Human Effort via LLM Tool-UseImad Eddine Toubal, Aditya Avinash, Neil Gordon Alldrin, Jan Dlabal 等CVPR 2024 · 被引用 9 次
- Rapid Image Labeling via Neuro-Symbolic LearningYifeng Wang, Zhi Tu, Yiwen Xiang, Shiyuan Zhou 等KDD 2023 · 被引用 3 次
- Supporting Co-Adaptive Machine Teaching through Human Concept Learning and Cognitive TheoriesSimret Araya Gebreegziabher, Yukun Yang, Elena L. Glassman, Toby Jia-Jun LiCHI 2025 · 被引用 5 次
- Usable and Fast Interactive Mental Face ReconstructionFlorian Strohm, Mihai Bâce, Andreas BullingUIST 2023 · 被引用 5 次
- What do You Mean? Interpreting Image Classification with Crowdsourced Concept Extraction and AnalysisAgathe Balayn, Panagiotis Soilis, Christoph Lofi, Jie Yang 等WWW 2021 · 被引用 31 次
