Model Sketching: Centering Concepts in Early-Stage Machine Learning Model Design
Michelle S. Lam, Zixian Ma, Anne Li, Izequiel Freitas, Dakuo Wang, James A. Landay, Michael S. Bernstein
Abstract
Machine learning practitioners often end up tunneling on low-level technical details like model architectures and performance metrics. Could early model development instead focus on high-level questions of which factors a model ought to pay attention to? Inspired by the practice of sketching in design, which distills ideas to their minimal representation, we introduce model sketching: a technical framework for iteratively and rapidly authoring functional approximations of a machine learning model’s decision-making logic. Model sketching refocuses practitioner attention on composing high-level, human-understandable concepts that the model is expected to reason over (e.g., profanity, racism, or sarcasm in a content moderation task) using zero-shot concept instantiation. In an evaluation with 17 ML practitioners, model sketching reframed thinking from implementation to higher-level exploration, prompted iteration on a broader range of model designs, and helped identify gaps in the problem formulation—all in a fraction of the time ordinarily required to build a model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6f484129-ba08-4954-bf81-e0e13de20ecdCited by top-tier papers10
- ChainForge: A Visual Toolkit for Prompt Engineering and LLM Hypothesis TestingIan Arawjo, Chelse Swoopes, Priyan Vaithilingam, Martin Wattenberg et al.CHI 2024 · 141 citations
- Concept Induction: Analyzing Unstructured Text with High-Level Concepts Using LLooMMichelle S. Lam, Janice Teoh, James A. Landay, Jeffrey Heer et al.CHI 2024 · 46 citations
- Sketching AI Concepts with Capabilities and Examples: AI Innovation in the Intensive Care UnitNur Yildirim, Susanna Zlotnikov, Deniz Sayar, Jeremy M. Kahn et al.CHI 2024 · 32 citations
- Policy Maps: Tools for Guiding the Unbounded Space of LLM BehaviorsMichelle S. Lam, Fred Hohman, Dominik Moritz, Jeffrey P. Bigham et al.UIST 2025 · 4 citations
- What We Talk About When We Talk About Frameworks in HCIShitao Fang, Koji Yatani, Kasper HornbækCHI 2026 · 4 citations
Builds on14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann et al.ICML 2020 · 1,233 citations
- The Hateful Memes Challenge: Detecting Hate Speech in Multimodal MemesDouwe Kiela, Hamed Firooz, Aravind Mohan, Vedanuj Goswami et al.NeurIPS 2020 · 1,022 citations
Related papers
- Computer-Aided Design as LanguageYaroslav Ganin, Sergey Bartunov, Yujia Li, Ethan Keller et al.NeurIPS 2021 · 129 citations
- Prompt Sketching for Large Language ModelsLuca Beurer-Kellner, Mark Niklas Müller, Marc Fischer, Martin T. VechevICML 2024 · 6 citations
- A study of UX practitioners roles in designing real-world, enterprise ML systemsSabah Zdanowska, Alex S. TaylorCHI 2022 · 38 citations
- An Exploratory Study of ML Sketches and Visual Code AssistantsLuís F. Gomes, Vincent J. Hellendoorn, Jonathan Aldrich, Rui AbreuICSE 2025 · 2 citations
- Unsupervised Causal Binary Concepts Discovery with VAE for Black-Box Model ExplanationThien Q. Tran, Kazuto Fukuchi, Youhei Akimoto, Jun SakumaAAAI 2022 · 11 citations
