GANSlider: How Users Control Generative Models for Images using Multiple Sliders with and without Feedforward Information
Hai Dang, Lukas Mecke, Daniel Buschek
Abstract
We investigate how multiple sliders with and without feedforward visualizations influence users’ control of generative models. In an online study (N=138), we collected a dataset of people interacting with a generative adversarial network (StyleGAN2) in an image reconstruction task. We found that more control dimensions (sliders) significantly increase task difficulty and user actions. Visual feedforward partly mitigates this by enabling more goal-directed interaction. However, we found no evidence of faster or more accurate task performance. This indicates a tradeoff between feedforward detail and implied cognitive costs, such as attention. Moreover, we found that visualizations alone are not always sufficient for users to understand individual control dimensions. Our study quantifies fundamental UI design factors and resulting interaction behavior in this context, revealing opportunities for improvement in the UI design for interactive applications of generative models. We close by discussing design directions and further aspects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 855bed57-8ee6-45ed-9aae-ca24f0d4b3f0Cited by top-tier papers12
- The Metacognitive Demands and Opportunities of Generative AILev Tankelevitch, Viktor Kewenig, Auste Simkute, Ava Elizabeth Scott et al.CHI 2024 · 279 citations
- DirectGPT: A Direct Manipulation Interface to Interact with Large Language ModelsDamien Masson, Sylvain Malacria, Géry Casiez, Daniel VogelCHI 2024 · 104 citations
- PromptPaint: Steering Text-to-Image Generation Through Paint Medium-like InteractionsJohn Joon Young Chung, Eytan AdarUIST 2023 · 94 citations
- Cells, Generators, and Lenses: Design Framework for Object-Oriented Interaction with Large Language ModelsTae Soo Kim, Yoonjoo Lee, Minsuk Chang, Juho KimUIST 2023 · 55 citations
- PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape RenderingRong Huang, Haichuan Lin, Chuanzhang Chen, Kang Zhang et al.CHI 2024 · 45 citations
Builds on8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan et al.CHI 2021 · 663 citations
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu et al.NeurIPS 2020 · 376 citations
Related papers
- SliderSpace: Decomposing the Visual Capabilities of Diffusion ModelsRohit Gandikota, Zongze Wu, Richard Zhang, David Bau et al.ICCV 2025 · 6 citations
- StyleFactory: Towards Better Style Alignment in Image Creation through Style-Strength-Based Control and EvaluationMingxu Zhou, Dengming Zhang, Weitao You, Ziqi Yu et al.UIST 2024 · 10 citations
- StyleSpace Analysis: Disentangled Controls for StyleGAN Image GenerationZongze Wu, Dani Lischinski, Eli ShechtmanCVPR 2021
- StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single VideoLizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang et al.SIGGRAPH 2023 · 49 citations
- ArtWhisperer: A Dataset for Characterizing Human-AI Interactions in Artistic CreationsKailas Vodrahalli, James ZouICML 2024 · 10 citations
