GANSlider: How Users Control Generative Models for Images using Multiple Sliders with and without Feedforward Information
Hai Dang, Lukas Mecke, Daniel Buschek
摘要
We investigate how multiple sliders with and without feedforward visualizations influence users’ control of generative models. In an online study (N=138), we collected a dataset of people interacting with a generative adversarial network (StyleGAN2) in an image reconstruction task. We found that more control dimensions (sliders) significantly increase task difficulty and user actions. Visual feedforward partly mitigates this by enabling more goal-directed interaction. However, we found no evidence of faster or more accurate task performance. This indicates a tradeoff between feedforward detail and implied cognitive costs, such as attention. Moreover, we found that visualizations alone are not always sufficient for users to understand individual control dimensions. Our study quantifies fundamental UI design factors and resulting interaction behavior in this context, revealing opportunities for improvement in the UI design for interactive applications of generative models. We close by discussing design directions and further aspects.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- The Metacognitive Demands and Opportunities of Generative AILev Tankelevitch, Viktor Kewenig, Auste Simkute, Ava Elizabeth Scott 等CHI 2024 · 被引用 279 次
- DirectGPT: A Direct Manipulation Interface to Interact with Large Language ModelsDamien Masson, Sylvain Malacria, Géry Casiez, Daniel VogelCHI 2024 · 被引用 104 次
- PromptPaint: Steering Text-to-Image Generation Through Paint Medium-like InteractionsJohn Joon Young Chung, Eytan AdarUIST 2023 · 被引用 94 次
- Cells, Generators, and Lenses: Design Framework for Object-Oriented Interaction with Large Language ModelsTae Soo Kim, Yoonjoo Lee, Minsuk Chang, Juho KimUIST 2023 · 被引用 55 次
- PlantoGraphy: Incorporating Iterative Design Process into Generative Artificial Intelligence for Landscape RenderingRong Huang, Haichuan Lin, Chuanzhang Chen, Kang Zhang 等CHI 2024 · 被引用 45 次
它引用的顶会 Paper8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Manipulating and Measuring Model InterpretabilityForough Poursabzi-Sangdeh, Daniel G. Goldstein, Jake M. Hofman, Jennifer Wortman Vaughan 等CHI 2021 · 被引用 663 次
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu 等NeurIPS 2020 · 被引用 376 次
相关 Paper
- SliderSpace: Decomposing the Visual Capabilities of Diffusion ModelsRohit Gandikota, Zongze Wu, Richard Zhang, David Bau 等ICCV 2025 · 被引用 6 次
- StyleFactory: Towards Better Style Alignment in Image Creation through Style-Strength-Based Control and EvaluationMingxu Zhou, Dengming Zhang, Weitao You, Ziqi Yu 等UIST 2024 · 被引用 10 次
- StyleSpace Analysis: Disentangled Controls for StyleGAN Image GenerationZongze Wu, Dani Lischinski, Eli ShechtmanCVPR 2021
- StyleAvatar: Real-time Photo-realistic Portrait Avatar from a Single VideoLizhen Wang, Xiaochen Zhao, Jingxiang Sun, Yuxiang Zhang 等SIGGRAPH 2023 · 被引用 49 次
- ArtWhisperer: A Dataset for Characterizing Human-AI Interactions in Artistic CreationsKailas Vodrahalli, James ZouICML 2024 · 被引用 10 次
