Screen2Vec: Semantic Embedding of GUI Screens and GUI Components
Toby Jia-Jun Li, Lindsay Popowski, Tom M. Mitchell, Brad A. Myers
摘要
Representing the semantics of GUI screens and components is crucial to data-driven computational methods for modeling user-GUI interactions and mining GUI designs. Existing GUI semantic representations are limited to encoding either the textual content, the visual design and layout patterns, or the app contexts. Many representation techniques also require significant manual data annotation efforts. This paper presents Screen2Vec, a new self-supervised technique for generating representations in embedding vectors of GUI screens and components that encode all of the above GUI features without requiring manual annotation using the context of user interaction traces. Screen2Vec is inspired by the word embedding method Word2Vec, but uses a new two-layer pipeline informed by the structure of GUIs and interaction traces and incorporates screen- and app-specific metadata. Through several sample downstream tasks, we demonstrate Screen2Vec’s key useful properties: representing between-screen similarity through nearest neighbors, composability, and capability to represent user tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper39
- Understanding Design Collaboration Between Designers and Artificial Intelligence: A Systematic Literature ReviewYang Shi, Tian Gao, Xiaohan Jiao, Nan CaoCSCW 2023 · 被引用 170 次
- Enabling Conversational Interaction with Mobile UI using Large Language ModelsBryan Wang, Gang Li, Yang LiCHI 2023 · 被引用 149 次
- Jury Learning: Integrating Dissenting Voices into Machine Learning ModelsMitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel 等CHI 2022 · 被引用 134 次
- CanvasVAE: Learning to Generate Vector Graphic DocumentsKota YamaguchiICCV 2021 · 被引用 103 次
- Screen2Words: Automatic Mobile UI Summarization with Multimodal LearningBryan Wang, Gang Li, Xin Zhou, Zhourong Chen 等UIST 2021 · 被引用 97 次
它引用的顶会 Paper6
- Unblind your apps: predicting natural-language labels for mobile GUI components by deep learningJieshan Chen, Chunyang Chen, Zhenchang Xing, Xiwei Xu 等ICSE 2020 · 被引用 101 次
- Multi-Modal Repairs of Conversational Breakdowns in Task-Oriented DialogsToby Jia-Jun Li, Jingya Chen, Haijun Xia, Tom M. Mitchell 等UIST 2020 · 被引用 98 次
- Mapping Natural Language Instructions to Mobile UI Action SequencesYang Li, Jiacong He, Xin Zhou, Yuan Zhang 等ACL 2020 · 被引用 75 次
- GUIComp: A GUI Design Assistant with Real-Time, Multi-Faceted FeedbackChunggi Lee, Sanghoon Kim, Dongyun Han, Hongjun Yang 等CHI 2020 · 被引用 61 次
- Widget Captioning: Generating Natural Language Description for Mobile User Interface ElementsYang Li, Gang Li, Luheng He, Jingjie Zheng 等EMNLP 2020 · 被引用 46 次
相关 Paper
- Mouse2Vec: Learning Reusable Semantic Representations of Mouse BehaviourGuanhua Zhang, Zhiming Hu, Mihai Bâce, Andreas BullingCHI 2024 · 被引用 5 次
- Graph4GUI: Graph Neural Networks for Representing Graphical User InterfacesYue Jiang, Changkong Zhou, Vikas Garg, Antti OulasvirtaCHI 2024 · 被引用 17 次
- CanvasEmb: Learning Layout Representation with Large-scale Pre-training for Graphic DesignYuxi Xie, Danqing Huang, Jinpeng Wang, Chin-Yew LinACM MM 2021 · 被引用 11 次
- Object detection for graphical user interface: old fashioned or deep learning or a combination?Jieshan Chen, Mulong Xie, Zhenchang Xing, Chunyang Chen 等FSE 2020 · 被引用 144 次
- GUIGAN: Learning to Generate GUI Designs Using Generative Adversarial NetworksTianming Zhao, Chunyang Chen, Yuanning Liu, Xiaodong ZhuICSE 2021 · 被引用 58 次
