Graph4GUI: Graph Neural Networks for Representing Graphical User Interfaces
Yue Jiang, Changkong Zhou, Vikas Garg, Antti Oulasvirta
Abstract
Present-day graphical user interfaces (GUIs) exhibit diverse arrangements of text, graphics, and interactive elements such as buttons and menus, but representations of GUIs have not kept up. They do not encapsulate both semantic and visuo-spatial relationships among elements. To seize machine learning’s potential for GUIs more efficiently, Graph4GUI exploits graph neural networks to capture individual elements’ properties and their semantic—visuo-spatial constraints in a layout. The learned representation demonstrated its effectiveness in multiple tasks, especially generating designs in a challenging GUI autocompletion task, which involved predicting the positions of remaining unplaced elements in a partially completed GUI. The new model’s suggestions showed alignment and visual appeal superior to the baseline method and received higher subjective ratings for preference. Furthermore, we demonstrate the practical benefits and efficiency advantages designers perceive when utilizing our model as an autocompletion plug-in.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 69c292b0-7771-4077-89d2-bbeceef1f6a2Cited by top-tier papers5
- Bridging Gulfs in UI Generation through Semantic GuidanceSeokhyeon Park, Soohyun Lee, Eugene Choi, Hyunwoo Kim et al.CHI 2026 · 3 citations
- DuetUI: A Bidirectional Context Loop for Human-Agent Co-Generation of Task-Oriented InterfacesYuan Xu, Shaowen Xiang, Yizhi Song, Ruoting Sun et al.CHI 2026 · 2 citations
- TaskAudit: Detecting Functiona11ity Errors in Mobile Apps via Agentic Task ExecutionMingyuan Zhong, Xia Chen, Davin Win Kyi, Chen Li et al.CHI 2026 · 1 citation
- AutoGameUI: Constructing High-Fidelity GameUI via Multimodal Correspondence MatchingZhongliang Tang, Qingrong Cheng, Mengchen Tan, Yongxiang Zhang et al.AAAI 2026 · 1 citation
- Belidor: A Specification Language for Operationalizing Structural Analogies Between User InterfacesMatthew T. Beaudouin-Lafon, Devamardeep Hayatpur, Arvind Satyanarayan, Haijun XiaCHI 2026 · 1 citation
Builds on11
- Generalization and Representational Limits of Graph Neural NetworksVikas K. Garg, Stefanie Jegelka, Tommi S. JaakkolaICML 2020 · 363 citations
- VINS: Visual Search for Mobile User Interface DesignSara Bunian, Kai Li, Chaima Jemmali, Casper Harteveld et al.CHI 2021 · 100 citations
- Multi-Modal Repairs of Conversational Breakdowns in Task-Oriented DialogsToby Jia-Jun Li, Jingya Chen, Haijun Xia, Tom M. Mitchell et al.UIST 2020 · 98 citations
- Screen2Words: Automatic Mobile UI Summarization with Multimodal LearningBryan Wang, Gang Li, Xin Zhou, Zhourong Chen et al.UIST 2021 · 97 citations
- Mapping Natural Language Instructions to Mobile UI Action SequencesYang Li, Jiacong He, Xin Zhou, Yuan Zhang et al.ACL 2020 · 75 citations
Related papers
- Optimizing User Interface Layouts via Gradient DescentPeitong Duan, Casimir Wierzynski, Lama NachmanCHI 2020 · 24 citations
- Screen2Vec: Semantic Embedding of GUI Screens and GUI ComponentsToby Jia-Jun Li, Lindsay Popowski, Tom M. Mitchell, Brad A. MyersCHI 2021 · 72 citations
- GUIGAN: Learning to Generate GUI Designs Using Generative Adversarial NetworksTianming Zhao, Chunyang Chen, Yuanning Liu, Xiaodong ZhuICSE 2021 · 58 citations
- EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement LearningYue Jiang, Zixin Guo, Hamed Rezazadegan Tavakoli, Luis A. Leiva et al.UIST 2024 · 15 citations
- Magic Layouts: Structural Prior for Component Detection in User Interface DesignsDipu Manandhar, Hailin Jin, John P. CollomosseCVPR 2021
