RL-L: A Deep Reinforcement Learning Approach Intended for AR Label Placement in Dynamic Scenarios
Chen Zhu-Tian, Daniele Chiappalupi, Tica Lin, Yalong Yang, Johanna Beyer, Hanspeter Pfister
Abstract
Labels are widely used in augmented reality (AR) to display digital information. Ensuring the readability of AR labels requires placing them in an occlusion-free manner while keeping visual links legible, especially when multiple labels exist in the scene. Although existing optimization-based methods, such as force-based methods, are effective in managing AR labels in static scenarios, they often struggle in dynamic scenarios with constantly moving objects. This is due to their focus on generating layouts optimal for the current moment, neglecting future moments and leading to sub-optimal or unstable layouts over time. In this work, we present RL-LABEL, a deep reinforcement learning-based method intended for managing the placement of AR labels in scenarios involving moving objects. RL-LABEL considers both the current and predicted future states of objects and labels, such as positions and velocities, as well as the user's viewpoint, to make informed decisions about label placement. It balances the trade-offs between immediate and long-term objectives. We tested RL-LABEL in simulated AR scenarios on two real-world datasets, showing that it effectively learns the decision-making process for long-term optimization, outperforming two baselines (i.e., no view management and a force-based method) by minimizing label occlusions, line intersections, and label movement distance. Additionally, a user study involving 18 participants indicates that, within our simulated environment, RL-LABEL excels over the baselines in aiding users to identify, compare, and summarize data on labels in dynamic scenes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on14
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- SemanticAdapt: Optimization-based Adaptation of Mixed Reality Layouts Leveraging Virtual-Physical Semantic ConnectionsYifei Cheng, Yukang Yan, Xin Yi, Yuanchun Shi et al.UIST 2021 · 143 citations
- ShuttleSpace: Exploring and Analyzing Movement Trajectory in Immersive VisualizationShuainan Ye, Chen Zhu-Tian, Xiangtong Chu, Yifan Wang et al.IEEE VIS 2020 · 94 citations
- AdapTutAR: An Adaptive Tutoring System for Machine Tasks in Augmented RealityGaoping Huang, Xun Qian, Tianyi Wang, Fagun Patel et al.CHI 2021 · 93 citations
- Towards an Understanding of Situated AR Visualization for Basketball Free-Throw TrainingTica Lin, Rishi Singh, Yalong Yang, Carolina Nobre et al.CHI 2021 · 81 citations
Related papers
- Move2Hear: Active Audio-Visual Source SeparationSagnik Majumder, Ziad Al-Halah, Kristen GraumanICCV 2021 · 48 citations
- Towards Automatic Oracle Prediction for AR Testing: Assessing Virtual Object Placement Quality under Real-World ScenesXiaoyi Yang, Yuxing Wang, Tahmid Rafi, Dongfang Liu et al.ISSTA 2024 · 8 citations
- Physics-based Scene Layout Generation from Human MotionJianan Li, Tao Huang, Qingxu Zhu, Tien-Tsin WongSIGGRAPH 2024 · 5 citations
- BAR - A Reinforcement Learning Agent for Bounding-Box Automated RefinementMorgane Ayle, Jimmy Tekli, Julia El Zini, Boulos El Asmar et al.AAAI 2020 · 9 citations
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang et al.AAAI 2021 · 131 citations
