Saliency in Augmented Reality
Huiyu Duan, Wei Shen, Xiongkuo Min, Danyang Tu, Jing Li, Guangtao Zhai
摘要
With the rapid development of multimedia technology, Augmented Reality (AR) has become a promising next-generation mobile platform. The primary theory underlying AR is human visual confusion, which allows users to perceive the real-world scenes and augmented contents (virtual-world scenes) simultaneously by superimposing them together. To achieve good Quality of Experience (QoE), it is important to understand the interaction between two scenarios, and harmoniously display AR contents. However, studies on how this superimposition will influence the human visual attention are lacking. Therefore, in this paper, we mainly analyze the interaction effect between background (BG) scenes and AR contents, and study the saliency prediction problem in AR. Specifically, we first construct a Saliency in AR Dataset (SARD), which contains 450 BG images, 450 AR images, as well as 1350 superimposed images generated by superimposing BG and AR images in pair with three mixing levels. A large-scale eye-tracking experiment among 60 subjects is conducted to collect eye movement data. To better predict the saliency in AR, we propose a vector quantized saliency prediction method and generalize it for AR saliency prediction. For comparison, three benchmark methods are proposed and evaluated together with our proposed method on our SARD. Experimental results demonstrate the superiority of our proposed method on both of the common saliency prediction problem and the AR saliency prediction problem over benchmark methods. Our dataset and code are available at: https://github.com/DuanHuiyu/ARSaliency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- ViDDAR: Vision Language Model-Based Task-Detrimental Content Detection for Augmented RealityYanming Xiu, Tim Scargill, Maria GorlatovaIEEE VR 2025 · 被引用 14 次
- ESIQA: Perceptual Quality Assessment of Vision-Pro-based Egocentric Spatial ImagesXilei Zhu, Liu Yang, Huiyu Duan, Xiongkuo Min 等IEEE VR 2025 · 被引用 10 次
- Unsupervised Ego- and Exo-centric Dense Procedural Activity Captioning via Gaze Consensus AdaptationZhaofeng Shi, Heqian Qiu, Lanxiao Wang, Qingbo Wu 等ACM MM 2025
它引用的顶会 Paper4
- Ego4D: Around the World in 3, 000 Hours of Egocentric VideoKristen Grauman, Andrew Westbury, Eugene Byrne, Zachary Chavis 等CVPR 2022 · 被引用 525 次
- End-to-End Human-Gaze-Target Detection with TransformersDanyang Tu, Xiongkuo Min, Huiyu Duan, Guodong Guo 等CVPR 2022 · 被引用 69 次
- Perceptual Quality Assessment of Omnidirectional ImagesYuming Fang, Liping Huang, Jiebin Yan, Xuelin Liu 等AAAI 2022 · 被引用 16 次
- Taming Transformers for High-Resolution Image SynthesisPatrick Esser, Robin Rombach, Björn OmmerCVPR 2021
相关 Paper
- Textured Mesh Saliency: Bridging Geometry and Texture for Human Perception in 3D GraphicsKaiwei Zhang, Dandan Zhu, Xiongkuo Min, Guangtao ZhaiAAAI 2025 · 被引用 1 次
- Comparison of Visual Saliency for Dynamic Point Clouds: Task-free vs. Task-dependentXuemei Zhou, Irene Viola, Silvia Rossi, Pablo CésarIEEE VR 2025 · 被引用 5 次
- FixationNet: Forecasting Eye Fixations in Task-Oriented Virtual EnvironmentsZhiming Hu, Andreas Bulling, Sheng Li, Guoping WangIEEE VR 2021 · 被引用 76 次
- SalBiNet360: Saliency Prediction on 360° Images with Local-Global Bifurcated Deep NetworkDongwen Chen, Chunmei Qing, Xiangmin Xu, Huansheng ZhuIEEE VR 2020 · 被引用 3 次
- SalientVR: saliency-driven mobile 360-degree video streaming with gaze informationShibo Wang, Shusen Yang, Hailiang Li, Xiaodan Zhang 等MobiCom 2022 · 被引用 43 次
