MemX: An Attention-Aware Smart Eyewear System for Personalized Moment Auto-capture
Yuhu Chang, Yingying Zhao, Mingzhi Dong, Yujiang Wang, Yutian Lu, Qin Lv, Robert P. Dick, Tun Lu, Ning Gu, Li Shang
Abstract
This work presents MemX : a biologically-inspired attention-aware eyewear system developed with the goal of pursuing the long-awaited vision of a personalized visual Memex. MemX captures human visual attention on the fly, analyzes the salient visual content, and records moments of personal interest in the form of compact video snippets. Accurate attentive scene detection and analysis on resource-constrained platforms is challenging because these tasks are computation and energy intensive. We propose a new temporal visual attention network that unifies human visual attention tracking and salient visual content analysis. Attention tracking focuses computation-intensive video analysis on salient regions, while video analysis makes human attention detection and tracking more accurate. Using the YouTube-VIS dataset and 30 participants, we experimentally show that MemX significantly improves the attention tracking accuracy over the eye-tracking-alone method, while maintaining high system energy efficiency. We have also conducted 11 in-field pilot studies across a range of daily usage scenarios, which demonstrate the feasibility and potential benefits of MemX.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- PANDALens: Towards AI-Assisted In-Context Writing on OHMD During TravelsRunze Cai, Nuwan Janaka, Yang Chen, Lucia J. Wang et al.CHI 2024 · 36 citations
- Can Large Language Models Be Good Companions?: An LLM-Based Eyewear System with Conversational Common GroundZhenyu Xu, Hailin Xu, Zhouyang Lu, Yingying Zhao et al.UbiComp 2024 · 20 citations
- Do Smart Glasses Dream of Sentimental Visions?: Deep Emotionship Analysis for Eyewear DevicesYingying Zhao, Yuhu Chang, Yutian Lu, Yujiang Wang et al.UbiComp 2022 · 18 citations
- Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture UnderstandingZhuoming Li, Aitong Liu, Mengxi Jia, Yubo Lu et al.UbiComp 2026 · 1 citation
- Optimizing Product Placement for Virtual StoresWei Liang, Luhui Wang, Xinzhe Yu, Changyang Li et al.IEEE VR 2023 · 1 citation
Builds on4
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 615 citations
- Dynamic Face Video Segmentation via Reinforcement LearningYujiang Wang, Mingzhi Dong, Jie Shen, Yang Wu et al.CVPR 2020
- Detecting Attended Visual Targets in VideoEunji Chong, Yongxin Wang, Nataniel Ruiz, James M. RehgCVPR 2020
Related papers
- ActiveEye: Enabling Continuous and Responsive Video Understanding for Smart Eyewear SystemsZhenyu Xu, Tianlin Lu, Yingying Zhao, Yujiang Wang et al.UbiComp 2026 · 1 citation
- CASES: A Cognition-Aware Smart Eyewear System for Understanding How People ReadXiangyao Qi, Qi Lu, Wentao Pan, Yingying Zhao et al.UbiComp 2023 · 9 citations
- MyDJ: Sensing Food Intakes with an Attachable on Your Eyeglass FrameJaemin Shin, Seungjoo Lee, Taesik Gong, Hyungjun Yoon et al.CHI 2022 · 30 citations
- Attention-Aware Visualization: Tracking and Responding to User Perception Over TimeArvind Srinivasan, Johannes Ellemose, Peter W. S. Butcher, Panagiotis D. Ritsos et al.IEEE VIS 2024 · 9 citations
- Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head AttentionUttaran Bhattacharya, Gang Wu, Stefano Petrangeli, Viswanathan Swaminathan et al.ACM MM 2022 · 5 citations
