CASES: A Cognition-Aware Smart Eyewear System for Understanding How People Read
Xiangyao Qi, Qi Lu, Wentao Pan, Yingying Zhao, Rui Zhu, Mingzhi Dong, Yuhu Chang, Qin Lv, Robert P. Dick, Fan Yang, Tun Lu, Ning Gu, Li Shang
摘要
The process of reading has attracted decades of scientific research. Work in this field primarily focuses on using eye gaze patterns to reveal cognitive processes while reading. However, eye gaze patterns suffer from limited resolution, jitter noise, and cognitive biases, resulting in limited accuracy in tracking cognitive reading states. Moreover, using sequential eye gaze data alone neglects the linguistic structure of text, undermining attempts to provide semantic explanations for cognitive states during reading. Motivated by the impact of the semantic context of text on the human cognitive reading process, this work uses both the semantic context of text and visual attention during reading to more accurately predict the temporal sequence of cognitive states. To this end, we present a Cognition-Aware Smart Eyewear System (CASES), which fuses semantic context and visual attention patterns during reading. The two feature modalities are time-aligned and fed to a temporal convolutional network based multi-task classification deep model to automatically estimate and further semantically explain the reading state timeseries. CASES is implemented in eyewear and its use does not interrupt the reading process, thus reducing subjective bias. Furthermore, the real-time association between visual and semantic information enables the interactions between visual attention and semantic context to be better interpreted and explained. Ablation studies with 25 subjects demonstrate that CASES improves multi-label reading state estimation accuracy by 20.90% for sentence compared to eye tracking alone. Using CASES, we develop an interactive reading assistance system. Three and a half months of deployment with 13 in-field studies enables several observations relevant to the study of reading. In particular, observed how individual visual history interacts with the semantic context at different text granularities. Furthermore, CASES enables just-in-time intervention when readers encounter processing difficulties, thus promoting self-awareness of the cognitive process involved in reading and helping to develop more effective reading habits.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- Can Large Language Models Be Good Companions?: An LLM-Based Eyewear System with Conversational Common GroundZhenyu Xu, Hailin Xu, Zhouyang Lu, Yingying Zhao 等UbiComp 2024 · 被引用 20 次
- Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture UnderstandingZhuoming Li, Aitong Liu, Mengxi Jia, Yubo Lu 等UbiComp 2026 · 被引用 1 次
相关 Paper
- Reading Recognition in the WildCharig Yang, Samiul Alam, Shakhrul Iman Siam, Michael J. Proulx 等NeurIPS 2025 · 被引用 9 次
- Do Smart Glasses Dream of Sentimental Visions?: Deep Emotionship Analysis for Eyewear DevicesYingying Zhao, Yuhu Chang, Yutian Lu, Yujiang Wang 等UbiComp 2022 · 被引用 18 次
- WearSE: Enabling Streaming Speech Enhancement on Eyewear Using Acoustic SensingQian Zhang, Kaiyi Guo, Yifei Yang, Dong WangUbiComp 2025 · 被引用 7 次
- Acoustic-based Upper Facial Action Recognition for Smart EyewearWentao Xie, Qian Zhang, Jin ZhangUbiComp 2021 · 被引用 31 次
- MemX: An Attention-Aware Smart Eyewear System for Personalized Moment Auto-captureYuhu Chang, Yingying Zhao, Mingzhi Dong, Yujiang Wang 等UbiComp 2021 · 被引用 15 次
