Integrating Gaze and Mouse Via Joint Cross-Attention Fusion Net for Students' Activity Recognition in E-learning
Rongrong Zhu, Liang Shi, Yunpeng Song, Zhongmin Cai
摘要
E-learning has emerged as an indispensable educational mode in the post-epidemic era. However, this mode makes it difficult for students to stay engaged in learning without appropriate activity monitoring. Our work explores a promising solution that combines gaze and mouse data to recognize students' activities, thereby facilitating activity monitoring and analysis during e-learning. We initially surveyed 200 students from a local university, finding more acceptance for eye trackers and mouse loggers compared to video surveillance. We then designed eight students' routine digital activities to collect a multimodal dataset and analyze the patterns and correlations between gaze and mouse across various activities. Our proposed Joint Cross-Attention Fusion Net, a multimodal activity recognition framework, leverages the gaze-mouse relationship to yield improved classification performance by integrating cross-modal representations through a cross-attention mechanism and integrating the joint features that characterize gaze-mouse coordination. Evaluation results show that our method can achieve up to 94.87% F1-score in predicting 8-classes activities, with an improvement of at least 7.44% over using gaze or mouse data independently. This research illuminates new possibilities for monitoring student engagement in intelligent education systems, also suggesting a promising strategy for melding perception and action modalities in behavioral analysis across a range of ubiquitous computing environments.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- HOIGaze: Gaze Estimation During Hand-Object Interactions in Extended Reality Exploiting Eye-Hand-Head CoordinationZhiming Hu, Daniel F. B. Haeufle, Syn Schmitt, Andreas BullingSIGGRAPH 2025 · 被引用 3 次
- Predictive Student Modeling in Educational Games with Multi-Task LearningMichael Geden, Andrew Emerson, Jonathan P. Rowe, Roger Azevedo 等AAAI 2020 · 被引用 22 次
- Reading the Room: Automated, Momentary Assessment of Student Engagement in the Classroom: Are We There Yet?Betsy DiSalvo, Dheeraj Bandaru, Qiaosi Wang, Hong Li 等UbiComp 2022 · 被引用 16 次
- EgoBrain: Synergizing Minds and Eyes For Human Action UnderstandingNie Lin, Yansen Wang, Dongqi Han, Wei-Bang Jiang 等ICLR 2026 · 被引用 2 次
- Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human ActivitiesMichele Mazzamuto, Antonino Furnari, Yoichi Sato, Giovanni Maria FarinellaCVPR 2025
