SenseSeek Dataset: Multimodal Sensing to Study Information Seeking Behaviors
Kaixin Ji, Danula Hettiachchi, Falk Scholer, Flora D. Salim, Damiano Spina
Abstract
Information processing tasks involve complex cognitive mechanisms that are shaped by various factors, including individual goals, prior experience, and system environments. Understanding such behaviors requires a sophisticated and personalized data capture of how one interacts with modern information systems (e.g., web search engines). Passive sensors, such as wearables, capturing physiological and behavioral data, have the potential to provide solutions in this context. This paper presents a novel dataset, SenseSeek, designed to evaluate the effectiveness of consumer-grade sensors in a complex information processing scenario: searching via systems (e.g., search engines), one of the common strategies users employ for information seeking. The SenseSeek dataset comprises data collected from 20 participants, 235 trials of the stimulated search process, 940 phases of stages in the search process, including the realization of Information Need (IN), Query Formulation (QF), Query Submission by Typing (QS-T) or Speaking (QS-S), and Relevance Judgment by Reading (RJ-R) or Listening (RJ-L). The data includes Electrodermal Activities (EDA), Electroencephalogram (EEG), PUPIL, GAZE, and MOTION data, which were captured using consumer-grade sensors. It also contains 258 features extracted from the sensor data, the gaze-annotated screen recordings, and task responses. We validate the usefulness of the dataset by providing baseline analyses on the impacts of different cognitive intents and interaction modalities on the sensor data, and effectiveness of the data in discriminating the search stages. To our knowledge, SenseSeek is the first dataset that characterizes multiple stages involved in information seeking with physiological signals collected from multiple sensors. We hope this dataset can serve as a reference for future research on information-seeking behaviors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on15
- VREED: Virtual Reality Emotion Recognition Dataset Using Eye Tracking & Physiological MeasuresLuma Tabbaa, Ryan Searle, Saber Mirzaee Bafti, Md. Moinul Hossain et al.UbiComp 2022 · 114 citations
- The Low/High Index of Pupillary ActivityAndrew T. Duchowski, Krzysztof Krejtz, Nina A. Gehrer, Tanya Bafna et al.CHI 2020 · 91 citations
- A Critique of Electrodermal Activity Practices at CHIEbrahim Babaei, Benjamin Tag, Tilman Dingler, Eduardo VellosoCHI 2021 · 64 citations
- n-Gage: Predicting in-class Emotional, Behavioural and Cognitive Engagement in the WildNan Gao, Wei Shao, Mohammad Saiedur Rahaman, Flora D. SalimUbiComp 2020 · 60 citations
- WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity RecognitionMarius Bock, Hilde Kuehne, Kristof Van Laerhoven, Michael MöllerUbiComp 2025 · 50 citations
Related papers
- Characterizing Information Seeking Processes with Multiple Physiological SignalsKaixin Ji, Danula Hettiachchi, Flora D. Salim, Falk Scholer et al.SIGIR 2024 · 23 citations
- CLUES: Cognitive Load Understanding through Experimental Sensing DatasetAna Krstevska, Shivalika Goyal, Linda Fiorini, Francesco Bombassei De Bona et al.UbiComp 2026
- Capturing Team Cognition: A Multimodal Dataset for Adaptive Collaborative InterfacesChristopher Micek, Lasse Warnke, Lourenço Abrunhosa Rodrigues, Felix Putze et al.CHI 2026 · 1 citation
- The EasyCog Dataset: Towards Easier Cognitive Assessment with Passive Video WatchingQingyong Hu, Yuxuan Zhou, Jinjian Wang, Yanbin Gong et al.UbiComp 2026 · 2 citations
- Sensor-Augmented Egocentric-Video Captioning with Dynamic Modal AttentionKatsuyuki Nakamura, Hiroki Ohashi, Mitsuhiro OkadaACM MM 2021 · 9 citations
