WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition
Marius Bock, Hilde Kuehne, Kristof Van Laerhoven, Michael Möller
摘要
Research has shown the complementarity of camera- and inertial-based data for modeling human activities, yet datasets with both egocentric video and inertial-based sensor data remain scarce. In this paper, we introduce WEAR, an outdoor sports dataset for both vision- and inertial-based human activity recognition (HAR). Data from 22 participants performing a total of 18 different workout activities was collected with synchronized inertial (acceleration) and camera (egocentric video) data recorded at 11 different outside locations. WEAR provides a challenging prediction scenario in changing outdoor environments using a sensor placement, in line with recent trends in real-world applications. Benchmark results show that through our sensor placement, each modality interestingly offers complementary strengths and weaknesses in their prediction performance. Further, in light of the recent success of single-stage Temporal Action Localization (TAL) models, we demonstrate their versatility of not only being trained using visual data, but also using raw inertial data and being capable to fuse both modalities by means of simple concatenation. The dataset and code to reproduce experiments is publicly available via: mariusbock.github.io/wear/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Temporal Action Localization for Inertial-based Human Activity RecognitionMarius Bock, Michael Möller, Kristof Van LaerhovenUbiComp 2025 · 被引用 14 次
- Feasibility and Utility of Multimodal Micro Ecological Momentary Assessment on a SmartwatchHa Le, Veronika Potter, Rithika Lakshminarayanan, Varun Mishra 等CHI 2025 · 被引用 11 次
- MobHAR: Source-free Knowledge Transfer for Human Activity Recognition on Mobile DevicesMeng Xue, Yinan Zhu, Wentao Xie, Zhixian Wang 等UbiComp 2025 · 被引用 7 次
- SenseSeek Dataset: Multimodal Sensing to Study Information Seeking BehaviorsKaixin Ji, Danula Hettiachchi, Falk Scholer, Flora D. Salim 等UbiComp 2025 · 被引用 7 次
- EgoLife: Towards Egocentric Life AssistantJingkang Yang, Shuai Liu, Hongming Guo, Yuhao Dong 等CVPR 2025
它引用的顶会 Paper26
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu 等ICCV 2021 · 被引用 2,462 次
- BMN: Boundary-Matching Network for Temporal Action Proposal GenerationTianwei Lin, Xiao Liu, Xin Li, Errui Ding 等ICCV 2019 · 被引用 709 次
- MViTv2: Improved Multiscale Vision Transformers for Classification and DetectionYanghao Li, Chao-Yuan Wu, Haoqi Fan, Karttikeya Mangalam 等CVPR 2022 · 被引用 699 次
相关 Paper
- Sensor-Augmented Egocentric-Video Captioning with Dynamic Modal AttentionKatsuyuki Nakamura, Hiroki Ohashi, Mitsuhiro OkadaACM MM 2021 · 被引用 9 次
- EMHI: A Multimodal Egocentric Human Motion Dataset with HMD and Body-Worn IMUsZhen Fan, Peng Dai, Zhuo Su, Xu Gao 等AAAI 2025 · 被引用 13 次
- DETACH : Decomposed Spatio-Temporal Alignment for Exocentric Video and Ambient Sensors with Staged LearningJunho Yoon, Jaemo Jeong, Hyunju Kim, Dongman LeeCVPR 2026
- A Systematic Study of Unsupervised Domain Adaptation for Robust Human-Activity RecognitionYoungjae Chang, Akhil Mathur, Anton Isopoussu, Junehwa Song 等UbiComp 2020 · 被引用 136 次
- MMAct: A Large-Scale Dataset for Cross Modal Human Action UnderstandingQuan Kong, Ziming Wu, Ziwei Deng, Martin Klinkigt 等ICCV 2019 · 被引用 108 次
