IF-ConvTransformer: A Framework for Human Activity Recognition Using IMU Fusion and ConvTransformer
Ye Zhang, Longguang Wang, Huiling Chen, Aosheng Tian, Shilin Zhou, Yulan Guo
摘要
Recent advances in sensor based human activity recognition (HAR) have exploited deep hybrid networks to improve the performance. These hybrid models combine Convolutional Neural Networks (CNNs) and Recurrent Neural Networks (RNNs) to leverage their complementary advantages, and achieve impressive results. However, the roles and associations of different sensors in HAR are not fully considered by these models, leading to insufficient multi-modal fusion. Besides, the commonly used RNNs in HAR suffer from the 'forgetting' defect, which raises difficulties in capturing long-term information. To tackle these problems, an HAR framework composed of an Inertial Measurement Unit (IMU) fusion block and an applied ConvTransformer subnet is proposed in this paper. Inspired by the complementary filter, our IMU fusion block performs multi-modal fusion of commonly used sensors according to their physical relationships. Consequently, the features of different modalities can be aggregated more effectively. Then, the extracted features are fed into the applied ConvTransformer subnet for classification. Thanks to its convolutional subnet and self-attention layers, ConvTransformer can better capture local features and construct long-term dependencies. Extensive experiments on eight benchmark datasets demonstrate the superior performance of our framework. The source code will be published soon.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper8
- AutoAugHAR: Automated Data Augmentation for Sensor-based Human Activity RecognitionYexu Zhou, Haibin Zhao, Yiran Huang, Tobias Röddiger 等UbiComp 2024 · 被引用 25 次
- Temporal Action Localization for Inertial-based Human Activity RecognitionMarius Bock, Michael Möller, Kristof Van LaerhovenUbiComp 2025 · 被引用 14 次
- CALANet: Cheap All-Layer Aggregation for Human Activity RecognitionJaegyun Park, Dae-Won Kim, Jaesung LeeNeurIPS 2024 · 被引用 11 次
- Deep Heterogeneous Contrastive Hyper-Graph Learning for In-the-Wild Context-Aware Human Activity RecognitionWen Ge, Guanyi Mou, Emmanuel O. Agu, Kyumin LeeUbiComp 2024 · 被引用 9 次
- IMUCoCo: Enabling Flexible On-Body IMU Placement for Human Pose Estimation and Activity RecognitionHaozhe Zhou, Riku Arakawa, Yuvraj Agarwal, Mayank GoelUIST 2025 · 被引用 5 次
相关 Paper
- MMTSA: Multi-Modal Temporal Segment Attention Network for Efficient Human Activity RecognitionZiqi Gao, Yuntao Wang, Jianguo Chen, Junliang Xing 等UbiComp 2023 · 被引用 22 次
- rTsfNet: A DNN Model with Multi-head 3D Rotation and Time Series Feature Extraction for IMU-based Human Activity RecognitionYu EnokiboriUbiComp 2025 · 被引用 10 次
- MoPFormer: Motion-Primitive Transformer for Wearable-Sensor Activity RecognitionHao Zhang, Zhan Zhuang, Xuehao Wang, Xiaodong Yang 等NeurIPS 2025 · 被引用 11 次
- GIobalFusion: A Global Attentional Deep Learning Framework for Multisensor Information FusionShengzhong Liu, Shuochao Yao, Jinyang Li, Dongxin Liu 等UbiComp 2020 · 被引用 57 次
- COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity RecognitionBaiyu Chen, Wilson Wongso, Zechen Li, Yonchanok Khaokaew 等UbiComp 2026 · 被引用 1 次
