Jointly Modeling Spatio-Temporal Features of Tactile Signals for Action Classification
Jimmy Lin, Junkai Li, Jiasi Gao, Weizhi Ma, Yang Liu
摘要
Tactile signals collected by wearable electronics are essential in modeling and understanding human behavior. One of the main applications of tactile signals is action classification, especially in healthcare and robotics. However, existing tactile classification methods fail to capture the spatial and temporal features of tactile signals simultaneously, which results in sub-optimal performances. In this paper, we design Spatio-Temporal Aware tactility Transformer (STAT) to utilize continuous tactile signals for action classification. We propose spatial and temporal embeddings along with a new temporal pretraining task in our model, which aims to enhance the transformer in modeling the spatio-temporal features of tactile signals. Specially, the designed temporal pretraining task is to differentiate the time order of tubelet inputs to model the temporal properties explicitly. Experimental results on a public action classification dataset demonstrate that our model outperforms state-of-the-art methods in all metrics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun 等ICCV 2021 · 被引用 2,947 次
- Is Space-Time Attention All You Need for Video Understanding?Gedas Bertasius, Heng Wang, Lorenzo TorresaniICML 2021 · 被引用 2,927 次
- VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-TrainingZhan Tong, Yibing Song, Jue Wang, Limin WangNeurIPS 2022 · 被引用 2,336 次
- Multiview Transformers for Video RecognitionShen Yan, Xuehan Xiong, Anurag Arnab, Zhichao Lu 等CVPR 2022 · 被引用 279 次
相关 Paper
- Shrinking Temporal Attention in Transformers for Video Action RecognitionBonan Li, Pengfei Xiong, Congying Han, Tiande GuoAAAI 2022 · 被引用 19 次
- TextToucher: Fine-Grained Text-to-Touch GenerationJiahang Tu, Hao Fu, Fengyu Yang, Hanbin Zhao 等AAAI 2025 · 被引用 16 次
- Time Series as Images: Vision Transformer for Irregularly Sampled Time SeriesZekun Li, Shiyang Li, Xifeng YanNeurIPS 2023 · 被引用 145 次
- Structural Action Transformer for 3D Dexterous ManipulationXiaohan Lei, Min Wang, Bohong Weng, Wengang Zhou 等CVPR 2026
- Hierarchical State Space Models for Continuous Sequence-to-Sequence ModelingRaunaq M. Bhirangi, Chenyu Wang, Venkatesh Pattabiraman, Carmel Majidi 等ICML 2024 · 被引用 22 次
