Hierarchical State Space Models for Continuous Sequence-to-Sequence Modeling
Raunaq M. Bhirangi, Chenyu Wang, Venkatesh Pattabiraman, Carmel Majidi, Abhinav Gupta, Tess Lee Hellebrekers, Lerrel Pinto
摘要
Reasoning from sequences of raw sensory data is a ubiquitous problem across fields ranging from medical devices to robotics. These problems often involve using long sequences of raw sensor data (e.g. magnetometers, piezoresistors) to predict sequences of desirable physical quantities (e.g. force, inertial measurements). While classical approaches are powerful for locally-linear prediction problems, they often fall short when using real-world sensors. These sensors are typically non-linear, are affected by extraneous variables (e.g. vibration), and exhibit data-dependent drift. For many problems, the prediction task is exacerbated by small labeled datasets since obtaining ground-truth labels requires expensive equipment. In this work, we present Hierarchical State-Space models (HiSS), a conceptually simple, new technique for continuous sequential prediction. HiSS stacks structured state-space models on top of each other to create a temporal hierarchy. Across six real-world sensor datasets, from tactile-based state prediction to accelerometer-based inertial measurement, HiSS outperforms state-of-the-art sequence models such as causal Transformers, LSTMs, S4, and Mamba by at least 23% on MSE. Our experiments further indicate that HiSS demonstrates efficient scaling to smaller datasets and is compatible with existing data-filtering techniques. Code, datasets and videos can be found on https://hiss-csp.github.io CSP Bench Transformer LSTM Mamba S4 HiSS Magnetometer signal Marker position in x (cm) Normalized Mean Squared Error on CSP-Bench ( ) ↓ Time (in seconds)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- FreqMamba: Viewing Mamba from a Frequency Perspective for Image DerainingZhen Zou, Hu Yu, Jie Huang, Feng ZhaoACM MM 2024 · 被引用 73 次
- Language Modeling by Language ModelsJunyan Cheng, Peter Clark, Kyle RichardsonNeurIPS 2025 · 被引用 11 次
- Task-Optimized Convolutional Recurrent Networks Align with Tactile Processing in the Rodent BrainTrinity Chung, Yuchen Shen, Nathan C. L. Kong, Aran NayebiNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper11
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Combining Recurrent, Convolutional, and Continuous-time Models with Linear State Space LayersAlbert Gu, Isys Johnson, Karan Goel, Khaled Saab 等NeurIPS 2021 · 被引用 1,280 次
- Hyena Hierarchy: Towards Larger Convolutional Language ModelsMichael Poli, Stefano Massaroli, Eric Nguyen, Daniel Y. Fu 等ICML 2023 · 被引用 481 次
- Resurrecting Recurrent Neural Networks for Long SequencesAntonio Orvieto, Samuel L. Smith, Albert Gu, Anushan Fernando 等ICML 2023 · 被引用 474 次
- It's Raw! Audio Generation with State-Space ModelsKaran Goel, Albert Gu, Chris Donahue, Christopher RéICML 2022 · 被引用 257 次
相关 Paper
- Oscillatory State-Space ModelsT. Konstantin Rusch, Daniela RusICLR 2025
- Block-State TransformersJonathan Pilault, Mahan Fathi, Orhan Firat, Chris Pal 等NeurIPS 2023 · 被引用 33 次
- Accurate and Steady Inertial Pose Estimation through Sequence Structure Learning and ModulationYinghao Wu, Chaoran Wang, Lu Yin, Shihui Guo 等NeurIPS 2024 · 被引用 11 次
- The Expressive Capacity of State Space Models: A Formal Language PerspectiveYash Raj Sarrof, Yana Veitsman, Michael HahnNeurIPS 2024 · 被引用 53 次
- Jointly Modeling Spatio-Temporal Features of Tactile Signals for Action ClassificationJimmy Lin, Junkai Li, Jiasi Gao, Weizhi Ma 等AAAI 2024 · 被引用 3 次
