Lune

ICML2026Top-tier venue

Dynamic Relational Priming Improves Transformer in Multivariate Time Series

Hunjae Lee, Corey Clark

2026Year

Abstract

Standard attention in transformers employ static token representations that remain unchanged across all pair-wise computations in each layer. This limits their representational alignment with the potentially diverse dynamics of each tokenpair interaction. While they excel in domains with relatively homogeneous relationships, standard attention may be inadequate in capturing heterogeneous inter-channel dependencies of multivariate time series (MTS) data where different channelpair interactions within a single system may be governed by entirely different physical laws or temporal dynamics. To better align the attention mechanism for such domain phenomena, we propose attention with dynamic relational priming (prime attention). Prime attention modulates token representations for each token-pair, optimizing each pair-wise interaction for that specific relationship. Results demonstrate that prime attention consistently outperforms standard attention across benchmarks, achieving up to 6.5% improvement in forecasting accuracy. In addition, prime attention achieves comparable performance using up to 40% less sequence length compared to standard attention, demonstrating its superior relational modeling capabilities and potential for data efficiency. Code is available at https://github.com/ timlee0131/Prime-Attention .

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext b72e1785-dc96-4422-8319-4034ae0bae5f

Builds on19

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines