GeoAggregator: An Efficient Transformer Model for Geo-Spatial Tabular Data
Rui Deng, Ziqi Li, Mingshu Wang
Abstract
Modeling geospatial tabular data with deep learning has become a promising alternative to traditional statistical and machine learning approaches. However, existing deep learning models often face challenges related to scalability and flexibility as datasets grow. To this end, this paper introduces GeoAggregator, an efficient and lightweight algorithm based on transformer architecture designed specifically for geospatial tabular data modeling. GeoAggregators explicitly account for spatial autocorrelation and spatial heterogeneity through Gaussian-biased local attention and global positional awareness. Additionally, we introduce a new attention mechanism that uses the Cartesian product to manage the size of the model while maintaining strong expressive power. We benchmark GeoAggregator against spatial statistical models, XGBoost, and several state-of-the-art geospatial deep learning methods using both synthetic and empirical geospatial datasets. The results demonstrate that GeoAggregators achieve the best or second-best performance compared to their competitors on nearly all datasets. GeoAggregator's efficiency is underscored by its reduced model size, making it both scalable and lightweight. Moreover, ablation experiments offer insights into the effectiveness of the Gaussian bias and Cartesian attention mechanism, providing recommendations for further optimizing the GeoAggregator's performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6c6cb400-5ce3-4513-8b35-6f1bd2d01557Cited by top-tier papers1
Ask how each one uses itBuilds on9
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- Perceiver IO: A General Architecture for Structured Inputs & OutputsAndrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch et al.ICLR 2022 · 797 citations
- STAN: Spatio-Temporal Attention Network for Next Location RecommendationYingtao Luo, Qiang Liu, Zhaocheng LiuWWW 2021 · 438 citations
- GeoCLIP: Clip-Inspired Alignment between Locations and Images for Effective Worldwide Geo-localizationVicente Vivanco Cepeda, Gaurav Kumar Nayak, Mubarak ShahNeurIPS 2023 · 303 citations
- Multi-Scale Representation Learning for Spatial Feature Distributions using Grid CellsGengchen Mai, Krzysztof Janowicz, Bo Yan, Rui Zhu et al.ICLR 2020 · 161 citations
Related papers
- Efficient Equivariant NetworkLingshen He, Yuxuan Chen, Zhengyang Shen, Yiming Dong et al.NeurIPS 2021 · 46 citations
- Towards Spatio- Temporal Aware Traffic Time Series ForecastingRazvan-Gabriel Cirstea, Bin Yang, Chenjuan Guo, Tung Kieu et al.ICDE 2022 · 137 citations
- CodedVTR: Codebook-based Sparse Voxel Transformer with Geometric GuidanceTianchen Zhao, Niansong Zhang, Xuefei Ning, He Wang et al.CVPR 2022 · 8 citations
- An Empirical Study of Spatial Attention Mechanisms in Deep NetworksXizhou Zhu, Dazhi Cheng, Zheng Zhang, Stephen Lin et al.ICCV 2019 · 522 citations
- GTA: A Geometry-Aware Attention Mechanism for Multi-View TransformersTakeru Miyato, Bernhard Jaeger, Max Welling, Andreas GeigerICLR 2024 · 51 citations
