MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers
Zhao Yanshun, Xiaoyu Peng, Jiamin Jiang, Congcong Zhu, Jingrun Chen
摘要
Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of local features. This inflexible tokenization scheme is inherently limited in its ability to efficiently represent and process solutions to complex PDEs. To address this, we propose MeshTok, an adaptive mesh refinement (AMR)-inspired tokenization and sequence modeling framework. This method selectively refines spatial regions exhibiting sharp gradients, transient features, or multiscale structures, generating a heterogeneous set of multiscale tokens defined on a fixed simulation grid. These tokens are processed within a unified Transformer sequence, enabling the model to simultaneously capture coarse-grained global context and fine-grained local details without requiring specialized architectural components. Although adaptive refinement moderately increases token count, it promotes a more targeted allocation of computational resources to physically informative regions, which we view as a practical inductive bias rather than a formal optimality guarantee. Experimental evaluations across multiple PDE families and benchmark datasets demonstrate that MeshTok consistently improves the efficiency-accuracy trade-off compared to uniform-grid baselines. This suggests adaptive multiscale tokenization as a scalable and generalizable design principle for neural PDE modeling. Code is available at https://github.com/ SCAILab-USTC/MeshTok .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image ClassificationChun-Fu (Richard) Chen, Quanfu Fan, Rameswar PandaICCV 2021 · 被引用 2,072 次
相关 Paper
- MSPT: Efficient Large-Scale Physical Modeling via Parallelized Multi-Scale AttentionPedro M. P. Curvo, Jan-Willem van de Meent, Maksim ZhdanovCVPR 2026 · 被引用 3 次
- AMR-Transformer: Enabling Efficient Long-range Interaction for Complex Neural Fluid SimulationZeyi Xu, Jinfan Liu, Kuangxu Chen, Ye Chen 等CVPR 2025
- Adaptive Physics Transformer with Fused Global-Local Attention for Subsurface Energy SystemsXin Ju, Hadrian Fung, Yuyan Zhang, Carl Jacquemyn 等ICML 2026 · 被引用 2 次
- SpiderSolver: A Geometry-Aware Transformer for Solving PDEs on Complex GeometriesKai Qi, Fan Wang, Zhewen Dong, Jian SunNeurIPS 2025 · 被引用 3 次
- Adaptive Patching for High-resolution Image Segmentation with TransformersEnzhi Zhang, Isaac Lyngaas, Peng Chen, Xiao Wang 等SC 2024 · 被引用 6 次
