Towards Scalable Unstructured Mesh Computations on Shared Memory Many-Cores
Haozhong Qiu, Chuanfu Xu, Jianbin Fang, Liang Deng, Jian Zhang, Qingsong Wang, Yue Ding, Zhe Dai, Yonggang Che, Shizhao Chen, Jie Liu
摘要
Due to data conflicts or data dependences, exploiting shared memory parallelism on unstructured mesh applications is highly challenging. The prior approaches are neither general nor scalable on emerging many-core processors. This paper presents a general and scalable shared memory approach for unstructured mesh computations. We recursively divide and reorder an unstructured mesh to construct a task dependency tree (TDT), where massive parallelism is exposed and data conflicts as well as data dependences are respected. We propose two recursion strategies to support popular programming models on both CPUs and GPUs for TDT. We evaluate our approach by applying it to an industrial unstructured Computational Fluid Dynamics (CFD) software. Experimental results show that our approach significantly outperforms the prior shared memory approaches, delivering up to 8.1× performance improvement over the engineer-tuned implementations.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- A Task-Parallel Approach for Localized Topological Data StructuresGuoxi Liu, Federico IuricichIEEE VIS 2023 · 被引用 9 次
- GALE: Leveraging Heterogeneous Systems for Efficient Unstructured Mesh Data AnalysisGuoxi Liu, Thomas Randall, Rong Ge, Federico IuricichIEEE VIS 2025 · 被引用 1 次
- Dynamic Mesh Processing on the GPUAhmed H. Mahmoud, Serban D. Porumbescu, John D. OwensSIGGRAPH 2025 · 被引用 4 次
- Designing a GPU-Accelerated Communication Layer for Efficient Fluid-Structure Interaction Computations on Heterogeneous SystemsAristotle X. Martin, Geng Liu, Bálint Joó, Runxin Wu 等SC 2024 · 被引用 1 次
- TD-NUCA: Runtime Driven Management of NUCA Caches in Task Dataflow Programming ModelsPaul Caheny, Lluc Alvarez, Marc Casas, Miquel MoretóSC 2022 · 被引用 4 次
