SC2021Top-tier venue
PAGANI: a parallel adaptive GPU algorithm for numerical integration
Ioannis Sakiotis, Kamesh Arumugam, Marc F. Paterno, Desh Ranjan, Balsa Terzic, Mohammad Zubair
Abstract
We present a new adaptive parallel algorithm for the challenging problem of multi-dimensional numerical integration on massively parallel architectures. Adaptive algorithms have demonstrated the best performance, but efficient many-core utilization is difficult to achieve because the adaptive work-load can vary greatly across the integration space and is impossible to predict a priori. Existing parallel algorithms utilize sequential computations on independent processors, which results in bottlenecks due to the need for data redistribution and processor synchronization. Our algorithm employs a high-throughput approach in which all existing sub-regions are processed and sub-divided in parallel. Repeated sub-region classification and filtering improves upon a brute-force approach and allows the algorithm to make efficient use of computation and memory resources. A CUDA implementation shows orders of magnitude speedup over the fastest open-source CPU method and extends the achievable accuracy for difficult integrands. Our algorithm typically outperforms other existing deterministic parallel methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- spECK: accelerating GPU sparse matrix-matrix multiplication through lightweight analysisMathias Parger, Martin Winter, Daniel Mlakar, Markus SteinbergerPPoPP 2020 · 48 citations
- Adaptive Workload-Balanced Scheduling Strategy for Global Ocean Data Assimilation on Massive GPUsJunmin Xiao, Chaoyang Shui, Di Cai, Kangyu Wang et al.SC 2023 · 1 citation
- Choosing the Best Parallelization and Implementation Styles for Graph Analytics Codes: Lessons Learned from 1106 ProgramsYiqian Liu, Noushin Azami, Avery Vanausdal, Martin BurtscherSC 2023 · 4 citations
- Towards Scalable Unstructured Mesh Computations on Shared Memory Many-CoresHaozhong Qiu, Chuanfu Xu, Jianbin Fang, Liang Deng et al.PPoPP 2024 · 8 citations
- A Programming Model for GPU Load BalancingMuhammad Osama, Serban D. Porumbescu, John D. OwensPPoPP 2023 · 10 citations
