SC2022Top-tier venue
Lessons Learned on MPI+Threads Communication
Rohit Zambre, Aparna Chandramowlishwaran
Abstract
Hybrid MPI+threads programming is gaining prominence, but, in practice, applications perform slower with it compared to the MPI everywhere model. The most critical challenge to the parallel efficiency of MPI+threads applications is slow MPI_THREAD_MULTIPLE performance. MPI libraries have recently made significant strides on this front, but to exploit their capabilities, users must expose the communication parallelism in their MPI+threads applications. Recent studies show that MPI 4.0 provides users with new performance-oriented options to do so, but our evaluation of these new mechanisms shows that they pose several challenges. An alternative design is MPI Endpoints. In this paper, we present a comparison of the different designs from the perspective of MPI's end-users: domain scientists and application developers. We evaluate the mechanisms on metrics beyond performance such as usability, scope, and portability. Based on the lessons learned, we make a case for a future direction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5d0fdf4-ce81-4a55-9db6-a2d558471634Cited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Improving all-to-many personalized communication in two-phase I/OQiao Kang, Robert B. Ross, Robert Latham, Sunwoo Lee et al.SC 2020 · 12 citations
- KaMPIng: Flexible and (Near) Zero-Overhead C++ Bindings for MPITim Niklas Uhl, Matthias Schimek, Lukas Hübner, Demian Hespe et al.SC 2024 · 12 citations
- Graphite: A NUMA-aware HPC System for Graph Analytics Based on a new MPI * X Parallelism ModelMohammad Hasanzadeh-Mofrad, Rami G. Melhem, Muhammad Yousuf Ahmad, Mohammad HammoudVLDB 2020 · 142 citations
- Embracing Irregular Parallelism in HPC with YGMTrevor Steil, Tahsin Reza, Benjamin Priest, Roger PearceSC 2023 · 8 citations
- swKokkos: An Athread Backend for Enhanced Kokkos with the Sunway Heterogeneous ArchitectureJunlin Wei, Jinrong Jiang, Wu Wang, Chen Li et al.EuroSys 2026
