SC2020Top-tier venue
Improving all-to-many personalized communication in two-phase I/O
Qiao Kang, Robert B. Ross, Robert Latham, Sunwoo Lee, Ankit Agrawal, Alok N. Choudhary, Wei-keng Liao
Abstract
As modern parallel computers enter the exascale era, the communication cost for redistributing requests becomes a significant bottleneck in MPIIO routines. The communication kernel for request redistribution, which has an all-to-many personalized communication pattern for application programs with a large number of noncontiguous requests, plays an essential role in the overall performance. This paper explores the available communication kernels for two-phase I/O communication. We generalize the spread-out algorithm to adapt to the all-to-many communication pattern of two-phase I/O by reducing the communication straggler effect. Communication throttling methods that reduce communication contention for asynchronous MPI implementation are adopted to improve communication performance further. Experimental results are presented using different communication kernels running on Cray XC40 Cori and IBM AC922 Summit supercomputers with different I/O patterns. Our study shows that adjusting communication kernel algorithms for different I/O patterns can improve the end-to-end performance up to 10 times compared with default MPI-IO implementations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9916383a-3460-4d64-964a-6289670d8aa2Cited by top-tier papers1
Ask how each one uses itRelated papers
- Efficient all-to-all Collective Communication Schedules for Direct-connect TopologiesPrithwish Basu, Liangyu Zhao, Jason Fantl, Siddharth Pal et al.HPDC 2024 · 7 citations
- Fine-grained Policy-driven I/O Sharing for Burst BuffersEd Karrels, Lei Huang, Yuhong Kan, Ishank Arora et al.SC 2023 · 5 citations
- Lessons Learned on MPI+Threads CommunicationRohit Zambre, Aparna ChandramowlishwaranSC 2022 · 5 citations
- Towards HPC I/O Performance Prediction through Large-scale Log AnalysisSunggon Kim, Alex Sim, Kesheng Wu, Suren Byna et al.HPDC 2020 · 34 citations
- Hitchhike: Efficient Request Submission via Deferred Enforcement of Address ContiguityXuda Zheng, Jian Zhou, Shuhan Bai, Runjin Wu et al.ASPLOS 2026
