Full-Duplex Inter-Group All-to-All Broadcast Algorithms with Optimal Bandwidth

Qiao Kang,Ankit Agrawal,Reda Al-Bahrani,Wei-Keng Liao,Alok Choudhary,Jesper Larsson Träff

doi:10.1145/3236367.3236374

Abstract

MPI inter-group collective communication patterns can be viewed as bipartite graphs that divide processes into two disjoint groups in which messages are transferred between but not within the groups. Such communication patterns can serve as basic operations for scientific application workflows. In this paper, we present parallel algorithms for inter-group all-to-all broadcast (Allgather) communication with optimal bandwidth for any message size and process number under single-port communication constraints. We implement the algorithms using MPI point-to-point and intra-group collective communication functions and evaluate their performance on the Cori supercomputer at NERSC. Using message sizes ranging from 256B to 64MB, the experiments show a significant performance improvement achieved by our algorithm, which is up to 9.27 times faster than production MPI libraries that adopt the so called root-gathering algorithm.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Full-Duplex Inter-Group All-to-All Broadcast Algorithms with Optimal Bandwidth

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Scalable Algorithms for MPI Intergroup Allgather and Allgatherv
Qiao Kang ... Wei-Keng Liao
Parallel Computing | VOL. 85
Qiao Kang, et. al.Qiao Kang ... Wei-Keng Liao
30 Apr 2019
Parallel Computing | VOL. 85

Proceedings of the 25th European MPI Users' Group Meeting
...
-
, et. al. ...
23 Sep 2018
23 Sep 2018

Adaptive Communication Algorithms for Distributed Heterogeneous Systems
Prashanth B Bhat ... C.S Raghavendra
Journal of Parallel and Distributed Computing | VOL. 59
Prashanth B Bhat, et. al.Prashanth B Bhat ... C.S Raghavendra
01 Nov 1999
Journal of Parallel and Distributed Computing | VOL. 59

Scalability analysis of parallel Particle-In-Cell codes on computational grids
Weifeng Tao ... Bertrand Lembege
Computer Physics Communications | VOL. 179
Weifeng Tao, et. al.Weifeng Tao ... Bertrand Lembege
26 Jul 2008
Computer Physics Communications | VOL. 179

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Full-Duplex Inter-Group All-to-All Broadcast Algorithms with Optimal Bandwidth

Abstract

Talk to us

Similar Papers