Article ID Journal Published Year Pages File Type
491729 Simulation Modelling Practice and Theory 2015 10 Pages PDF
Abstract

Significant research has been conducted in collective communication operations, in particular in MPI broadcast, on distributed memory platforms. Most of the research efforts aim to optimize the collective operations for particular architectures by taking into account either their topology or platform parameters. In this work we propose a simple but general approach to optimization of the legacy MPI broadcast algorithms, which are widely used in MPICH and Open MPI. The proposed optimization technique is designed to address the challenge of extreme scale of future HPC platforms. It is based on hierarchical transformation of the traditionally flat logical arrangement of communicating processors. Theoretical analysis and experimental results on IBM BlueGene/P and a cluster of the Grid’5000 platform are presented.

Related Topics
Physical Sciences and Engineering Computer Science Computer Science (General)
Authors
, , ,