Design and Implementation of Broadcast Algorithms for Extreme-Scale Systems
Conference
·
OSTI ID:1042820
- ORNL
- Oak Ridge National Laboratory (ORNL)
The scalability and performance of collective communication operations limit the scalability and performance of many scientific applications. This paper presents two new blocking and nonblocking Broadcast algorithms for communicators with arbitrary communication topology, and studies their performance. These algorithms benefit from increased concurrency and a reduced memory footprint, making them suitable for use on large-scale systems. Measuring small, medium, and large data Broadcasts on a Cray-XT5, using 24,576 MPI processes, the Cheetah algorithms outperform the native MPI on that system by 51%, 69%, and 9%, respectively, at the same process count. These results demonstrate an algorithmic approach to the implementation of the important class of collective communications, which is high performing, scalable, and also uses resources in a scalable manner.
- Research Organization:
- Oak Ridge National Laboratory (ORNL)
- Sponsoring Organization:
- SC USDOE - Office of Science (SC)
- DOE Contract Number:
- AC05-00OR22725
- OSTI ID:
- 1042820
- Country of Publication:
- United States
- Language:
- English
Similar Records
Collective Framework and Performance Optimizations to Open MPI for Cray XT Platforms
Optimizing Blocking and Nonblocking Reduction Operations for Multicore Systems: Hierarchical Design and Implementation
Optimizing blocking and nonblocking reduction operations for multicore systems: Hierarchical design and implementation
Conference
·
Fri Dec 31 23:00:00 EST 2010
·
OSTI ID:1035529
Optimizing Blocking and Nonblocking Reduction Operations for Multicore Systems: Hierarchical Design and Implementation
Conference
·
Mon Dec 31 23:00:00 EST 2012
·
OSTI ID:1095156
Optimizing blocking and nonblocking reduction operations for multicore systems: Hierarchical design and implementation
Conference
·
Sun Sep 01 00:00:00 EDT 2013
· 2013 IEEE International Conference on Cluster Computing (CLUSTER)
·
OSTI ID:1567567