Design and Implementation of Broadcast Algorithms for Extreme-Scale Systems

Shamis, Pavel; Graham, Richard L; Gorentla Venkata, Manjunath; Ladd, Joshua

Title: Design and Implementation of Broadcast Algorithms for Extreme-Scale Systems

Conference · Sat Jan 01 00:00:00 EST 2011

OSTI ID:1042820

Shamis, Pavel ^[1]; Graham, Richard L ^[1]; Gorentla Venkata, Manjunath ^[2]; Ladd, Joshua ^[2]

ORNL
Oak Ridge National Laboratory (ORNL)

The scalability and performance of collective communication operations limit the scalability and performance of many scientific applications. This paper presents two new blocking and nonblocking Broadcast algorithms for communicators with arbitrary communication topology, and studies their performance. These algorithms benefit from increased concurrency and a reduced memory footprint, making them suitable for use on large-scale systems. Measuring small, medium, and large data Broadcasts on a Cray-XT5, using 24,576 MPI processes, the Cheetah algorithms outperform the native MPI on that system by 51%, 69%, and 9%, respectively, at the same process count. These results demonstrate an algorithmic approach to the implementation of the important class of collective communications, which is high performing, scalable, and also uses resources in a scalable manner.

OSTI does not have a digital full text copy available. For more information, please see document availability, search WorldCat, or search Google Scholar.

Cite

Export

Save

Research Organization:: Oak Ridge National Lab. (ORNL), Oak Ridge, TN (United States)

Sponsoring Organization:: USDOE Office of Science (SC)

DOE Contract Number:: DE-AC05-00OR22725

OSTI ID:: 1042820

Resource Relation:: Conference: IEEE Clusters 2011, Austin, TX, USA, 20110926, 20110926

Country of Publication:: United States

Language:: English

Similar Records

Optimizing Blocking and Nonblocking Reduction Operations for Multicore Systems: Hierarchical Design and Implementation

Conference · Tue Jan 01 00:00:00 EST 2013 · OSTI ID:1042820

Gorentla Venkata, Manjunath; Shamis, Pavel; Graham, Richard L; +2 more

Optimizing blocking and nonblocking reduction operations for multicore systems: Hierarchical design and implementation

Conference · Sun Sep 01 00:00:00 EDT 2013 · 2013 IEEE International Conference on Cluster Computing (CLUSTER) · OSTI ID:1042820

Venkata, Manjunath Gorentla; Shamis, Pavel; Sampath, Rahul; +2 more

ConnectX-2 CORE-Direct Enabled Asynchronous Broadcast Collective Communications

Conference · Sat Jan 01 00:00:00 EST 2011 · OSTI ID:1042820

Gorentla Venkata, Manjunath; Graham, Richard L; Ladd, Joshua S; +4 more

Related Subjects

99 GENERAL AND MISCELLANEOUS//MATHEMATICS, COMPUTING, AND INFORMATION SCIENCE
ALGORITHMS
COMMUNICATIONS
DESIGN
IMPLEMENTATION
PERFORMANCE
TOPOLOGY

Title: Design and Implementation of Broadcast Algorithms for Extreme-Scale Systems

Citation Formats

Similar Records

Related Subjects