Title
A scalable implementation of the NAS Parallel Benchmark BT on distributed memory systems
Abstract
In this paper, we describe an efficient and scalable implementation of the NAS Parallel Benchmark BT suitable for distributed memory systems such as the IBM Scalable POWERparallel Systems®. After describing the parallelization and data partitioning methods used, we outline some of the optimization steps used to realize good performance on individual processors and to reduce the communication overheads on the IBM SP1™ and SP2™ systems. We present performance results on up to 128 nodes of the SP1, and on the SP2 with wide nodes. We describe the performance on the standard Class A and Class B problem sets. To show the scalability of our parallelization methods, we present the performance of two additional data sets.
Year
DOI
Venue
1995
10.1147/sj.342.0273
IBM Systems Journal
Keywords
Field
DocType
memory system,parallel benchmark bt,scalable implementation,distributed memory
IBM,Computer architecture,Data set,Computer science,Parallel computing,Data partitioning,Distributed memory systems,Overhead (business),Scalability
Journal
Volume
Issue
ISSN
34
2
0018-8670
Citations 
PageRank 
References 
12
4.06
9
Authors
1
Name
Order
Citations
PageRank
V. K. Naik18014.68