arrow
Return

A framework for exploiting task and data parallelism on distributed memory multicomputers

delete1997-01-01
delete56
PRE
AI
S
S. Ramaswamy
S
Sachin S. Sapatnekar
P
P. Banerjee
DOI:10.1109/71.642945delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Distributed Memory Multicomputers (DMMs), such as the IBM SP-2, the Intel Paragon. and the Thinking Machines CM-5, offer significant advantages over shared memory multiprocessors in terms of cost and scalability, Unfortunately, the utilization of all the available computational power in these machines involves a tremendous programming effort on the part of users, which creates a need for sophisticated compiler and run-time support for distributed memory machines. In this paper, we explore a new compiler optimization for regular scientific applications-the simultaneous exploitation of task and data parallelism. Our optimization is implemented as part of the PARADIGM HPF compiler framework we have developed. The intuitive idea behind the optimization is the use of task parallelism to control the degree of data parallelism of individual tasks. The reason this provides increased performance is that data parallelism provides diminishing returns as the number of processors used is increased. By controlling the number of processors used for each data parallel task in an application and by concurrently executing these tasks, we make program execution more efficient and, therefore, faster. A practical implementation of a task and data parallel scheme of execution for an application on a distributed memory multicomputer also involves data redistribution. This data redistribution causes an overhead. However, as our experimental results show, this overhead is not a problem; execution of a program using task and data parallelism together can be significantly faster than its execution using data parallelism alone. This makes our proposed optimization practical and extremely useful.
Keywords:
task parallel
data parallel
allocation
scheduling
HPF
distributed memory
convex programming
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Parallel and Distributed Systems cover
IEEE Transactions on Parallel and Distributed Systems
IF:
6
Papers:
5.2K
Citations:
1.1W

Organization

No organization information available
Cited Papers

Cited Papers

The AMPTE IRM Spacecraft
err1985-05-01
err0
PREAI
errB. Hausler; F. Melzner; J. Stocker; A. Valenzuela; O. Bauer; P. Parigger; K. Sigritz; R. Schoning; E. Seidenschwang; F. Eberl; K.-h. Kaiser; W. Lieb; B. Merz; U. Pagel; E. Wiezorrek; J. Genzel
errShare
errSave
researcher View more