arrow
返回

Computing programs containing band linear recurrences on vector supercomputers

delete1996-01-01
delete4
PRE
AI
A
A. Nicolau
S
Siu, KYS
DOI:10.1109/71.532109delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Many large-scale scientific and engineering computations, e.g., some of the Grand Challenge problems [1], spend a major portion of execution time in their core loops computing band linear recurrences (BLRs). Conventional compiler parallelization techniques [4] cannot generate scalable parallel code for this type of computation because they respect loop-carried dependences (LCDs) in programs, and there is a limited amount of parallelism in a BLR with respect to LCDs. For many applications, using library routines to replace the core BLR requires the separation of BLR from its dependent computation, which usually incurs significant overhead. In this paper, we present a new scalable algorithm, called the Regular Schedule, for parallel evaluation of BLRs. We describe our implementation of the Regular Schedule and discuss how to obtain maximum memory throughput in implementing the schedule on vector supercomputers. We also illustrate our approach, based on our Regular Schedule, to parallelizing programs containing BLR and other kinds of code. Significant improvements in CPU performance for a range of programs containing BLR implemented using the Regular Schedule in C over the same programs implemented using highly optimized coded-in-assembly BLAS routines [11] are demonstrated on Convex C240. Our approach can be used both at the user level in parallel programming code containing BLRs, and in compiler parallelization of such programs combined with recurrence recognition techniques for vector supercomputers.
Keyword:
band linear recurrences (BLRs)
parallel evaluation of BLRs with resource constraints
programs with BLRs
parallel programming
vector supercomputer
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Parallel and Distributed Systems 封面图
IEEE Transactions on Parallel and Distributed Systems
IF:
6
论文数:
5.2K
被引数:
1.1W

机构

暂无机构信息
引用论文

引用论文

Growth of indium phosphide from indium rich melts
err1983-11-01
err0
PREAI
errI. Grant; L. Li; D. Rumsby; R.M. Ware
err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
AUTOMATIC PROGRAM PARALLELIZATION
err1993-01-01
err166
PREAI
errBANERJEE, U; EIGENMANN, R; NICOLAU, A; PADUA, DA
err分享
err收藏
学者 查看更多内容