arrow
返回

Efficient computation of address sequences in data parallel programs using closed forms for basis vectors

delete1996-11-01
delete16
PRE
AI
A
A. Thirumalai *
J
J. Ramanujam
DOI:10.1006/jpdc.1996.0140delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Arrays are mapped to processors through a two-step process-alignment followed by distribution-in data-parallel languages such as High Performance Fortran. This process of mapping creates disjoint pieces of the array that are locally owned by each processor. An HPF compiler that generates code for array statements must compute the sequence of local memory addresses accessed by each processor and the sequence of sends and receives for a given processor to access nonlocal data. In this paper, we present an approach to the address sequence generation problem using the theory of integer lattices. The set of elements referenced can be generated by integer linear combinations of basis vectors. Unlike other work on this problem, we derive closed form expressions for the basis vectors as a function of the mapping of data. Using these basis vectors and exploiting the fact that there is a repeating pattern in the access sequence, we derive highly optimized code that generates the pattern at runtime. The code generated uses table-lookup of the pattern. Experimental results show that our approach is faster than other solutions to this problem. (C) 1997 Academic Press, Inc.
Keyword:
COMMUNICATION

期刊

Journal of Parallel and Distributed Computing 封面图
Journal of Parallel and Distributed Computing
IF:
4
论文数:
3.8K
被引数:
4.8K

机构

暂无机构信息
引用论文

引用论文

暂无论文信息