返回
CRONUS: A platform for parallel code generation based on computational geometry methods
DOI:10.1016/j.jss.2007.11.715.png)
摘要
En 中文
This paper describes CRONUS, a platform for parallelizing general nested loops. General nested loops contain complex loop bodies (assignments, conditionals, repetitions) and exhibit uniform loop-carried dependencies. The novelty Of CRONUS is twofold: ( 1) it determines the optimal scheduling hyperplane using the QuickHull algorithm, which is more efficient than previously used methods, and (2) it implements a simple and efficient dynamic rule (successive dynamic scheduling) for the runtime scheduling of the loop iterations along the optimal hyperplane. This scheduling policy enhances data locality and improves the makespan. CRONUS provides an efficient runtime library, specifically designed for communication minimization, that performs better than more generic systems, such as Berkeley UPC. Its performance was evaluated through extensive testing. Three representative case studies are examined: the Floyd-Steinberg dithering algorithm, the Transitive Closure algorithm, and the FSBM motion estimation algorithm. The experimental results corroborate the efficiency of the parallel code. The tests show speedup ranging from 1.18 (Out of the ideal 4) to 12.29 (Out of the ideal 16) on distributed-systems and 3.60 (out of 4) to 15.79 (out of 16) on shared-memory systems. CRONUS Outperforms UPC by 5-95% depending on the test case. (C) 2007 Elsevier Inc. All rights reserved.
Keyword:
general loops
dynamic scheduling
code generation
shared and distributed memory architectures
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
4.1
论文数:
5.5K
被引数:
8.4K
机构
引用论文
Single motor unit activity in human extraocular muscles during the vestibulo‐ocular reflex前庭眼反射过程中人眼外肌的单个运动单位活动
没有更多内容


