返回
Optimizing dataflow applications on heterogeneous environments
DOI:10.1007/s10586-010-0151-6.png)
摘要
En 中文
The increases in multi-core processor parallelism and in the flexibility of many-core accelerator processors, such as GPUs, have turned traditional SMP systems into hierarchical, heterogeneous computing environments. Fully exploiting these improvements in parallel system design remains an open problem. Moreover, most of the current tools for the development of parallel applications for hierarchical systems concentrate on the use of only a single processor type (e.g., accelerators) and do not coordinate several heterogeneous processors. Here, we show that making use of all of the heterogeneous computing resources can significantly improve application performance. Our approach, which consists of optimizing applications at run-time by efficiently coordinating application task execution on all available processing units is evaluated in the context of replicated dataflow applications. The proposed techniques were developed and implemented in an integrated run-time system targeting both intra- and inter-node parallelism. The experimental results with a real-world complex biomedical application show that our approach nearly doubles the performance of the GPU-only implementation on a distributed heterogeneous accelerator cluster.
Keyword:
GPGPU
Run-time optimizations
Filter-stream
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
C
IF:
4.1
论文数:
5.0K
被引数:
7.5K
机构
引用论文
What a hawkmoth remembers after hibernation depends on innate preferences and conditioning situation
Computer-aided prognosis of neuroblastoma on whole-slide images: Classification of stromal development全载玻片图像上神经母细胞瘤的计算机辅助预后: 基质发育的分类
PATTERN RECOGNITION
IF7.6

