arrow
返回

Taskgraph: A Low Contention OpenMP Tasking Framework

delete2023-08-01
delete3
delete
OA
AI
C
Chenle Yu *
S
Sara Royuela
E
Eduardo Quiñones
DOI:10.1109/TPDS.2023.3284219delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
OpenMP is the de-facto standard for shared memory systems in High-Performance Computing (HPC). It includes a tasking model that offers a high-level of abstraction to effectively exploit structured (loop-based) and highly dynamic unstructured (task-based) parallelism in an easy and flexible way. Unfortunately, the run-time overheads introduced to manage tasks are (very) high in most common OpenMP frameworks (e.g., GCC, LLVM), which defeats the potential benefits of the tasking model, and makes it suitable for coarse-grained tasks only. This paper presents taskgraph, a framework that uses a task dependency graph (TDG) to represent a region of code implemented with OpenMP tasks in order to reduce the run-time overheads associated with the management of tasks, i.e., contention and parallel orchestration, including task creation and synchronization. The TDG avoids the overheads related to the resolution of task dependencies and greatly reduces those deriving from accesses to shared resources. Moreover, the taskgraph framework introduces in OpenMP the record-and-replay execution model that accelerates the taskgraph region from its second execution. Overall, the multiple optimizations presented in this paper allow exploiting fine-grained OpenMP tasks to cope with the trend in current applications pointing to leverage massive on-node parallelism, fine-grained and dynamic scheduling paradigms. The framework is implemented on LLVM 15.0. Results show that the taskgraph implementation outperforms the vanilla OpenMP system in terms of performance and scalability, for all structured and unstructured parallelism, and considering coarse and fine grained tasks. Furthermore, the proposed framework makes the tasking model a competitive alternative to the OpenMP thread model in most cases.
Keyword:
OpenMP tasking
run-time overhead
fine-grained parallelism

期刊

IEEE Transactions on Parallel and Distributed Systems 封面图
IEEE Transactions on Parallel and Distributed Systems
IF:
6
论文数:
5.2K
被引数:
1.1W

机构

B
barcelona supercomputer center (bsc-cns)
学者数:
1.2K
论文数: 825
被引数: 5
U
universitat politecnica de catalunya
学者数:
1.9W
论文数: 1.6W
被引数: 17
引用论文

引用论文

Time-Dependent Statistical and Correlation Properties of Neural Signals during Handwriting
err2012-09-11
err0
errOAAI
errValery I. Rupasov; Mikhail A. Lebedev; Joseph S. Erlichman; Stephen L. Lee; James C. Leiter; Michael Linderman
err分享
err收藏
Disrupted Value-Directed Strategic Processing in Individuals with Mild Cognitive Impairment: Behavioral and Neural Correlates
err2022-05-11
err0
errOAAI
errLydia T. Nguyen; Elizabeth A. Lydon; Shraddha A. Shende; Daniel A. Llano; Raksha A. Mudar
err分享
err收藏
A 5D gyrokinetic full-f global semi-Lagrangian code for flux-driven ion turbulence simulations用于通量驱动离子湍流模拟的5D陀螺动力学全f全局半拉格朗日代码
err2016-10-01
err87
errOAAI
errGrandgirard, V.; Abiteboul, J.; Bigot, J.; Cartier-Michaud, T.; Crouseilles, N.; Dif-Pradalier, G.; Ehrlacher, Ch.; Esteve, D.; Garbet, X.; Ghendrih, Ph.; Latu, G.; Mehrenberger, M.; Norscini, C.; Passeron, Ch.; Rozar, F.; Sarazin, Y.; Sonnendruecker, E.; Strugarek, A.; Zarzoso, D.
err分享
err收藏
Xanthogranulomatous pituitary adenoma: A case report and literature review
err2018-01-10
err0
errOAAI
errGuihong Li; Chaochao Zhang; Yuxue Sun; Qingchun Mu; Haiyan Huang
err分享
err收藏
学者 查看更多内容