arrow
Return

Fine grain algorithm parallelization on a hybrid control-flow and dataflow processor

delete2025-02-22
delete0
delete
OA
AI
N
Nenad Korolija *
DOI:10.1186/s40537-024-01021-5delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The execution time of a high-performance computing algorithm is influenced by various factors, including the algorithm's scalability, the selected hardware for processing elements, and the communication speed between these elements. This study utilizes a hybrid architecture that integrates both control-flow and dataflow hardware. Specifically, the control-flow hardware encompasses both multicore and manycore architectures. Guidance for dataflow programmers is provided to enable them to anticipate the level of acceleration achievable with a hybrid control-flow and dataflow architecture before developing dataflow hardware algorithms. Furthermore, the methodology developed is introduced, offering a structured approach for programmers to decompose algorithms and optimize each segment by leveraging the most suitable architectural type. The prerequisite for a programmer is not to know hardware description languages, but he must be well-versed in estimating the complexity of an algorithm. This study represents the culmination of over a decade of expertise in hybrid control-flow and dataflow architectures. It provides a detailed methodology for decomposing a control-flow algorithm into segments optimized for dataflow architectures and those better suited to control-flow architectures. The Lattice-Boltzmann method is employed as a representative example, implemented on both control-flow and dataflow hardware. The estimated total acceleration factor of the decomposed Lattice-Boltzmann method on the hybrid architecture, relative to execution times using control-flow and dataflow hardware, is approximately two for a given matrix dimension. The findings underscore the advantages of employing a hybrid architecture, demonstrating significant acceleration potential even for algorithms traditionally optimized for dataflow architectures. The primary benefit of the hybrid architecture lies in its capacity to accelerate algorithms where only specific portions are suitable for dataflow hardware.
Keywords:
Control-flow architectures
Dataflow architectures
Hybrid architectures
Algorithm parallelization

Journal

Journal of Big Data cover
Journal of Big Data
IF:
6.4
Papers:
1.5K
Citations:
1.1W

Organization

U
Univ Belgrade
Scholars:
1.4K
Papers: 529
Citations: 149
Cited Papers

Cited Papers

errShare
errSave
Dataflow architectures and multithreading
err1994-08-01
err0
errOAAI
errB. Lee; A.R. Hurson
errShare
errSave
An efficient dataflow accelerator for scientific applications
err2020-11-01
err9
PREAI
errYe, Xiaochun; Tan, Xu; Wu, Meng; Feng, Yujing; Wang, Da; Zhang, Hao; Pei, Songwen; Fan, Dongrui
errShare
errSave
errShare
errSave
The DataFlow Paradigm
err2015-01-01
err0
PREAI
errVeljko Milutinović; Jakob Salom; Nemanja Trifunovic; Roberto Giorgi
errShare
errSave
A Systematic Approach to Generation of New Ideas for PhD Research in Computing
err2017-01-01
err0
PREAI
errV. Blagojević; D. Bojić; M. Bojović; M. Cvetanović; J. Đorđević; Đ. Đurđević; B. Furlan; S. Gajin; Z. Jovanović; D. Milićev; V. Milutinović; B. Nikolić; J. Protić; M. Punt; Z. Radivojević; Ž. Stanisavljević; S. Stojanović; I. Tartalja; M. Tomašević; P. Vuletić
errShare
errSave
researcher View more