arrow
Return

Revisiting workflow execution in HPC: a data-flow approach

delete2025-11-01
delete0
PRE
AI
陈滔 (Tao Chen)
X
Xiao‐Ning Wang *
G
Guanlong Li
Y
Yining Zhao
H
Haili Xiao
DOI:10.1007/s42514-025-00256-9delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Scientific workflows are essential to modern scientific computing, yet traditional execution approaches-based on control-flow paradigms and disk-based data transfers-struggle as data movement, rather than computation, emerges as the dominant performance bottleneck. These methods suffer from long latency due to centralized orchestration, sequential task triggering, and inefficient disk-mediated exchanges. We propose HPCFlow, a data-flow-oriented workflow framework designed for high-performance computing (HPC) environments. HPCFlow supports decentralized, input-driven execution. Functions are decomposed into computation and data transmission, enabling asynchronous data propagation and efficient overlap. HPCFlow incorporates context-aware data transfer strategies and alleviates small-file I/O inefficiencies through mini-batching. Additionally, HPCFlow implements an input synchronization mechanism to guarantee data completeness during parallel execution under elastic scaling conditions. Empirical results from a production HPC environment demonstrate that compared to a control-flow baseline, HPCFlow significantly reduces makespan and end-to-end latency, achieves efficient overlap, and alleviates pressure on network file systems, thereby validating its effectiveness for data-intensive scientific workflows.
Keywords:
Scientific Workflows
High-Performance Computing (HPC)
Cloud Computing
Data-Flow Execution
Data Movement Optimization

Journal

C
CCF Transactions on High Performance Computing
IF:
1.9
Papers:
38
Citations:
253

Organization

C
Chinese Academy of Sciences
Scholars:
3.9W
Papers: 1.5W
Citations: 58.4W