arrow
Return

Efficient Distributed Graph Neural Network Training With Source Chunking and Moving Aggregation

delete2025-09-24
delete0
PRE
AI
W
Wenjie Huang
T
Tongya Zheng
汪睿 (Rui Wang)
T
Tongtian Zhu
B
Bingde Hu
S
Shuibing He
宋明黎 (Mingli Song)
王新宇 (Xinyu Wang)
S
Sai Wu
C
Chun Chen
DOI:10.1109/TKDE.2025.3613787delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Graph neural networks (GNNs) are effective models for analyzing graph-structured data, but encounter challenges when training on large distributed graphs. Existing GNN training frameworks use sampling parallelism and historical embedding methods to support distributed training and enhance efficiency. However, these methods suffer from issues like stale historical embeddings, imbalanced communication messages, and redundant storage and computation costs. In this paper, we present Emma, a distributed GNN training framework that incorporates source node centric chunking for frequent updates of embeddings and balanced communication, as well as a moving message aggregation technique to boost training efficiency and reduce storage costs. Experimental results show that Emma significantly enhances training efficiency by reducing computation and communication overhead, leading to a notable speedup while maintaining convergence accuracy compared to state-of-the-art distributed GNN training methods.
Keywords:
Graph neural networks (GNNs)
distributed GNN training
graph message aggregation
historical embedding

Journal

IEEE Transactions on Knowledge and Data Engineering cover
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
Papers:
6.7K
Citations:
3.2W

Organization

H
Hangzhou City University
Scholars:
2.2K
Papers: 2.0K
Citations: 1.0K
Z
zhejiang university
Scholars:
17.4W
Papers: 12.0W
Citations: 152