arrow
Return

HOME: A Holistic GPU Memory Management Framework for Deep Learning

delete2022-01-01
delete5
PRE
AI
S
Shuibing He
P
Ping Chen *
Z
Zheng Li
S
Siling Yang
W
Weijian Chen
L
Lidan Shou
DOI:10.1109/TC.2022.3180991delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We propose HOlistic MEmory management (HOME), a new framework for performing tensor placements in large DNN training when GPU memory space is not enough. HOME combines tensor swapping with tensor recomputation to reduce GPU memory footprint. Different from existing work that only considers partial DNN model information, HOME takes the holistic DNN model information into account in tensor placement decisions. More specifically, HOME uses a custom-designed particle swarm optimization algorithm to achieve the globally optimized placement for each tensor of the DNN model with a greatly reduced searching space. This holistic awareness of the whole model information enables HOME to obtain high performance under the given GPU memory constraint. We implement HOME in PyTorch and conduct our experiments using six popular DNN models. Experimental results show that HOME can outperform vDNN and Capuchin by up to 5.7x and 1.3x in throughput. Furthermore, HOME can improve the maximum batch size by up to 2.8x than the original PyTorch and up to 1.3x than Capuchin.
Keywords:
DNN
GPU
recomputation
swapping
tensor

Journal

IEEE Transactions on Computers cover
IEEE Transactions on Computers
IF:
3.8
Papers:
5.3K
Citations:
9.8K

Organization

Stockton University cover
Stockton University
Scholars:
281
Papers: 248
Citations: 257
Z
zhejiang university
Scholars:
17.5W
Papers: 12.0W
Citations: 152