arrow
Return

Optimizing for In-Memory Deep Learning With Emerging Memory Technology

delete2024-11-01
delete1
delete
OA
AI
Z
Zhehui Wang
T
Tao Luo *
R
Rick Siow Mong Goh
W
Wei Zhang
W
Weng‐Fai Wong
DOI:10.1109/TNNLS.2023.3285488delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In-memory deep learning executes neural network models where they are stored, thus avoiding long-distance communication between memory and computation units, resulting in considerable savings in energy and time. In-memory deep learning has already demonstrated orders of magnitude higher performance density and energy efficiency. The use of emerging memory technology (EMT) promises to increase density, energy, and performance even further. However, EMT is intrinsically unstable, resulting in random data read fluctuations. This can translate to nonnegligible accuracy loss, potentially nullifying the gains. In this article, we propose three optimization techniques that can mathematically overcome the instability problem of EMT. They can improve the accuracy of the in-memory deep learning model while maximizing its energy efficiency. Experiments show that our solution can fully recover most models' state-of-the-art (SOTA) accuracy and achieves at least an order of magnitude higher energy efficiency than the SOTA.
Keywords:
Fluctuations
Deep learning
Computational modeling
Optimization
Energy efficiency
In-memory computing
Phase change random access memory
Deep learning
emerging memory technology (EMT)
in-memory computing
optimization

Journal

IEEE Transactions on Neural Networks and Learning Systems cover
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
Papers:
7.5K
Citations:
7.2W

Organization

A
a*star - institute of high performance computing (ihpc)
Scholars:
1.5K
Papers: 1.3K
Citations: 3
A
agency for science technology & research (a*star)
Scholars:
2.2W
Papers: 1.9W
Citations: 57