arrow
Return

A 3.77TOPS/W Convolutional Neural Network Processor With Priority-Driven Kernel Optimization

delete2019-02-01
delete18
PRE
AI
J
Jinshan Yue
Y
Yongpan Liu *
袁哲 cover
袁哲 (Zimo Yuan)
王志波 (Zhibo Wang)
Q
Qingwei Guo
J
Jinyang Li
C
Chengmo Yang
H
Huazhong Yang
DOI:10.1109/TCSII.2018.2846698delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Convolutional neural network (CNN) has become very popular in image classification tasks. With the increasing demand on intelligent classification on battery-powered devices, energy-efficient ASICs for CNN are badly needed. While previous CNN ASIC processors support operations of different kernel sizes, they sacrifice efficiency to support flexible convolution operations. In fact, convolution operations with a certain kernel size are dominating in many real-case CNNs. This brief proposes a kernel-optimized architecture for 3 x 3 kernels (KOP3), which are dominating operations in mainstream image classification CNNs. Although KOP3 aims at 3 x 3 kernel operations, it also provides programmability to support arbitrary kernel sizes. KOP3 achieves average energy efficiency of 3.77TOPS/W, which is 4.01x better than the best state-of-the-art CNN ASIC processor.
Keywords:
Convolutional neural network
CNN processor
module-parallel IS
priority-driven kernel optimization
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

I
IEEE Transactions on Circuits and Systems and Express Briefs
IF:
4.9
Papers:
8.8K
Citations:
2.5W

Organization

T
tsinghua university
Scholars:
11.8W
Papers: 10.0W
Citations: 137
U
University of Delaware
Scholars:
1.3W
Papers: 1.3W
Citations: 2.0W