arrow
返回

Efficient AI System Design With Cross-Layer Approximate Computing

delete2020-12-01
delete32
PRE
AI
S
Swagath Venkataramani *
X
Xiao Sun
N
Naigang Wang
C
Chia‐Yu Chen
J
Jungwook Choi
M
Min‐Gu Kang
A
Ankur Agarwal
J
Jinwook Oh
S
Shubham Jain
T
Tina Babinsky
N
Nianzheng Cao
T
Thomas Fox
B
Bruce Fleischer
G
George Gristede
M
Michael Guillorn
H
Howard Haynie
H
Hiroshi Inoue
K
Kazuaki Ishizaki
M
Michael J. Klaiber
S
Shih-Hsien Lo
G
Gary Maier
S
Silvia Melitta Mueller
M
M. Scheuermann
E
Eri Ogawa
M
Marcel Schaal
M
Maurício Serrano
J
J. A. Silberman
C
Christos Vezyrtzis
W
Wei Wang
F
Fanchieh Yee
J
Jintao Zhang
M
Matthew M. Ziegler
C
Ching Zhou
M
Moriyoshi Ohara
P
Pong-Fei Lu
B
Brian Curran
S
Sunil Shukla
V
Vijayalakshmi Srinivasan
L
Leland Chang
K
Kailash Gopalakrishnan
DOI:10.1109/JPROC.2020.3029453delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Advances in deep neural networks (DNNs) and the availability of massive real-world data have enabled superhuman levels of accuracy on many AI tasks and ushered the explosive growth of AI workloads across the spectrum of computing devices. However, their superior accuracy comes at a high computational cost, which necessitates approaches beyond traditional computing paradigms to improve their operational efficiency. Leveraging the application-level insight of error resilience, we demonstrate how approximate computing (AxC) can significantly boost the efficiency of AI platforms and play a pivotal role in the broader adoption of AI-based applications and services. To this end, we present RaPiD, a multi-tera operations per second (TOPS) AI hardware accelerator core (fabricated at 14-nm technology) that we built from the ground-up using AxC techniques across the stack including algorithms, architecture, programmability, and hardware. We highlight the workload-guided systematic explorations of AxC techniques for AI, including custom number representations, quantization/pruning methodologies, mixed-precision architecture design, instruction sets, and compiler technologies with quality programmability, employed in the RaPiD accelerator.
Keyword:
Neural networks
Approximate computing
Approximation methods
Artificial intelligence
Computer architecture
Computational efficiency
Approximate computing (AxC)
artificial intelligence (AI)
deep neural networks (DNNs)
hardware acceleration
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Proceedings of the IEEE 封面图
Proceedings of the IEEE
IF:
25.9
论文数:
9.9K
被引数:
4.5W

机构

I
international business machines (ibm)
学者数:
5.7K
论文数: 4.5K
被引数: 4