arrow
返回

Towards an Efficient CNN Inference Architecture Enabling In-Sensor Processing

delete2021-03-10
delete7
delete
OA
AI
M
Md Jubaer Hossain Pantho
P
Pankaj Bhowmik
C
Christophe Bobda *
DOI:10.3390/s21061955delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
The astounding development of optical sensing imaging technology, coupled with the impressive improvements in machine learning algorithms, has increased our ability to understand and extract information from scenic events. In most cases, Convolution neural networks (CNNs) are largely adopted to infer knowledge due to their surprising success in automation, surveillance, and many other application domains. However, the convolution operations' overwhelming computation demand has somewhat limited their use in remote sensing edge devices. In these platforms, real-time processing remains a challenging task due to the tight constraints on resources and power. Here, the transfer and processing of non-relevant image pixels act as a bottleneck on the entire system. It is possible to overcome this bottleneck by exploiting the high bandwidth available at the sensor interface by designing a CNN inference architecture near the sensor. This paper presents an attention-based pixel processing architecture to facilitate the CNN inference near the image sensor. We propose an efficient computation method to reduce the dynamic power by decreasing the overall computation of the convolution operations. The proposed method reduces redundancies by using a hierarchical optimization approach. The approach minimizes power consumption for convolution operations by exploiting the Spatio-temporal redundancies found in the incoming feature maps and performs computations only on selected regions based on their relevance score. The proposed design addresses problems related to the mapping of computations onto an array of processing elements (PEs) and introduces a suitable network structure for communication. The PEs are highly optimized to provide low latency and power for CNN applications. While designing the model, we exploit the concepts of biological vision systems to reduce computation and energy. We prototype the model in a Virtex UltraScale+ FPGA and implement it in Application Specific Integrated Circuit (ASIC) using the TSMC 90nm technology library. The results suggest that the proposed architecture significantly reduces dynamic power consumption and achieves high-speed up surpassing existing embedded processors' computational capabilities.
Keyword:
CNN
embedded vision
FPGA
pixel-parallel processing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Sensors 封面图
Sensors
IF:
3.5
论文数:
7.2W
被引数:
20.9W

机构

State University System of Florida 封面图
State University System of Florida
学者数:
12.8W
论文数: 10.9W
被引数: 130
引用论文

引用论文

The DEAD-Box RNA Helicase DDX3 Interacts with NF-κB Subunit p65 and Suppresses p65-Mediated Transcription
err2016-10-13
err0
errOAAI
errNian Xiang; Miao He; Musarat Ishaq; Yu Gao; Feifei Song; Liang Guo; Li Ma; Guihong Sun; Dan Liu; Deyin Guo; Yu Chen
err分享
err收藏
A Self-Powered Triboelectric Nanosensor for PH Detection
err2016-01-01
err0
errOAAI
errYing Wu; Yuanjie Su; Junjie Bai; Guang Zhu; Xiaoyun Zhang; Zhanolin Li; Yi Xiang; Jingliang Shi
err分享
err收藏
Deep learning for visual understanding: A review视觉理解的深度学习: 综述
err2016-04-01
err1.6K
PREAI
errGuo, Yanming; Liu, Yu; Oerlemans, Ard; Lao, Songyang; Wu, Song; Lew, Michael S.
err分享
err收藏
Crucial experiment to resolve Abraham–Minkowski controversy
err2011-11-01
err0
errOAAI
errZhong-Yue Wang; Pin-Yu Wang; Yan-Rong Xu
err分享
err收藏
学者 查看更多内容