arrow
返回

Simulating cortical networks on heterogeneous multi-GPU systems

delete2013-07-01
delete7
PRE
AI
A
Andrew Nere *
S
Sean Franey
A
Atif Hashmi
M
Mikko H. Lipasti
DOI:10.1016/j.jpdc.2012.02.006delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recent advances in neuroscientific understanding have highlighted the highly parallel computation power of the mammalian neocortex. In this paper we describe a GPGPU-accelerated implementation of an intelligent learning model inspired by the structural and functional properties of the neocortex. Furthermore, we consider two inefficiencies inherent to our initial implementation and propose software optimizations to mitigate such problems. Analysis of our application's behavior and performance provides important insights into the GPGPU architecture, including the number of cores, the memory system, atomic operations, and the global thread scheduler. Additionally, we create a runtime profiling tool for the cortical network that proportionally distributes work across the host CPU as well as multiple GPGPUs available to the system. Using the profiling tool with these optimizations on Nvidia's CUDA framework, we achieve up to 60 x speedup over a single-threaded CPU implementation of the model. (c) 2012 Elsevier Inc. All rights reserved.
Keyword:
Cortical learning algorithms
CUDA
GPGPU
Profiling systems

期刊

Journal of Parallel and Distributed Computing 封面图
Journal of Parallel and Distributed Computing
IF:
4
论文数:
3.8K
被引数:
4.8K

机构

University of Wisconsin System 封面图
University of Wisconsin System
学者数:
6.7W
论文数: 5.8W
被引数: 382
引用论文

引用论文

Always returning: feedback and sensory processing in visual cortex and thalamus
err2006-06-01
err218
errOAAI
errSillito, Adam M.; Cudeiro, Javier; Jones, Helen E.
err分享
err收藏
The columnar organization of the neocortex
errBRAIN
IF11.7
err1997-04-01
err1.7K
errOAAI
errMountcastle, VB
err分享
err收藏
err分享
err收藏
学者 查看更多内容