返回
Simulating cortical networks on heterogeneous multi-GPU systems
DOI:10.1016/j.jpdc.2012.02.006.png)
摘要
En 中文
Recent advances in neuroscientific understanding have highlighted the highly parallel computation power of the mammalian neocortex. In this paper we describe a GPGPU-accelerated implementation of an intelligent learning model inspired by the structural and functional properties of the neocortex. Furthermore, we consider two inefficiencies inherent to our initial implementation and propose software optimizations to mitigate such problems. Analysis of our application's behavior and performance provides important insights into the GPGPU architecture, including the number of cores, the memory system, atomic operations, and the global thread scheduler. Additionally, we create a runtime profiling tool for the cortical network that proportionally distributes work across the host CPU as well as multiple GPGPUs available to the system. Using the profiling tool with these optimizations on Nvidia's CUDA framework, we achieve up to 60 x speedup over a single-threaded CPU implementation of the model. (c) 2012 Elsevier Inc. All rights reserved.
Keyword:
Cortical learning algorithms
CUDA
GPGPU
Profiling systems
期刊
IF:
4
论文数:
3.8K
被引数:
4.8K
机构
引用论文
A configurable simulation environment for the efficient simulation of large-scale spiking neural networks on graphics processors可配置的仿真环境,用于在图形处理器上有效仿真大规模尖峰神经网络
NEURAL NETWORKS
IF6.3
Always returning: feedback and sensory processing in visual cortex and thalamus
TRENDS IN NEUROSCIENCES
IF15.1


