arrow
返回

Parallel-vector algorithms for particle simulations on shared-memory multiprocessors

delete2011-03-01
delete58
PRE
AI
D
Daisuke Nishiura *
H
Hide Sakaguchi
DOI:10.1016/j.jcp.2010.11.040delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Over the last few decades, the computational demands of massive particle-based simulations for both scientific and industrial purposes have been continuously increasing. Hence, considerable efforts are being made to develop parallel computing techniques on various platforms. In such simulations, particles freely move within a given space, and so on a distributed-memory system, load balancing, i.e., assigning an equal number of particles to each processor, is not guaranteed. However, shared-memory systems achieve better load balancing for particle models, but suffer from the intrinsic drawback of memory access competition, particularly during (1) paring of contact candidates from among neighboring particles and (2) force summation for each particle. Here, novel algorithms are proposed to overcome these two problems. For the first problem, the key is a pre-conditioning process during which particle labels are sorted by a cell label in the domain to which the particles belong. Then, a list of contact candidates is constructed by pairing the sorted particle labels. For the latter problem, a table comprising the list indexes of the contact candidate pairs is created and used to sum the contact forces acting on each particle for all contacts according to Newton's third law. With just these methods, memory access competition is avoided without additional redundant procedures. The parallel efficiency and compatibility of these two algorithms were evaluated in discrete element method (DEM) simulations on four types of shared-memory parallel computers: a multicore multiprocessor computer, scalar supercomputer, vector supercomputer, and graphics processing unit. The computational efficiency of a DEM code was found to be drastically improved with our algorithms on all but the scalar supercomputer. Thus, the developed parallel algorithms are useful on shared-memory parallel computers with sufficient memory bandwidth. (C) 2010 Elsevier Inc. All rights reserved.
Keyword:
Discrete element method
GPO computing
Vectorization
Parallelization
High performance computing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Journal of Computational Physics 封面图
Journal of Computational Physics
IF:
3.8
论文数:
1.6W
被引数:
7.4W

机构

J
japan agency for marine-earth science & technology (jamstec)
学者数:
2.9K
论文数: 3.2K
被引数: 0
引用论文

引用论文

err分享
err收藏
学者 查看更多内容