返回
Parallelization of torsion finite element code using compressed stiffness matrix algorithm
DOI:10.1007/s00366-020-00952-w.png)
摘要
En 中文
In the current study, the problems of elastic and elastoplastic torsion were formulated by finite element method. The finite element code was parallelized on both shared and distributed memory architectures. An assembling method with high parallelism ability and consuming minimum memory was proposed to obtain compressed global stiffness matrix directly. Parallel programming principles were expressed in two shared memory and distributed memory approaches; moreover, parallel well-known mathematical libraries were briefly expressed. In this paper, the main focus was on a lucid explanation of parallelization mechanisms in detail on two memory architectures such as some settings of Linux operating system for large-scale problems. To verify the ability of the proposed method and its parallel performance, several benchmark examples were represented with different mesh sizes and were compared with their respective analytical solutions. Considering the obtained results, the proposed sparse assembling algorithm decreased required memory significantly (about 10(3.5) to 10(5.5) times) and the obtained speedup was about 3.4 for the elastoplastic torsion problem in a simple multicore computer.
Keyword:
Parallelization
OpenMP
SuperLU
Assembling
Finite element method
Elastoplastic torsion
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
4.9
论文数:
2.6K
被引数:
9.3K
机构
引用论文
Turbocharger motor-generator for improvement of transient performance in an internal combustion engine用于改善内燃机瞬态性能的涡轮增压器电动发电机

