arrow
返回

Efficient multicore-aware parallelization strategies for iterative stencil computations

delete2011-05-01
delete31
delete
OA
AI
J
Jan Treibig
G
Gerhard Wellein
G
Georg Hager *
DOI:10.1016/j.jocs.2011.01.010delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Stencil computations consume a major part of runtime in many scientific simulation codes. As prototypes for this class of algorithms we consider the iterative Jacobi and Gauss-Seidel smoothers and aim at highly efficient parallel implementations for cache-based multicore architectures. Temporal cache blocking is a known advanced optimization technique, which can reduce the pressure on the memory bus significantly. We apply and refine this optimization for a recently presented temporal blocking strategy designed to explicitly utilize multicore characteristics. Especially for the case of Gauss-Seidel smoothers we show that simultaneous multi-threading (SMT) can yield substantial performance improvements for our optimized algorithm on some architectures. (C) 2011 Elsevier B.V. All rights reserved.
Keyword:
Stencil computations
Spatial blocking
Temporal blocking
Wavefront parallelization
Multicore
Simultaneous multi-threading
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Nature Computational Science 封面图
Nature Computational Science
IF:
18.3
论文数:
3.1K
被引数:
4.0K

机构

U
University of Erlangen Nuremberg
学者数:
3.2W
论文数: 2.6W
被引数: 29
引用论文

引用论文

Heavy metal removal from water by adsorption using a low-cost geopolymer
err2020-04-18
err0
PREAI
errLaxmipriya Panda; Sandeep K. Jena; Swagat S. Rath; Pramila K. Misra
err分享
err收藏
Optimization and Performance Modeling of Stencil Computations on Modern Microprocessors
err2009-02-05
err153
errOAAI
errDatta, Kaushik; Kamil, Shoaib; Williams, Samuel; Oliker, Leonid; Shalf, John; Yelick, Katherine
err分享
err收藏
err分享
err收藏