arrow
返回

Amdahl's law for multithreaded multicore processors

delete2014-10-01
delete15
PRE
AI
H
Hao Che
M
Minh Nguyen *
DOI:10.1016/j.jpdc.2014.06.012delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this paper, we conduct performance scaling analysis of multithreaded multicore processors (MMPs) for parallel computing. We propose a thread-level closed-queuing network model covering a fairly large design space, accounting for hardware scaling models, coarse-grain, fine-grain, and simultaneous multithreading (SMT) cores, shared resources, including cache, memory, and critical sections. We then derive a closed-form solution for this model in terms of speedup performance measure. This solution makes it possible to analyze performance scaling properties of MMPs along multiple dimensions. In particular, we show that for the parallelizable part of the workload, the speedup, in the absence of resource contention, is no longer just a linear function of parallel processing unit counts, as predicted by Amdahl's law, but also a strong function of workload characteristics, ranging from strong memory-bound to strong CPU-bound workloads. We also find that with core multithreading, super linear speedup, higher than that predicted by Amdahl's law, may be achieved for the parallelizable part of the workload, if core threads exhibit strong cache affinity and the workload is strongly memory-bound. Then, we derive a tight speedup upper bound in the presence of both memory resource contention and critical section for multicore processors with single-threaded cores. This speedup upper bound indicates that with resource contention among threads, whether it is due to shared memory or critical section, a sequential term is guaranteed to emerge from the parallelizable part of the workload, fundamentally limiting the scalability of multicore processors for parallel computing, in addition to the sequential part of the workload, as dictated by Amdahl's law. As a result, to improve speedup performance for MMPs, one should strive to enhance memory parallelism and confine critical sections as locally as possible, e.g., to the smallest possible number of threads in the same core. (C) 2014 Elsevier Inc. All rights reserved.
Keyword:
Amdahl's law
Multithreaded multicore processor
Closed queuing network
Speedup
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Journal of Parallel and Distributed Computing 封面图
Journal of Parallel and Distributed Computing
IF:
4
论文数:
3.8K
被引数:
4.8K

机构

U
university of texas system
学者数:
18.5W
论文数: 15.6W
被引数: 210
引用论文

引用论文

Effects of Temperature on the Age-Stage, Two-Sex Life Table of Bradysia odoriphaga (Diptera: Sciaridae)
err2015-01-22
err0
PREAI
errW. Li; Y. Yang; W. Xie; Q. Wu; B. Xu; S. Wang; X. Zhu; S. Wang; Y. Zhang
err分享
err收藏
Stag hunt contests and alliance formation
err2018-06-18
err0
PREAI
errJames W. Boudreau; Lucas Rentschler; Shane Sanders
err分享
err收藏
Integer and Combinatorial Optimization
err
IF0
err2014-08-22
err0
PREAI
errGeorge Nemhauser; Laurence Wolsey
err分享
err收藏
err分享
err收藏
Interannual oxygen isotope variability in Indian summer monsoon precipitation reflects changes in moisture sources
err2021-05-20
err0
errOAAI
errGayatri Kathayat; Ashish Sinha; Masahiro Tanoue; Kei Yoshimura; Hanying Li; Haiwei Zhang; Hai Cheng
err分享
err收藏
The Root Herbivore History of the Soil Affects the Productivity of a Grassland Plant Community and Determines Plant Response to New Root Herbivore Attack
err2013-02-18
err0
errOAAI
errIlja Sonnemann; Stefan Hempel; Maria Beutel; Nicola Hanauer; Stefan Reidinger; Susanne Wurst
err分享
err收藏
学者 查看更多内容