arrow
Return

Architectural support for efficient message passing on shared memory multi-cores

delete2016-09-01
delete3
PRE
AI
R
Rubén Titos-Gil *
O
Oscar Palomar
O
Osman Ünsal
A
Adrián Cristal
DOI:10.1016/j.jpdc.2016.02.005delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Thanks to programming approaches like actor-based models, message passing is regaining popularity outside large-scale scientific computing for building scalable distributed applications in multi-core processors. Unfortunately, the mismatch between message passing models and today's shared-memory hardware provided by commercial vendors results in suboptimal performance and a waste of energy. This paper presents a set of architectural extensions to reduce the overheads incurred by message passing workloads running on shared memory multi-core architectures. It describes the instruction set extensions and the hardware implementation. In order to facilitate programmability, the proposed extensions are used by a message passing library, allowing programs to take advantage of them transparently. As a proof-of-concept, we use modified MPI libraries and unmodified MPI programs to evaluate the proposal. Experimental results show that a best-effort design can eliminate over 60% of cache accesses caused by message data transmission and reduce the cycles spent in such task by 75%, while the addition of a simple coprocessor can completely off-load data movement from the CPU to avoid up to 92% of cache accesses, and a reduction of 12% of network traffic on average. The design achieves an improvement of 11%-12% in the energy-delay product of on-chip caches. (C) 2016 Elsevier Inc. All rights reserved.
Keywords:
Message passing
Shared memory
Multicore
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Journal of Parallel and Distributed Computing cover
Journal of Parallel and Distributed Computing
IF:
4
Papers:
3.8K
Citations:
4.8K

Organization

B
barcelona supercomputer center (bsc-cns)
Scholars:
1.2K
Papers: 825
Citations: 5
U
universitat politecnica de catalunya
Scholars:
1.9W
Papers: 1.6W
Citations: 17