返回
Optimal broadcast for fully connected processor-node networks
DOI:10.1016/j.jpdc.2007.12.001.png)
摘要
En 中文
We develop and implement an optimal broadcast algorithm for fully connected processor networks under a bidirectional communication model in which each processor can simultaneously send a message to one processor and receive a message from another, possibly different processor. For any number of processors p the algorithm requires N - 1 + [log p] communication rounds to broadcast N blocks of data from a root processor to the remaining processors, meeting the lower bound in the model. For data of size in, assuming that sending and receiving data of size m' takes time alpha + beta m', the best running time that can be achieved by the division of m into equal-sized blocks is (root([log p] - 1)alpha + root beta m)(2). The algorithm uses a regular, circulant graph communication pattern, and degenerates into a binomial tree broadcast when the number of blocks to be broadcast is one. The algorithm is furthermore well suited to fully connected clusters of SMP (Symmetric Multi-Processor) nodes. The algorithm is implemented as part of an MPI (Message Passing Interface) library. We demonstrate significant practical bandwidth improvements of up to a factor 1.5 over several other, commonly used broadcast algorithms on both a small SMP cluster and a 72 node NEC SX vector supercomputer. (c) 2008 Elsevier Inc. All rights reserved.
Keyword:
broadcast
fully connected communication network
bidirectional communication model
SMP cluster
MPI (Message Passing Interface)
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
4
论文数:
3.8K
被引数:
4.8K
机构
引用论文
Spitz/Reed nevi: a review of clinical-dermatoscopic and histological correlationSpitz/Reed痣:临床-皮肤镜和病理学相关性综述

