返回
Multiclass classification of distributed memory parallel computations
DOI:10.1016/j.patrec.2012.10.007.png)
摘要
En 中文
High Performance Computing (HPC) is a field concerned with solving large-scale problems in science and engineering. However, the computational infrastructure of HPC systems can also be misused as demonstrated by the recent commoditization of cloud computing resources on the black market As a first step towards addressing this, we introduce a machine learning approach for classifying distributed parallel computations based on communication patterns between compute nodes. We first provide relevant background on message passing and computational equivalence classes called dwarfs and describe our exploratory data analysis using self organizing maps. We then present our classification results across 29 scientific codes using Bayesian networks and compare their performance against Random Forest classifiers. These models, trained with hundreds of gigabytes of communication logs collected at Lawrence Berkeley National Laboratory, perform well without any a priori information and address several shortcomings of previous approaches. (C) 2012 Elsevier B.V. All rights reserved.
Keyword:
Multiclass classification
Bayesian networks
Random forests
Self-organizing maps
High performance computing
Communication patterns
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.3
论文数:
8.0K
被引数:
1.6W
机构
引用论文
Physiological and behavioral responses to corticotropin-releasing factor administration: is CRF a mediator of anxiety or stress responses?促肾上腺皮质激素释放因子的生理和行为反应: CRF是焦虑或应激反应的中介物吗?
Efficient iterative schemes for ab initio total-energy calculations using a plane-wave basis set使用平面波基础集从头计算总能量的有效迭代方案
PHYSICAL REVIEW B
IF3.7

