arrow
返回

A data placement strategy in scientific cloud workflows

delete2010-10-01
delete241
PRE
AI
D
Dong Yuan *
Y
Yun Yang
刘笑 封面图
刘笑 (Xiao Liu)
J
Jinjun Chen
DOI:10.1016/j.future.2010.02.004delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In scientific cloud workflows, large amounts of application data need to be stored in distributed data centres. To effectively store these data, a data manager must intelligently select data centres in which these data will reside. This is, however, not the case for data which must have a fixed location. When one task needs several datasets located in different data centres, the movement of large volumes of data becomes a challenge. In this paper, we propose a matrix based k-means clustering strategy for data placement in scientific cloud workflows. The strategy contains two algorithms that group the existing datasets in k data centres during the workflow build-time stage, and dynamically clusters newly generated datasets to the most appropriate data centres - based on dependencies - during the runtime stage. Simulations show that our algorithm can effectively reduce data movement during the workflow's execution. (c) 2010 Elsevier B.V. All rights reserved.
Keyword:
Data management
Scientific workflow
Cloud computing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Future Generation Computer Systems-The International Journal of eScience
IF:
6.1
论文数:
6.9K
被引数:
2.3W

机构

S
Swinburne University of Technology
学者数:
9.3K
论文数: 1.2W
被引数: 2.0W
引用论文

引用论文

err分享
err收藏
学者 查看更多内容