返回
Adaptive priority-based data placement and multi-task scheduling in geo-distributed cloud systems
DOI:10.1016/j.knosys.2021.107050.png)
摘要
En 中文
With the rapid development and the widespread use of cloud computing in various applications, the number of users distributed in different regions has grown exponentially. Therefore, the Geo-distributed cloud systems have become a research hotspot and big data processing technology has also emerged. Nowadays, the most widely used big data processing framework is Spark. However, massive amounts of data are generated every moment, and the processing procedure becomes more and more complex, the execution efficiency of Spark has been greatly affected. In the Spark frame of geo-distributed cloud systems, aiming at the data placement problem, the data placement strategy based on RDD dynamic weight is introduced. The target node is selected with a strong computation capacity to place the data. Aiming at the problems of multi-task scheduling, a task will be scheduled to a node whose computation capacity can satisfy the requirement of this task. And then considering job classification and computing node performance, the optimized task scheduling strategy is in traduced. Experiments show that our algorithms can effectively adjust the weight of node data placement according to the actual task execution information, and shorten the task execution time. (C) 2021 Elsevier B.V. All rights reserved.
Keyword:
Distributed cloud
Data stream
Spark frame
Multi-task scheduling
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.2W
被引数:
4.5W
机构
引用论文
Identification of coenzyme A-related tolmetin metabolites in rats: relationship with reactive drug metabolites大鼠中与辅酶a相关的托美泰代谢产物的鉴定: 与反应性药物代谢产物的关系
Xenobiotica
IF0
Topology-Aware Resource-Efficient Placement for High Availability Clusters Over Geo-Distributed Cloud Infrastructure
IEEE ACCESS
IF3.6

