arrow
返回

A reference architecture for serverless big data processing

delete2024-06-01
delete3
delete
OA
AI
S
Sebastian Werner *
S
Stefan Tai
DOI:10.1016/j.future.2024.01.029delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Despite significant advances in data management systems in recent decades, the processing of big data at scale remains very challenging. While cloud computing has been well -accepted as a solution to address scalability needs, cloud configuration and operation complexity persist and often present themselves as entry barriers, especially for novice data analysts. Serverless computing and Function -as -a -Service (FaaS) platforms have been suggested to reduce such entry barriers by shifting configuration and operational responsibilities from the application developer to the FaaS platform provider. Naturally, serverless data processing (SDP)'', that is, using FaaS for (big) data processing, has received increasing interest in recent years. However, FaaS platforms were never intended to support large data processing tasks primarily. SDP, therefore, manifests itself through workarounds and adaptations on the application level, addressing various quirks and limitations of the FaaS platforms in use for data processing needs. This, in turn, creates tensions between the platforms and the applications using them, again encouraging the constant (re -)design of both. Consequently, we present lessons learned from a series of application and platform re -designs that address these tensions, leading to the development of an SDP reference architecture and a platform instantiation and implementation thereof called CREW. Mitigating the tensions through the process of application platform codesign proves to reduce both entry barriers and costs significantly. In some experiments, CREW outperforms traditional, non -SDP big data processing frameworks by factors.
Keyword:
Serverless data processing
Application platform co -design
Serverless reference architecture
Function as a Service
Software engineering
Cloud computing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Future Generation Computer Systems-The International Journal of eScience
IF:
6.1
论文数:
6.9K
被引数:
2.3W

机构

T
Technical University of Berlin
学者数:
1.3W
论文数: 1.1W
被引数: 18
引用论文

引用论文

err分享
err收藏
Classification of Granular Material in an Impact with a Separation Surface
err2015-04-22
err0
PREAI
errS. A. Lyaptsev; V. Ya. Potapov; S. Ya. Davydov; V. V. Potapov; L. A. Semerikov; E. A. Vasil’ev
err分享
err收藏
The Rise of Serverless Computing
err2019-11-21
err181
PREAI
errCastro, Paul; Ishakian, Vatche; Muthusamy, Vinod; Slominski, Aleksander
err分享
err收藏
Serverless computing for container-based architectures
err2018-06-01
err75
errOAAI
errPerez, Alfonso; Molto, German; Caballer, Miguel; Calatrava, Amanda
err分享
err收藏
A mixed-method empirical study of Function-as-a-Service software development in industrial practice
err2019-03-01
err72
PREAI
errLeitner, Philipp; Wittern, Erik; Spillner, Josef; Hummer, Waldemar
err分享
err收藏
Recent Patents on Roll Crushing Mills for Selective Crushing of Coal and Gangue
err2020-02-12
err0
PREAI
errDaolong Yang; Yanxiang Wang; Bangsheng Xing; Yanting Yu; Yuntao Wang; Youtao Xia
err分享
err收藏
学者 查看更多内容