arrow
返回

Towards a Novel Framework for Automatic Big Data Detection

delete2020-01-01
delete4
delete
OA
AI
H
Hameeza Ahmed *
M
Muhammad Ali Ismail
DOI:10.1109/ACCESS.2020.3030562delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Big data is a relative concept. It is the combination of data, application, and platform properties. Recently, big data specific technologies have emerged, including software frameworks, databases, hardware accelerators, storage technologies, etc. However, the automatic selection of these solutions for big data computations remains a non-trivial task. Presently, the big data tools are selected by analyzing the problem manually, or by using several performance prediction techniques. The manual identification is based on the data properties only, whereas the performance predictors only estimate basic execution metrics without linking them with big data (3Vs) thresholds. Hence, both ways of identification are mostly incorrect, which can lead to inefficient use of 3Vs optimizations, resulting into global inefficiency, reduced system performance, increasing power consumption, requiring greater effort on the part of the programming team, and misallocation of the hardware resources required for the task. In this regard, a novel framework has been proposed for automatic detection of 3Vs (Volume, Velocity, Variety) of big data, using machine learning. The detection is done through static code features, data, and platform properties, leading to relevant tool selection, and code generation, with minimal overheads, lesser programmer interventions, higher usability, and portability. Instead of handling each application with big data specialized solutions, or manually identifying the 3Vs, the framework can automatically detect and link the 3Vs to the relevant optimizations. Several standard applications have been tested using the proposed framework. In the case of volume, the average detection accuracy is up to 97.8% for seen and 95.9% for unseen applications. In the case of velocity, the average detection accuracy is up to 97.3% for seen and 92.6%; for unseen applications. There is no margin of error in variety detection, as it has straightforward computations without any predictions. Furthermore, an airline recommendation system case study strengthens the effectiveness of the proposed approach.
Keyword:
Big Data
Feature extraction
Tools
Hardware
Software
Measurement
Optimization
Big data (3Vs)
detection
LLVM
machine learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

N
ned university of engineering and technology
学者数:
834
论文数: 637
被引数: 0
引用论文

引用论文

Organozinnverbindungen
err1984-08-01
err0
PREAI
errJianxi Fu; Wilhelm P. Neumann
err分享
err收藏
Impact of liver fibrosis and clinical characteristics on dose-adjusted serum methadone concentrations
err2022-03-31
err0
errOAAI
errFatemeh Chalabianloo; Gudrun Høiseth; Jørn Henrik Vold; Kjell Arne Johansson; Marianne K. Kringen; Olav Dalgard; Christian Ohldieck; Karl Trygve Druckrey-Fiskaaen; Christer Aas; Else-Marie Løberg; Jørgen G. Bramness; Lars Thore Fadnes
err分享
err收藏
err分享
err收藏
Development and evaluation of a novel seawater-based viscoelastic fracturing fluid system
err2019-12-01
err0
PREAI
errXin Sun; Zhibin Gao; Mingwei Zhao; Mingwei Gao; Mingyong Du; Caili Dai
err分享
err收藏
Camptothecin exhibits topoisomerase1-independent KMT1A suppression and myogenic differentiation in alveolar rhabdomyosarcoma cells
err2018-05-25
err0
errOAAI
errDavid W. Wolff; Min-Hyung Lee; Mathivanan Jothi; Munmun Mal; Fengzhi Li; Asoke K. Mal
err分享
err收藏
学者 查看更多内容