返回
Improving classification accuracy using data augmentation on small data sets
DOI:10.1016/j.eswa.2020.113696.png)
摘要
En 中文
Data augmentation (DA) is a key element in the success of Deep Learning (DL) models, as its use can lead to better prediction accuracy values when large size data sets are used. DA was not very much used with earlier neural network models before 2012, and the reason might be related to the type of models and the size of the data sets used. We investigate in this work, applying several state-of-the-art models based on Variational Autoencoders (VAEs) and Generative Adversarial Networks (GANs), the effect of DA when using small size data sets, analyzing the results in terms of the prediction accuracy obtained according to the different characteristics of the training samples (number of instances and features, and class unbalance degree). We further introduce modifications to the standard methods used to generate the synthetic samples to alter the class balance representation, and the overall results indicate that with some computational effort a significant increase in prediction accuracy can be obtained when small data sets are considered. (C) 2020 Elsevier Ltd. All rights reserved.
Keyword:
Deep Learning
Data augmentation
GAN
VAE
Unbalanced sets
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.5
论文数:
2.9W
被引数:
10.2W
机构
引用论文
A dynamic over-sampling procedure based on sensitivity for multi-class problems
PATTERN RECOGNITION
IF7.6
Winter wheat grain yield and its components in the North China Plain: Irrigation management, cultivation, and climate华北平原冬小麦籽粒产量及其组成: 灌溉管理,栽培和气候
Arsenic removal from aqueous solutions by adsorption using novel MIL-53(Fe) as a highly efficient adsorbent使用新型MIL-53(Fe) 作为高效吸附剂通过吸附从水溶液中去除砷
RSC Advances
IF0
Noise injection for training artificial neural networks: A comparison with weight decay and early stopping
MEDICAL PHYSICS
IF3.2
Generative adversarial networks for data augmentation in machine fault diagnosis机器故障诊断中用于数据增强的生成对抗网络

