arrow
Return

Dimensionality reduction in data mining: A Copula approach

delete2016-12-01
delete50
PRE
AI
A
Ahcène Bounceur *
T
Tahar Kechadi
R
Reinhardt Euler
DOI:10.1016/j.eswa.2016.07.041delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The recent trends in collecting huge and diverse datasets have created a great challenge in data analysis. One of the characteristics of these gigantic datasets is that they often have significant amounts of redundancies. The use of very large multi-dimensional data will result in more noise, redundant data, and the possibility of unconnected data entities. To efficiently manipulate data represented in a high dimensional space and to address the impact of redundant dimensions on the final results, we propose a new technique for the dimensionality reduction using Copulas and the LU-decomposition (Forward Substitution) method. The proposed method is compared favorably with existing approaches on real-world datasets: Diabetes, Waveform, two versions of Human Activity Recognition based on Smartphone, and Thyroid Datasets taken from machine learning repository in terms of dimensionality reduction and efficiency of the method, which are performed on statistical and classification measures. (C) 2016 Elsevier Ltd. All rights reserved.
Keywords:
Data mining
Data pre-processing
Multi-dimensional sampling
Copulas
Dimensionality reduction
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Expert Systems with Applications cover
Expert Systems with Applications
IF:
7.5
Papers:
2.9W
Citations:
10.2W

Organization

U
universite de bejaia
Scholars:
1.6K
Papers: 1.1K
Citations: 0
U
universite de bretagne occidentale
Scholars:
7.2K
Papers: 5.0K
Citations: 6
U
university college dublin
Scholars:
2.6W
Papers: 2.2W
Citations: 22
researcher View more organizations