arrow
Return

DK-means: a deterministic K-means clustering algorithm for gene expression analysis

delete2017-12-28
delete33
PRE
AI
R
R. Jothi *
S
Sraban Kumar Mohanty
A
Aparajita Ojha
DOI:10.1007/s10044-017-0673-0delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Clustering has been widely applied in interpreting the underlying patterns in microarray gene expression profiles, and many clustering algorithms have been devised for the same. K-means is one of the popular algorithms for gene data clustering due to its simplicity and computational efficiency. But, K-means algorithm is highly sensitive to the choice of initial cluster centers. Thus, the algorithm easily gets trapped with local optimum if the initial centers are chosen randomly. This paper proposes a deterministic initialization algorithm for K-means (DK-means) by exploring a set of probable centers through a constrained bi-partitioning approach. The proposed algorithm is compared with classical K-means with random initialization and improved K-means variants such as K-means++ and MinMax algorithms. It is also compared with three deterministic initialization methods. Experimental analysis on gene expression datasets demonstrates that DK-means achieves improved results in terms of faster and stable convergence, and better cluster quality as compared to other algorithms.
Keywords:
K-means clustering algorithm
Initial cluster centers
Gene expression clustering
Microarray data analysis
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Pattern Analysis and Applications cover
Pattern Analysis and Applications
IF:
2
Papers:
1.9K
Citations:
1.9K

Organization

P
pandit deendayal energy university
Scholars:
1.6K
Papers: 1.3K
Citations: 26