arrow
Return

A Tabu search based clustering algorithm and its parallel implementation on Spark

delete2018-02-01
delete31
delete
OA
AI
C
César Rego
F
Fred Glover
DOI:10.1016/j.asoc.2017.11.038delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
The well-known K-means clustering algorithm has been employed widely in different application domains ranging from data analytics to logistics applications. However, the K-means algorithm can be affected by factors such as the initial choice of centroids and can readily become trapped in a local optimum. In this paper, we propose an improved K-means clustering algorithm that is augmented by a Tabu Search strategy, and which is better adapted to meet the needs of big data applications. Our design focuses on enhancements to take advantage of parallel processing based on the Spark framework. Computational experiments demonstrate the superiority of our parallel Tabu Search based clustering algorithm over a widely used version of the K-means approach embodied in the parallel Spark MLlib system, comparing the algorithms in terms of scalability, accuracy, and effectiveness. (C) 2017 Elsevier B.V. All rights reserved.
Keywords:
Clustering
K-means
Tabu search
Parallel computing
Spark
Big data
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Applied Soft Computing cover
Applied Soft Computing
IF:
6.6
Papers:
1.4W
Citations:
4.8W

Organization

University of Colorado System cover
University of Colorado System
Scholars:
6.3W
Papers: 5.5W
Citations: 1.8K
T
tongji university
Scholars:
7.7W
Papers: 5.9W
Citations: 98
U
University of Mississippi
Scholars:
9.5K
Papers: 7.9K
Citations: 5.8K
researcher View more organizations