arrow
Return

Accelerating k-medoid-based algorithms through metric access methods

delete2008-03-01
delete19
PRE
AI
M
Maria Camila N. Barioni *
H
Humberto Razente
A
Agma J. M. Traina
C
Caetano Traina
DOI:10.1016/j.jss.2007.06.019delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Scalable data mining algorithms have become crucial to efficiently support KDD processes on large databases. In this paper, we address the task of scaling up k-medoid-based algorithms through the utilization of metric access methods, allowing clustering algorithms to be executed by database management systems in a fraction of the time usually required by the traditional approaches. We also present an optimization strategy that can be applied as an additional step of the proposed algorithm in order to achieve better clustering solutions. Experimental results based on several datasets, including synthetic and real ones, show that the proposed algorithm can reduce the number of distance calculations by a factor of more than three thousand times when compared to existing algorithms, while producing clusters of equivalent quality. (C) 2007 Elsevier Inc. All rights reserved.
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Journal of Systems and Software cover
Journal of Systems and Software
IF:
4.1
Papers:
5.5K
Citations:
8.4K

Organization

U
universidade de sao paulo
Scholars:
10.6W
Papers: 6.7W
Citations: 93
Cited Papers

Cited Papers

Data clustering: A review
err1999-09-01
err9.6K
errOAAI
errJain, AK; Murty, MN; Flynn, PJ
errShare
errSave
no more