arrow
Return

Sequence clustering algorithm based on weighted vector identification

delete2015-06-03
delete2
PRE
AI
吴
吴迪 (Di Wu) *
任
任家东 (Jiadong Ren)
DOI:10.1007/s13042-015-0381-2delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Sequence clustering has become an important topic that experts in data mining are currently investigating. However, clustering quality is typically significantly affected by both the selection of initial centers and the mean sequences. In this study, the sequence clustering algorithm based on weighted vector identification (SCAWVI) algorithm is developed based on sequence element composite similarity and the weight of a sequence in its corresponding cluster. Based on the weighted sequence element, all sequences in the sequence database are preprocessed into M-dimensional weighted vector identifications. Then, using Huffman-based initial clustering centers optimization algorithm, the initial clustering centers are optimized. In addition, the weighted vector identification and the weight of a sequence in its corresponding cluster are used to update the clustering centers. The theoretical experimental results and the analysis results in this study show that the SCAWVI algorithm has a higher rate of accurate results in its clustering results and higher execution efficiency.
Keywords:
Sequence clustering
Huffman
Weighted vector
Similarity

Journal

International Journal of Machine Learning and Cybernetics cover
International Journal of Machine Learning and Cybernetics
IF:
2.7
Papers:
3.2K
Citations:
5.6K

Organization

Y
Yanshan University
Scholars:
1.7W
Papers: 1.1W
Citations: 1.3W
H
Hebei University of Engineering
Scholars:
3.3K
Papers: 2.1K
Citations: 2.7K
Cited Papers

Cited Papers

no more