arrow
返回

Sequence clustering algorithm based on weighted vector identification

delete2015-06-03
delete2
PRE
AI
吴
吴迪 (Di Wu) *
任
任家东 (Jiadong Ren)
DOI:10.1007/s13042-015-0381-2delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Sequence clustering has become an important topic that experts in data mining are currently investigating. However, clustering quality is typically significantly affected by both the selection of initial centers and the mean sequences. In this study, the sequence clustering algorithm based on weighted vector identification (SCAWVI) algorithm is developed based on sequence element composite similarity and the weight of a sequence in its corresponding cluster. Based on the weighted sequence element, all sequences in the sequence database are preprocessed into M-dimensional weighted vector identifications. Then, using Huffman-based initial clustering centers optimization algorithm, the initial clustering centers are optimized. In addition, the weighted vector identification and the weight of a sequence in its corresponding cluster are used to update the clustering centers. The theoretical experimental results and the analysis results in this study show that the SCAWVI algorithm has a higher rate of accurate results in its clustering results and higher execution efficiency.
Keyword:
Sequence clustering
Huffman
Weighted vector
Similarity

期刊

International Journal of Machine Learning and Cybernetics 封面图
International Journal of Machine Learning and Cybernetics
IF:
2.7
论文数:
3.2K
被引数:
5.6K

机构

Y
Yanshan University
学者数:
1.7W
论文数: 1.1W
被引数: 1.3W
H
Hebei University of Engineering
学者数:
3.3K
论文数: 2.1K
被引数: 2.7K
引用论文

引用论文

没有更多内容