返回
CBCG: A Clustering Algorithm Based on Bidirectional Conical Information Granularity
DOI:10.1109/TFUZZ.2024.3397808.png)
摘要
En 中文
In this article, we propose a novel center-based clustering algorithm based on bidirectional conical information granularity. The main purpose is to fully absorb the semantic information of the ordinal relationship between objects to improve the performance of central clustering in identifying interleaved and imbalanced data. The proposed algorithm includes two main stages: first, the stage of determining the cluster center and second, the division stage. In the stage of determining the cluster center, the first cluster center is determined by using the number of conical information granularity in the data, and the remaining cluster centers are determined by defining the statistical measure of fuzzy importance degree. In the division stage, we divide the points to be clustered into stable and active areas. The former quickly and accurately identifies and assigns the objects belonging to a cluster by measuring the fuzzy similarity between the objects to be clustered and the cluster center, and the latter assigns the objects in the active area by using the information of the points already assigned. This method describes the position and sorting relationship of objects that are granulated through ordinal relationships more accurately in the global environment, thereby gaining a more comprehensive understanding of the structural characteristics of the data. This helps to improve the accuracy and stability of clustering algorithms in handling interleaved and imbalanced data. This article uses three clustering validity indicators to test the performance of our algorithm. We compare the results with those of six different types of popular clustering algorithms and new algorithms proposed in recent years. The experimental results show that the algorithm proposed in this article can identify clusters more accurately on the datasets with a complex and staggered distribution. It is significantly better than the clustering algorithm participating in the comparison and has good robustness on datasets with added noise.
Keyword:
Clustering algorithms
Granular computing
Semantics
Fuzzy systems
Data mining
Clustering methods
Data models
Bidirectional conical information granularity
k-bidirectional conical information granularity
center-based clustering
fuzzy importance degree (FID)
two-step division
期刊
IF:
11.9
论文数:
5.0K
被引数:
2.9W
机构
引用论文
A unified inference procedure for a class of measures to assess improvement in risk prediction systems with survival data一类具有生存数据的风险预测系统评估改进措施的统一推理程序

