Cluster's Quality Evaluation and Selective Clustering Ensemble

被引:23
|
作者
Li, Feijiang [1 ]
Qian, Yuhua [1 ,2 ]
Wang, Jieting [1 ]
Dang, Chuangyin [3 ]
Liu, Bing [4 ]
机构
[1] Shanxi Univ, Inst Big Data Sci & Ind, Taiyuan 030006, Shanxi, Peoples R China
[2] Minist Educ, Key Lab Computat Intelligence & Chinese Informat, Taiyuan 030006, Shanxi, Peoples R China
[3] City Univ Hong Kong, Dept Manufacture Engn & Engn Management, Hong Kong, Hong Kong, Peoples R China
[4] Univ Illinois, Dept Comp Sci, Chicago, IL 60607 USA
关键词
Clustering ensemble; selective clustering ensemble; weighted clustering ensemble; cluster quality; DIVERSITY; STABILITY; CONSENSUS;
D O I
10.1145/3211872
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Clustering ensemble has drawn much attention in recent years due to its ability to generate a high quality and robust partition result. Weighted clustering ensemble and selective clustering ensemble are two general ways to further improve the performance of a clustering ensemble method. Existing weighted clustering ensemble methods assign the same weight to each cluster in a partition of the ensemble. Since the qualities of the clusters in a partition are different, the clusters should be weighted differently. To address this issue, this article proposes a new measure to calculate the similarity between a cluster and a partition. Theoretically, this measure is effective in handling two problems in measuring the quality of a cluster, which are defined as the symmetric problem and the context meaning problem. In addition, some properties of the proposed measure are analyzed. This measure can be easily expanded to a clustering performance measure that calculates the similarity between two partitions. As a result of this measure, we propose a novel selective clustering ensemble framework, which considers the differences between the objective of the ensemble selection stage and the object of the ensemble integration stage in the selective clustering ensemble. To verify the performance of the new measure, we compare the performance of the measure with the two existing measures in weighting clusters. The experiments show that the proposed measure is more effective. To verify the performance of the novel framework, four existing state-of-the-art selective clustering ensemble frameworks are employed as references. The experiments show that the proposed framework is statistically better than the others on 17 UCI benchmark datasets, 8 document datasets, and the Olivetti Face Database.
引用
收藏
页数:27
相关论文
共 50 条
  • [21] A clustering ensemble algorithm based on cluster-mode
    Jia, Rui-Yu
    Geng, Jin-Wei
    International Journal of Digital Content Technology and its Applications, 2012, 6 (19) : 17 - 24
  • [22] DICLENS: Divisive Clustering Ensemble with Automatic Cluster Number
    Mimaroglu, Selim
    Aksehirli, Emin
    IEEE-ACM TRANSACTIONS ON COMPUTATIONAL BIOLOGY AND BIOINFORMATICS, 2012, 9 (02) : 408 - 420
  • [23] Tumor Clustering based on Hybrid Cluster Ensemble Framework
    Yu, Zhiwen
    You, Jane
    Chen, Hantao
    Li, Le
    Wang, Xiaowei
    2012 INTERNATIONAL CONFERENCE ON COMPUTERIZED HEALTHCARE (ICCH), 2012, : 99 - +
  • [24] Multiple clustering and selecting algorithms with combining strategy for selective clustering ensemble
    Ma, Tinghuai
    Yu, Te
    Wu, Xiuge
    Cao, Jie
    Al-Abdulkarim, Alia
    Al-Dhelaan, Abdullah
    Al-Dhelaan, Mohammed
    SOFT COMPUTING, 2020, 24 (20) : 15129 - 15141
  • [25] Multiple clustering and selecting algorithms with combining strategy for selective clustering ensemble
    Tinghuai Ma
    Te Yu
    Xiuge Wu
    Jie Cao
    Alia Al-Abdulkarim
    Abdullah Al-Dhelaan
    Mohammed Al-Dhelaan
    Soft Computing, 2020, 24 : 15129 - 15141
  • [26] A fuzzy clustering ensemble based on cluster clustering and iterative Fusion of base clusters
    Musa Mojarad
    Samad Nejatian
    Hamid Parvin
    Majid Mohammadpoor
    Applied Intelligence, 2019, 49 : 2567 - 2581
  • [27] A fuzzy clustering ensemble based on cluster clustering and iterative Fusion of base clusters
    Mojarad, Musa
    Nejatian, Samad
    Parvin, Hamid
    Mohammadpoor, Majid
    APPLIED INTELLIGENCE, 2019, 49 (07) : 2567 - 2581
  • [28] Germplasm Evaluation using Cluster Ensemble
    Ahuja, Sangeeta
    Raiger, H. L.
    Febrice, Mudenge
    Choubey, A. K.
    Sharma, O. P.
    PROCEEDINGS OF THE 10TH INDIACOM - 2016 3RD INTERNATIONAL CONFERENCE ON COMPUTING FOR SUSTAINABLE GLOBAL DEVELOPMENT, 2016, : 1741 - 1742
  • [29] Unsupervised Evaluation of Cluster Ensemble Solutions
    Zhang, Shaohong
    Yang, Liu
    Xie, Dongqing
    2015 SEVENTH INTERNATIONAL CONFERENCE ON ADVANCED COMPUTATIONAL INTELLIGENCE (ICACI), 2015, : 101 - 106
  • [30] Clustering ensemble selection considering quality and diversity
    Abbasi, Sadr-olah
    Nejatian, Samad
    Parvin, Hamid
    Rezaie, Vahideh
    Bagherifard, Karamolah
    ARTIFICIAL INTELLIGENCE REVIEW, 2019, 52 (02) : 1311 - 1340