Cluster's Quality Evaluation and Selective Clustering Ensemble

被引:23
|
作者
Li, Feijiang [1 ]
Qian, Yuhua [1 ,2 ]
Wang, Jieting [1 ]
Dang, Chuangyin [3 ]
Liu, Bing [4 ]
机构
[1] Shanxi Univ, Inst Big Data Sci & Ind, Taiyuan 030006, Shanxi, Peoples R China
[2] Minist Educ, Key Lab Computat Intelligence & Chinese Informat, Taiyuan 030006, Shanxi, Peoples R China
[3] City Univ Hong Kong, Dept Manufacture Engn & Engn Management, Hong Kong, Hong Kong, Peoples R China
[4] Univ Illinois, Dept Comp Sci, Chicago, IL 60607 USA
关键词
Clustering ensemble; selective clustering ensemble; weighted clustering ensemble; cluster quality; DIVERSITY; STABILITY; CONSENSUS;
D O I
10.1145/3211872
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Clustering ensemble has drawn much attention in recent years due to its ability to generate a high quality and robust partition result. Weighted clustering ensemble and selective clustering ensemble are two general ways to further improve the performance of a clustering ensemble method. Existing weighted clustering ensemble methods assign the same weight to each cluster in a partition of the ensemble. Since the qualities of the clusters in a partition are different, the clusters should be weighted differently. To address this issue, this article proposes a new measure to calculate the similarity between a cluster and a partition. Theoretically, this measure is effective in handling two problems in measuring the quality of a cluster, which are defined as the symmetric problem and the context meaning problem. In addition, some properties of the proposed measure are analyzed. This measure can be easily expanded to a clustering performance measure that calculates the similarity between two partitions. As a result of this measure, we propose a novel selective clustering ensemble framework, which considers the differences between the objective of the ensemble selection stage and the object of the ensemble integration stage in the selective clustering ensemble. To verify the performance of the new measure, we compare the performance of the measure with the two existing measures in weighting clusters. The experiments show that the proposed measure is more effective. To verify the performance of the novel framework, four existing state-of-the-art selective clustering ensemble frameworks are employed as references. The experiments show that the proposed framework is statistically better than the others on 17 UCI benchmark datasets, 8 document datasets, and the Olivetti Face Database.
引用
收藏
页数:27
相关论文
共 50 条
  • [1] A Comparative Study of Selective Cluster Ensemble for Document Clustering
    Xu, Sen
    Gao, Jun
    Xu, Xiufang
    Li, Xianfeng
    Yu, Hualong
    2015 7TH INTERNATIONAL CONFERENCE ON INTELLIGENT HUMAN-MACHINE SYSTEMS AND CYBERNETICS IHMSC 2015, VOL I, 2015, : 308 - 311
  • [2] Similar Manifold Learning Based on Selective Cluster Ensemble for Image Clustering
    Luo X.-H.
    Li F.-Z.
    Zhang L.
    Gao J.-J.
    Ruan Jian Xue Bao/Journal of Software, 2020, 31 (04): : 991 - 1001
  • [3] A comprehensive study of clustering ensemble weighting based on cluster quality and diversity
    Nazari, Ahmad
    Dehghan, Ayob
    Nejatian, Samad
    Rezaie, Vahideh
    Parvin, Hamid
    PATTERN ANALYSIS AND APPLICATIONS, 2019, 22 (01) : 133 - 145
  • [4] A comprehensive study of clustering ensemble weighting based on cluster quality and diversity
    Ahmad Nazari
    Ayob Dehghan
    Samad Nejatian
    Vahideh Rezaie
    Hamid Parvin
    Pattern Analysis and Applications, 2019, 22 : 133 - 145
  • [5] From Ensemble Clustering to Subspace Clustering: Cluster Structure Encoding
    Tao, Zhiqiang
    Li, Jun
    Fu, Huazhu
    Kong, Yu
    Fu, Yun
    IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2023, 34 (05) : 2670 - 2681
  • [6] A New Selective Clustering Ensemble Algorithm
    Liu Limin
    Fan Xiaoping
    2012 NINTH IEEE INTERNATIONAL CONFERENCE ON E-BUSINESS ENGINEERING (ICEBE), 2012, : 45 - 49
  • [7] Selective Affinity Propagation Ensemble Clustering
    Lei, Qi
    Li, Ting
    2019 3RD IEEE CONFERENCE ON CONTROL TECHNOLOGY AND APPLICATIONS (IEEE CCTA 2019), 2019, : 889 - 894
  • [8] A Clustering Ensemble Method Based on Cluster Selection and Cluster Splitting
    Tang, Yuyang
    Liu, Xiabi
    PROCEEDINGS OF 2018 10TH INTERNATIONAL CONFERENCE ON MACHINE LEARNING AND COMPUTING (ICMLC 2018), 2018, : 54 - 58
  • [9] Cluster Quality Based Performance Evaluation of Hierarchical Clustering Method
    Nisha
    Kaur, Puneet Jai
    2015 1ST INTERNATIONAL CONFERENCE ON NEXT GENERATION COMPUTING TECHNOLOGIES (NGCT), 2015, : 649 - 653
  • [10] Fuzzy clustering ensemble considering cluster dependability
    School of Information Engineering, China University of Geosciences , Beijing, China
    不详
    不详
    不详
    不详
    不详
    不详
    Int. J. on Artif. Intell. Tools, 2021, 2