Cluster's Quality Evaluation and Selective Clustering Ensemble

被引:23
|
作者
Li, Feijiang [1 ]
Qian, Yuhua [1 ,2 ]
Wang, Jieting [1 ]
Dang, Chuangyin [3 ]
Liu, Bing [4 ]
机构
[1] Shanxi Univ, Inst Big Data Sci & Ind, Taiyuan 030006, Shanxi, Peoples R China
[2] Minist Educ, Key Lab Computat Intelligence & Chinese Informat, Taiyuan 030006, Shanxi, Peoples R China
[3] City Univ Hong Kong, Dept Manufacture Engn & Engn Management, Hong Kong, Hong Kong, Peoples R China
[4] Univ Illinois, Dept Comp Sci, Chicago, IL 60607 USA
关键词
Clustering ensemble; selective clustering ensemble; weighted clustering ensemble; cluster quality; DIVERSITY; STABILITY; CONSENSUS;
D O I
10.1145/3211872
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Clustering ensemble has drawn much attention in recent years due to its ability to generate a high quality and robust partition result. Weighted clustering ensemble and selective clustering ensemble are two general ways to further improve the performance of a clustering ensemble method. Existing weighted clustering ensemble methods assign the same weight to each cluster in a partition of the ensemble. Since the qualities of the clusters in a partition are different, the clusters should be weighted differently. To address this issue, this article proposes a new measure to calculate the similarity between a cluster and a partition. Theoretically, this measure is effective in handling two problems in measuring the quality of a cluster, which are defined as the symmetric problem and the context meaning problem. In addition, some properties of the proposed measure are analyzed. This measure can be easily expanded to a clustering performance measure that calculates the similarity between two partitions. As a result of this measure, we propose a novel selective clustering ensemble framework, which considers the differences between the objective of the ensemble selection stage and the object of the ensemble integration stage in the selective clustering ensemble. To verify the performance of the new measure, we compare the performance of the measure with the two existing measures in weighting clusters. The experiments show that the proposed measure is more effective. To verify the performance of the novel framework, four existing state-of-the-art selective clustering ensemble frameworks are employed as references. The experiments show that the proposed framework is statistically better than the others on 17 UCI benchmark datasets, 8 document datasets, and the Olivetti Face Database.
引用
收藏
页数:27
相关论文
共 50 条
  • [31] Clustering ensemble selection considering quality and diversity
    Sadr-olah Abbasi
    Samad Nejatian
    Hamid Parvin
    Vahideh Rezaie
    Karamolah Bagherifard
    Artificial Intelligence Review, 2019, 52 : 1311 - 1340
  • [32] Subspace Selective Ensemble Algorithm Based on Feature Clustering
    Tao, Hui
    Ma, Xiao-ping
    Qiao, Mei-ying
    JOURNAL OF COMPUTERS, 2013, 8 (02) : 509 - 516
  • [33] Clustering-based selective neural network ensemble
    Fu Q.
    Hu S.-X.
    Zhao S.-Y.
    Journal of Zhejiang University-SCIENCE A, 2005, 6 (5): : 387 - 392
  • [34] Selective SVM Ensemble Based on Improved Spectral Clustering
    Chen, Tao
    INTERNATIONAL CONFERENCE ON ENGINEERING AND BUSINESS MANAGEMENT (EBM2011), VOLS 1-6, 2011, : 2636 - 2639
  • [35] A Point-Cluster-Partition Architecture for Weighted Clustering Ensemble
    Li, Na
    Xu, Sen
    Xu, Heyang
    Xu, Xiufang
    Guo, Naixuan
    Cai, Na
    NEURAL PROCESSING LETTERS, 2024, 56 (03)
  • [36] Fast Constrained Spectral Clustering and Cluster Ensemble with Random Projection
    Liu, Wenfen
    Ye, Mao
    Wei, Jianghong
    Hu, Xuexian
    COMPUTATIONAL INTELLIGENCE AND NEUROSCIENCE, 2017, 2017
  • [37] Heterogeneous clustering ensemble method for combining different cluster results
    Yoon, Hye-Sung
    Ahn, Sun-Young
    Lee, Sang-Ho
    Cho, Sung-Bum
    Kim, Ju Han
    DATA MINING FOR BIOMEDICAL APPLICATIONS, PROCEEDINGS, 2006, 3916 : 82 - 92
  • [38] A New Assessment of Cluster Tendency Ensemble approach for Data Clustering
    Pham Van Nha
    Ngo Thanh Long
    Pham The Long
    Pham Van Hai
    PROCEEDINGS OF THE NINTH INTERNATIONAL SYMPOSIUM ON INFORMATION AND COMMUNICATION TECHNOLOGY (SOICT 2018), 2018, : 216 - 221
  • [39] The Core Cluster-Based Subspace Weighted Clustering Ensemble
    Huang, Xuan
    Qin, Fang
    Lin, Lin
    WIRELESS COMMUNICATIONS & MOBILE COMPUTING, 2022, 2022
  • [40] Clustering Ensemble Algorithm with Cluster Connection Based on Wisdom of Crowds
    Zhang H.
    Gao Y.
    Chen Y.
    Wang Z.
    Gao, Yukun (821566504@qq.com), 2018, Science Press (55): : 2611 - 2619