A semi-supervised hierarchical ensemble clustering framework based on a novel similarity metric and stratified feature sampling

被引:1
|
作者
Shi, Hui [1 ,2 ]
Peng, Qiang [1 ]
Xie, Zhiming [1 ,2 ]
Wang, Jian [1 ]
机构
[1] Shanwei Inst Technol, Shanwei 516600, Guangdong, Peoples R China
[2] Shanwei Innovat Ind Design & Res Inst, Shanwei 516600, Guangdong, Peoples R China
关键词
Hierarchical clustering; Ensemble clustering; Semi-supervised clustering; Stratified feature sampling; Similarity metric; SYSTEMS;
D O I
10.1016/j.jksuci.2023.101687
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Recently, both ensemble clustering and semi-supervised clustering have emerged as important paradigms of traditional clustering. Ensemble clustering seeks to integrate multiple clustering results from different methods or the same methods with different parameters. Semi-supervised clustering involves using a small amount of class membership information in some samples for the learning process. Meanwhile, Semi-Supervised Ensemble Clustering (SSEC) has attracted increasing attention due to its high performance. However, most SSEC algorithms are configured based on partitional clustering techniques, and there are few attempts on hierarchical clustering techniques. Even in existing hierarchybased SSEC algorithms, prior knowledge is not sufficiently used and is often applied to create primary partitions. To address these problems, we propose a Semi-supervised Hierarchical Ensemble Clustering framework based on a novel Similarity metric and stratified feature Sampling, which we call SHECSS. SHECSS uses the information of all primary partitions according to their strength to calculate the similarity between samples. Also, SHECSS is equipped with a stratified feature sampling mechanism that can improve the diversity of primary partitions and deal with high-dimensional data. Here, the primary partitions are created based on multiple hierarchical clustering techniques, and the target partition is configured by a consensus function based on the clusters clustering policy. Experimental results show the effectiveness and efficiency of SHECSS compared to representative clustering methods.(c) 2023 The Author(s). Published by Elsevier B.V. on behalf of King Saud University. This is an open access article under the CC BY license (http://creativecommons.org/licenses/by/4.0/).
引用
收藏
页数:11
相关论文
共 50 条
  • [1] Stratified Feature Sampling for Semi-Supervised Ensemble Clustering
    Tian, Jialin
    Ren, Yazhou
    Cheng, Xiang
    IEEE ACCESS, 2019, 7 : 128669 - 128675
  • [2] Semi-supervised hierarchical ensemble clustering based on an innovative distance metric and constraint information
    Shen, Baohua
    Jiang, Juan
    Qian, Feng
    Li, Daoguo
    Ye, Yanming
    Ahmadi, Gholamreza
    ENGINEERING APPLICATIONS OF ARTIFICIAL INTELLIGENCE, 2023, 124
  • [3] Developing ensemble clustering through similarity measures: A semi-supervised hierarchical clustering learning
    Wang, Dandan
    Li, Qi
    CONCURRENCY AND COMPUTATION-PRACTICE & EXPERIENCE, 2024, 36 (16):
  • [4] Semi-supervised hierarchical clustering ensemble and its application
    Xiao, Wenchao
    Yang, Yan
    Wang, Hongjun
    Li, Tianrui
    Xing, Huanlai
    NEUROCOMPUTING, 2016, 173 : 1362 - 1376
  • [5] Combined constraint-based with metric-based in semi-supervised clustering ensemble
    Siting Wei
    Zhixin Li
    Canlong Zhang
    International Journal of Machine Learning and Cybernetics, 2018, 9 : 1085 - 1100
  • [6] Combined constraint-based with metric-based in semi-supervised clustering ensemble
    Wei, Siting
    Li, Zhixin
    Zhang, Canlong
    INTERNATIONAL JOURNAL OF MACHINE LEARNING AND CYBERNETICS, 2018, 9 (07) : 1085 - 1100
  • [7] A semi-supervised framework for concept-based hierarchical document clustering
    Seyed Mojtaba Sadjadi
    Hoda Mashayekhi
    Hamid Hassanpour
    World Wide Web, 2023, 26 : 3861 - 3890
  • [8] A semi-supervised framework for concept-based hierarchical document clustering
    Sadjadi, Seyed Mojtaba
    Mashayekhi, Hoda
    Hassanpour, Hamid
    WORLD WIDE WEB-INTERNET AND WEB INFORMATION SYSTEMS, 2023, 26 (06): : 3861 - 3890
  • [9] Hierarchical Text Clustering and Categorisation using A Semi-Supervised Framework
    Mahyoub, Mohamed
    Hind, Jade
    Woods, David
    Wong, Carl
    Hussain, Abir
    Aljumeily, Dhiya
    12TH INTERNATIONAL CONFERENCE ON THE DEVELOPMENTS IN ESYSTEMS ENGINEERING (DESE 2019), 2019, : 153 - 159
  • [10] Semi-supervised spectral clustering ensemble
    1600, ICIC Express Letters Office (10):