A semi-supervised hierarchical ensemble clustering framework based on a novel similarity metric and stratified feature sampling

被引:1
|
作者
Shi, Hui [1 ,2 ]
Peng, Qiang [1 ]
Xie, Zhiming [1 ,2 ]
Wang, Jian [1 ]
机构
[1] Shanwei Inst Technol, Shanwei 516600, Guangdong, Peoples R China
[2] Shanwei Innovat Ind Design & Res Inst, Shanwei 516600, Guangdong, Peoples R China
关键词
Hierarchical clustering; Ensemble clustering; Semi-supervised clustering; Stratified feature sampling; Similarity metric; SYSTEMS;
D O I
10.1016/j.jksuci.2023.101687
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Recently, both ensemble clustering and semi-supervised clustering have emerged as important paradigms of traditional clustering. Ensemble clustering seeks to integrate multiple clustering results from different methods or the same methods with different parameters. Semi-supervised clustering involves using a small amount of class membership information in some samples for the learning process. Meanwhile, Semi-Supervised Ensemble Clustering (SSEC) has attracted increasing attention due to its high performance. However, most SSEC algorithms are configured based on partitional clustering techniques, and there are few attempts on hierarchical clustering techniques. Even in existing hierarchybased SSEC algorithms, prior knowledge is not sufficiently used and is often applied to create primary partitions. To address these problems, we propose a Semi-supervised Hierarchical Ensemble Clustering framework based on a novel Similarity metric and stratified feature Sampling, which we call SHECSS. SHECSS uses the information of all primary partitions according to their strength to calculate the similarity between samples. Also, SHECSS is equipped with a stratified feature sampling mechanism that can improve the diversity of primary partitions and deal with high-dimensional data. Here, the primary partitions are created based on multiple hierarchical clustering techniques, and the target partition is configured by a consensus function based on the clusters clustering policy. Experimental results show the effectiveness and efficiency of SHECSS compared to representative clustering methods.(c) 2023 The Author(s). Published by Elsevier B.V. on behalf of King Saud University. This is an open access article under the CC BY license (http://creativecommons.org/licenses/by/4.0/).
引用
收藏
页数:11
相关论文
共 50 条
  • [31] Semi-supervised sentiment classification based on sentiment feature clustering
    Li, Suke
    Jiang, Yanbing
    Jisuanji Yanjiu yu Fazhan/Computer Research and Development, 2013, 50 (12): : 2570 - 2577
  • [32] Semi-Supervised Clustering Algorithm Based on Deep Feature Mapping
    Xu, Xiong
    Zhou, Chun
    Wang, Chenggang
    Zhang, Xiaoyan
    Meng, Hua
    INTELLIGENT AUTOMATION AND SOFT COMPUTING, 2023, 37 (01): : 815 - 831
  • [33] A Feature Space Learning Model Based on Semi-Supervised Clustering
    Guan, Renchu
    Wang, Xu
    Marchese, Maurizio
    Liang, Yanchun
    Yang, Chen
    2017 IEEE INTERNATIONAL CONFERENCE ON COMPUTATIONAL SCIENCE AND ENGINEERING (CSE) AND IEEE/IFIP INTERNATIONAL CONFERENCE ON EMBEDDED AND UBIQUITOUS COMPUTING (EUC), VOL 1, 2017, : 403 - 409
  • [34] Semi-supervised affinity propagation clustering algorithm based on stratified combination
    Zhang, Z. (zhangzhen2096@163.com), 2013, Science Press (35):
  • [35] Semi-Supervised Fuzzy Clustering with Feature Discrimination
    Li, Longlong
    Garibaldi, Jonathan M.
    He, Dongjian
    Wang, Meili
    PLOS ONE, 2015, 10 (09):
  • [36] Eigenvectors selection for spectral clustering based on semi-supervised selective ensemble
    Wang, X. (wangxingliang0911@163.com), 1600, Binary Information Press (10):
  • [37] Deep semi-supervised clustering based on pairwise constraints and sample similarity
    Qin, Xiao
    Yuan, Changan
    Jiang, Jianhui
    Chen, Long
    PATTERN RECOGNITION LETTERS, 2024, 178 : 1 - 6
  • [38] Semi-Supervised Clustering with Multi-Viewpoint based Similarity Measure
    Yan, Yang
    Chen, Lihui
    Duc Thang Nguyen
    2012 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS (IJCNN), 2012,
  • [39] Constraint projections for semi-supervised spectral clustering ensemble
    Yang, Jingya
    Sun, Linfu
    Wu, Qishi
    CONCURRENCY AND COMPUTATION-PRACTICE & EXPERIENCE, 2019, 31 (20):
  • [40] RAPID CLUSTERING WITH SEMI-SUPERVISED ENSEMBLE DENSITY CENTERS
    Kadhim, Mustafa R.
    Tian, Wenhong
    Khan, Tahseen
    2019 16TH INTERNATIONAL COMPUTER CONFERENCE ON WAVELET ACTIVE MEDIA TECHNOLOGY AND INFORMATION PROCESSING (ICWAMTIP), 2019, : 230 - 235