A Unifying Probabilistic Framework for Partially Labeled Data Learning

被引:3
|
作者
Gong, Xiuwen [1 ]
Yuan, Dong [1 ]
Bao, Wei [1 ]
Luo, Fulin [2 ]
机构
[1] Univ Sydney, Fac Engn, Camperdown, NSW 2006, Australia
[2] Chongqing Univ, Coll Comp Sci, Chongqing 400044, Peoples R China
基金
中国国家自然科学基金;
关键词
Phase locked loops; Correlation; Training; Probabilistic logic; Testing; Task analysis; Noise measurement; Partially labeled data learning (PLDL); partial label learning (PLL); partial multi-label learning (PML); classification;
D O I
10.1109/TPAMI.2022.3228755
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Partially labeled data learning (PLDL), including partial label learning (PLL) and partial multi-label learning (PML), has been widely used in nowadays data science. Researchers attempt to construct different specific models to deal with the different classification tasks for PLL and PML scenarios respectively. The main challenge in training classifiers for PLL and PML is how to deal with ambiguities caused by the noisy false-positive labels in the candidate label set. The state-of-the-art strategy for both scenarios is to perform disambiguation by identifying the ground-truth label(s) directly from the candidate label set, which can be summarized into two categories: 'the identifying method' and 'the embedding method'. However, both kinds of methods are constructed by hand-designed heuristic modeling under considerations like feature/label correlations with no theoretical interpretation. Instead of adopting heuristic or specific modeling, we propose a novel unifying framework called A Unifying Probabilistic Framework for Partially Labeled Data Learning (UPF-PLDL), which is derived from a clear probabilistic formulation, and brings existing research on PLL and PML under one theoretical interpretation with respect to information theory. Furthermore, the proposed UPF-PLDL also unifies 'the identifying method' and 'the embedding method' into one integrated framework, which naturally incorporates the feature and label correlation considerations. Comprehensive experiments on synthetic and real-world datasets for both PLL and PML scenarios clearly demonstrate the superiorities of the derived framework.
引用
收藏
页码:8036 / 8048
页数:13
相关论文
共 50 条
  • [41] Wrapper feature selection with partially labeled data
    Feofanov, Vasilii
    Devijver, Emilie
    Amini, Massih-Reza
    [J]. APPLIED INTELLIGENCE, 2022, 52 (11) : 12316 - 12329
  • [42] Wrapper feature selection with partially labeled data
    Vasilii Feofanov
    Emilie Devijver
    Massih-Reza Amini
    [J]. Applied Intelligence, 2022, 52 : 12316 - 12329
  • [43] Cost-sensitive ensemble learning: a unifying framework
    George Petrides
    Wouter Verbeke
    [J]. Data Mining and Knowledge Discovery, 2022, 36 : 1 - 28
  • [44] A Deep Probabilistic Transfer Learning Framework for Soft Sensor Modeling With Missing Data
    Chai, Zheng
    Zhao, Chunhui
    Huang, Biao
    Chen, Hongtian
    [J]. IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2022, 33 (12) : 7598 - 7609
  • [45] Cost-sensitive ensemble learning: a unifying framework
    Petrides, George
    Verbeke, Wouter
    [J]. DATA MINING AND KNOWLEDGE DISCOVERY, 2022, 36 (01) : 1 - 28
  • [46] Learning in real-time search: A unifying framework
    Bulitko, V
    Lee, G
    [J]. JOURNAL OF ARTIFICIAL INTELLIGENCE RESEARCH, 2006, 25 : 119 - 157
  • [47] Toward Deep Supervised Anomaly Detection: Reinforcement Learning from Partially Labeled Anomaly Data
    Pang, Guansong
    van den Hengel, Anton
    Shen, Chunhua
    Cao, Longbing
    [J]. KDD '21: PROCEEDINGS OF THE 27TH ACM SIGKDD CONFERENCE ON KNOWLEDGE DISCOVERY & DATA MINING, 2021, : 1298 - 1308
  • [48] UnifyDR: A Generic Framework for Unifying Data and Replica Placement
    Atrey, Ankita
    Van Seghbroeck, Gregory
    Mora, Higinio
    Volckaert, Bruno
    De Turck, Filip
    [J]. IEEE ACCESS, 2020, 8 : 216894 - 216910
  • [49] A framework for reading and unifying heliophysics time series data
    Vandegriff, Jon
    Brown, Lawrence
    [J]. EARTH SCIENCE INFORMATICS, 2010, 3 (1-2) : 75 - 86
  • [50] A framework for reading and unifying heliophysics time series data
    Jon Vandegriff
    Lawrence Brown
    [J]. Earth Science Informatics, 2010, 3 : 75 - 86