Variational Bayes approach for model aggregation in unsupervised classification with Markovian dependency

被引:5
|
作者
Volant, Stevenn [1 ,2 ]
Magniette, Marie-Laure Martin [1 ,2 ,3 ,4 ,5 ]
Robin, Stephane [1 ,2 ]
机构
[1] AgroParisTech, F-75231 Paris 05, France
[2] INRA, UMR MIA 518, F-75231 Paris, France
[3] INRA, URGV, UMR 1165, F-91057 Evry, France
[4] URGV, UEVE, F-91057 Evry, France
[5] CNRS, ERL 8196, URGV, F-91057 Evry, France
关键词
Model averaging; Variational Bayes inference; Markov chain; Unsupervised classification;
D O I
10.1016/j.csda.2012.01.027
中图分类号
TP39 [计算机的应用];
学科分类号
081203 ; 0835 ;
摘要
A binary unsupervised classification problem where each observation is associated with an unobserved label that needs to be retrieved is considered. More precisely, it is assumed that there are two groups of observation: normal and abnormal. The 'normal' observations are coming from a known distribution whereas the distribution of the 'abnormal' observations is unknown. Several models have been developed to fit this unknown distribution. An alternative based on a mixture of Gaussian distributions is proposed. The inference is performed within a variational Bayesian framework and the aim is to infer the posterior probability of belonging to the class of interest. To this end, it makes little sense to estimate the number of mixture components since each mixture model provides more or less relevant information to the posterior probability estimation. By computing a weighted average (named aggregated estimator) over the model collection, Bayesian Model Averaging (BMA) is one way of combining models in order to account for information provided by each model. An aim is then the estimation of the weights and the posterior probability for a specific model. Optimal approximations of these quantities from the variational theory are derived; other approximations of the weights are also proposed. It is assumed that the data are dependent (Markovian dependency) and hence a Hidden Markov Model is considered. A simulation study is carried out to evaluate the accuracy of the estimates in terms of classification performance. An illustration on both epidemiologic and genetic datasets is presented. (C) 2012 Elsevier B.V. All rights reserved.
引用
收藏
页码:2375 / 2387
页数:13
相关论文
共 50 条
  • [41] Unsupervised Relation Extraction: A Variational Autoencoder Approach
    Yuan, Chenhan
    Eldardiry, Hoda
    2021 CONFERENCE ON EMPIRICAL METHODS IN NATURAL LANGUAGE PROCESSING (EMNLP 2021), 2021, : 1929 - 1938
  • [42] Naive Bayes Approach for Website Classification
    Rajalakshmi, R.
    Aravindan, C.
    INFORMATION TECHNOLOGY AND MOBILE COMMUNICATION, 2011, 147 : 323 - 326
  • [43] Variational Bayes based approach to robust subspace learning
    Okatani, Takayuki
    Deguchi, Koichiro
    2007 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, VOLS 1-8, 2007, : 1004 - +
  • [44] Unsupervised Deep Learning based Variational Autoencoder Model for COVID-19 Diagnosis and Classification
    Mansour, Romany F.
    Escorcia-Gutierrez, Jose
    Gamarra, Margarita
    Gupta, Deepak
    Castillo, Oscar
    Kumar, Sachin
    PATTERN RECOGNITION LETTERS, 2021, 151 : 267 - 274
  • [45] Unsupervised Evaluation and Weighted Aggregation of Ranked Classification Predictions
    Ahsen, Mehmet Eren
    Vogel, Robert M.
    Stolovitzky, Gustavo A.
    JOURNAL OF MACHINE LEARNING RESEARCH, 2019, 20
  • [46] Unsupervised evaluation and weighted aggregation of ranked classification predictions
    Ahsen, Mehmet Eren
    Vogel, Robert M.
    Stolovitzky, Gustavo A.
    Journal of Machine Learning Research, 2019, 20
  • [47] Unsupervised classification of hyperspectral data: an ICA mixture model based approach
    Shah, CA
    Arora, MK
    Varshney, PK
    INTERNATIONAL JOURNAL OF REMOTE SENSING, 2004, 25 (02) : 481 - 487
  • [48] Unsupervised Sentiment Classification: A Hybrid Sentiment-Topic Model Approach
    Blair, Stuart J.
    Bi, Yaxin
    Mulvenna, Maurice D.
    2017 IEEE 29TH INTERNATIONAL CONFERENCE ON TOOLS WITH ARTIFICIAL INTELLIGENCE (ICTAI 2017), 2017, : 453 - 460
  • [49] Computer classification of injury narratives using a fuzzy Bayes approach: Improving the model
    Marucci, Helen R.
    Lehto, Mark R.
    Corns, Helen L.
    Human Interface and the Management of Information: Methods, Techniques and Tools in Information Design, Pt 1, Proceedings, 2007, 4557 : 500 - 506
  • [50] A Variational Bayes Spatiotemporal Model for Electromagnetic Brain Mapping
    Nathoo, F. S.
    Babul, A.
    Moiseev, A.
    Virji-Babul, N.
    Beg, M. F.
    BIOMETRICS, 2014, 70 (01) : 132 - 143