Classification using generalized partial least squares

被引:49
|
作者
Ding, BY
Gentleman, R
机构
[1] Amgen Inc, Med Affairs Biostat, Newbury Pk, CA 91320 USA
[2] Fred Hutchinson Canc Res Ctr, Program Computat Biol, Div Publ Hlth Sci, Seattle, WA 98104 USA
关键词
cross-validation; Firth's procedure; gene expression; iteratively reweighted partial least squares; (quasi) separation; two-stage PLS;
D O I
10.1198/106186005X47697
中图分类号
O21 [概率论与数理统计]; C8 [统计学];
学科分类号
020208 ; 070103 ; 0714 ;
摘要
Advances in computational biology have made simultaneous monitoring of thousands of features possible. The high throughput technologies not only bring about a much richer information context in which to study various aspects of gene function, but they also present the challenge of analyzing data with a large number of covariates and few samples. As an integral part of machine learning, classification of samples into two or more categories is almost always of interest to scientists. We address the question of classification in this setting by extending partial least squares (PLS), a popular dimension reduction tool in chemometrics, in the context of generalized linear regression, based on a previous approach, iteratively reweighted partial least squares, that is, IRWPLS. We compare our results with two-stage PLS and with other classifiers. We show that by phrasing the problem in a generalized linear model setting and by applying Firth's procedure to avoid (quasi)separation, we often get lower classification error rates.
引用
收藏
页码:280 / 298
页数:19
相关论文
共 50 条
  • [1] Classification using partial least squares with penalized logistic regression
    Fort, G
    Lambert-Lacroix, S
    [J]. BIOINFORMATICS, 2005, 21 (07) : 1104 - 1111
  • [2] Protein family classification with partial least squares
    Opiyo, Stephen O.
    Moriyama, Etsuko N.
    [J]. JOURNAL OF PROTEOME RESEARCH, 2007, 6 (02) : 846 - 853
  • [3] MULTIVARIATE FUNCTIONAL PARTIAL LEAST SQUARES FOR CLASSIFICATION USING LONGITUDINAL DATA
    Dembowska, Sonia
    Frangi, Alex
    Houwing-Duistermaat, Jeanine
    Liu, Haiyan
    [J]. THEORETICAL BIOLOGY FORUM, 2021, 114 (01) : 75 - 88
  • [4] Materials classification by partial least squares using S-parameters
    Turgut Ozturk
    İhsan Uluer
    İlhami Ünal
    [J]. Journal of Materials Science: Materials in Electronics, 2016, 27 : 12701 - 12706
  • [5] Materials classification by partial least squares using S-parameters
    Ozturk, Turgut
    Uluer, Ihsan
    Unal, Ilhami
    [J]. JOURNAL OF MATERIALS SCIENCE-MATERIALS IN ELECTRONICS, 2016, 27 (12) : 12701 - 12706
  • [6] Document-Zone Classification using Partial Least Squares and Hybrid Classifiers
    Abd-Almageed, Wael
    Agrawal, Mudit
    Seo, Wontaek
    Doermann, David
    [J]. 19TH INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION, VOLS 1-6, 2008, : 1869 - 1872
  • [7] PARTIAL LEAST-SQUARES AND CLASSIFICATION AND REGRESSION TREES
    YEH, CH
    SPIEGELMAN, CH
    [J]. CHEMOMETRICS AND INTELLIGENT LABORATORY SYSTEMS, 1994, 22 (01) : 17 - 23
  • [8] Multi-label Classification Using Hypergraph Orthonormalized Partial Least Squares
    Luo, Gaofeng
    Huang, Tongcheng
    Shi, Zijuan
    [J]. JOURNAL OF COMPUTERS, 2014, 9 (06) : 1364 - 1370
  • [9] Partial least squares classification for high dimensional data using the PCOUT algorithm
    Turkmen, Asuman
    Billor, Nedret
    [J]. COMPUTATIONAL STATISTICS, 2013, 28 (02) : 771 - 788
  • [10] Principal balances of compositional data for regression and classification using partial least squares
    Nesrstova, V.
    Wilms, I.
    Palarea-Albaladejo, J.
    Filzmoser, P.
    Martin-Fernandez, J. A.
    Friedecky, D.
    Hron, K.
    [J]. JOURNAL OF CHEMOMETRICS, 2023, 37 (12)