Quantile-Composited Feature Screening for Ultrahigh-Dimensional Data

被引:0
|
作者
Chen, Shuaishuai [1 ]
Lu, Jun [2 ]
机构
[1] Shandong Univ, Sch Math, Jinan 250100, Peoples R China
[2] Natl Univ Def & Technol, Sch Sci, Changsha 410000, Peoples R China
基金
中国国家自然科学基金;
关键词
feature screening; discriminative analysis; quantile-composited; CLASSIFICATION;
D O I
10.3390/math11102398
中图分类号
O1 [数学];
学科分类号
0701 ; 070101 ;
摘要
Ultrahigh-dimensional grouped data are frequently encountered by biostatisticians working on multi-class categorical problems. To rapidly screen out the null predictors, this paper proposes a quantile-composited feature screening procedure. The new method first transforms the continuous predictor to a Bernoulli variable, by thresholding the predictor at a certain quantile. Consequently, the independence between the response and each predictor is easy to judge, by employing the Pearson chi-square statistic. The newly proposed method has the following salient features: (1) it is robust against high-dimensional heterogeneous data; (2) it is model-free, without specifying any regression structure between the covariate and outcome variable; (3) it enjoys a low computational cost, with the computational complexity controlled at the sample size level. Under some mild conditions, the new method was shown to achieve the sure screening property without imposing any moment condition on the predictors. Numerical studies and real data analyses further confirmed the effectiveness of the new screening procedure.
引用
收藏
页数:21
相关论文
共 50 条