Online and Distribution-Free Robustness: Regression and Contextual Bandits with Huber Contamination

被引：9

作者：

Chen, Sitan ^{[1
]}

Koehler, Frederic ^{[1
]}

Moitra, Ankur ^{[2
]}

Yau, Morris ^{[2
]}

机构：

[1] Univ Calif Berkeley, Berkeley, CA 94720 USA

[2] MIT, 77 Massachusetts Ave, Cambridge, MA 02139 USA

来源：

2021 IEEE 62ND ANNUAL SYMPOSIUM ON FOUNDATIONS OF COMPUTER SCIENCE (FOCS 2021) | 2022年

关键词：

robust statistics; regression; contextual bandits; online learning; Huber contamination; ASYMPTOTICS;

D O I：

10.1109/FOCS52979.2021.00072

中图分类号：

TP301 [理论、方法];

学科分类号：

081202 ;

摘要：

In this work we revisit two classic high-dimensional online learning problems, namely linear regression and contextual bandits, from the perspective of adversarial robustness. Existing works in algorithmic robust statistics make strong distributional assumptions that ensure that the input data is evenly spread out or comes from a nice generative model. Is it possible to achieve strong robustness guarantees even without distributional assumptions altogether, where the sequence of tasks we are asked to solve is adaptively and adversarially chosen? We answer this question in the affirmative for both linear regression and contextual bandits. In fact our algorithms succeed where conventional methods fail. In particular we show strong lower bounds against Huber regression and more generally any convex M-estimator. Our approach is based on a novel alternating minimization scheme that interleaves ordinary least-squares with a simple convex program that finds the optimal reweighting of the distribution under a spectral constraint. Our results obtain essentially optimal dependence on the contamination level eta, reach the optimal breakdown point, and naturally apply to infinite dimensional settings where the feature vectors are represented implicitly via a kernel map.

引用

页码：684 / 695

页数：12

共 50 条

[41] THE ROBUSTNESS OF MAXIMUM-LIKELIHOOD AND DISTRIBUTION-FREE ESTIMATORS TO NONNORMALITY IN CONFIRMATORY FACTOR-ANALYSIS
BENSON, J
FLEISHMAN, JA
QUALITY & QUANTITY, 1994, 28 (02) : 117 - 136
[42] KL-UCB-Switch: Optimal Regret Bounds for Stochastic Bandits from Both a Distribution-Dependent and a Distribution-Free Viewpoints
Garivier, Aurelien
Hadiji, Hedi
Menard, Pierre
Stoltz, Gilles
JOURNAL OF MACHINE LEARNING RESEARCH, 2022, 23 : 1 - 66
[43] KL-UCB-Switch: Optimal Regret Bounds for Stochastic Bandits from Both a Distribution-Dependent and a Distribution-Free Viewpoints
Garivier, Aurélien
Hadiji, Hédi
Ménard, Pierre
Stoltz, Gilles
Journal of Machine Learning Research, 2022, 23
[44] EFFICIENCY BOUNDS FOR DISTRIBUTION-FREE ESTIMATORS OF THE BINARY CHOICE AND THE CENSORED REGRESSION-MODELS
COSSLETT, SR
ECONOMETRICA, 1987, 55 (03) : 559 - 585
[45] Distribution-free strong consistency for nonparametric kernel regression involving nonlinear time series
Lu, ZD
Cheng, P
JOURNAL OF STATISTICAL PLANNING AND INFERENCE, 1997, 65 (01) : 67 - 86
[46] Distribution-free strong consistency for nonparametric kernel regression involving nonlinear time series
Lu, Z.
Cheng, P.
Journal of Statistical Planning and Inference, 65 (01):
[47] A DISTRIBUTION-FREE LEAST-SQUARES ESTIMATOR FOR CENSORED LINEAR-REGRESSION MODELS
HOROWITZ, JL
JOURNAL OF ECONOMETRICS, 1986, 32 (01) : 59 - 84
[48] Distribution-free predictive inference for partial least squares regression with applications to molecular descriptors datasets
Lin, Youwu
Xu, Congcong
Zhou, Zhaojun
Shen, Liang
Huang, Shuai
JOURNAL OF CHEMOMETRICS, 2022, 36 (12)
[49] Abstract: The Normal-Theory and Asymptotic Distribution-Free Covariance Matrix of Standardized Regression Coefficients
Jones, Jeff A.
Waller, Niels G.
MULTIVARIATE BEHAVIORAL RESEARCH, 2013, 48 (01) : 161 - 161
[50] A distribution-free robust method for monitoring linear profiles using rank-based regression
Zi, Xuemin
Zou, Changliang
Tsung, Fugee
IIE TRANSACTIONS, 2012, 44 (11) : 949 - 963

← 1 2 3 4 5 →