Online and Distribution-Free Robustness: Regression and Contextual Bandits with Huber Contamination

Chen, Sitan; Koehler, Frederic; Moitra, Ankur; Yau, Morris

Computer Science > Machine Learning

arXiv:2010.04157 (cs)

[Submitted on 8 Oct 2020 (v1), last revised 10 Jun 2021 (this version, v3)]

Title:Online and Distribution-Free Robustness: Regression and Contextual Bandits with Huber Contamination

Authors:Sitan Chen, Frederic Koehler, Ankur Moitra, Morris Yau

View PDF

Abstract:In this work we revisit two classic high-dimensional online learning problems, namely linear regression and contextual bandits, from the perspective of adversarial robustness. Existing works in algorithmic robust statistics make strong distributional assumptions that ensure that the input data is evenly spread out or comes from a nice generative model. Is it possible to achieve strong robustness guarantees even without distributional assumptions altogether, where the sequence of tasks we are asked to solve is adaptively and adversarially chosen?
We answer this question in the affirmative for both linear regression and contextual bandits. In fact our algorithms succeed where conventional methods fail. In particular we show strong lower bounds against Huber regression and more generally any convex M-estimator. Our approach is based on a novel alternating minimization scheme that interleaves ordinary least-squares with a simple convex program that finds the optimal reweighting of the distribution under a spectral constraint. Our results obtain essentially optimal dependence on the contamination level $\eta$, reach the optimal breakdown point, and naturally apply to infinite dimensional settings where the feature vectors are represented implicitly via a kernel map.

Comments:	66 pages, 1 figure, v3: refined exposition and improved rates
Subjects:	Machine Learning (cs.LG); Data Structures and Algorithms (cs.DS); Machine Learning (stat.ML)
Cite as:	arXiv:2010.04157 [cs.LG]
	(or arXiv:2010.04157v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2010.04157

Submission history

From: Sitan Chen [view email]
[v1] Thu, 8 Oct 2020 17:59:05 UTC (85 KB)
[v2] Mon, 8 Feb 2021 02:12:07 UTC (103 KB)
[v3] Thu, 10 Jun 2021 22:54:35 UTC (353 KB)

Computer Science > Machine Learning

Title:Online and Distribution-Free Robustness: Regression and Contextual Bandits with Huber Contamination

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Online and Distribution-Free Robustness: Regression and Contextual Bandits with Huber Contamination

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators