Explorer›Data Science›Statistics
Research PaperResearchia:202610.08032

Oracle-Efficient and Parameter-Free Agnostic Smoothed Online Learning

Sasha Voitovych

Abstract

Online learning is an attractive framework in many domains because it permits well-defined learning even when data are dependent or chosen adversarially. This generality, however, comes at a steep price, introducing significant statistical and computational barriers. Recently, smoothed online learning has emerged as a promising framework that interpolates between the fully adversarial and fully stochastic settings by assuming that the conditional law of each covariate has density at most $1/σ$ w...

Submitted: October 8, 2026Subjects: Statistics; Data Science

Description / Details

Online learning is an attractive framework in many domains because it permits well-defined learning even when data are dependent or chosen adversarially. This generality, however, comes at a steep price, introducing significant statistical and computational barriers. Recently, smoothed online learning has emerged as a promising framework that interpolates between the fully adversarial and fully stochastic settings by assuming that the conditional law of each covariate has density at most 1/σ1/σ with respect to some fixed base measure μμ, and it is known to match the statistical and computational guarantees of classical learning while still allowing for much of the flexibility of online learning. However, existing oracle-efficient algorithms require either (i) sampling access to the base measure μμ or (ii) labels that are perfectly predicted by a fixed hypothesis. Both assumptions limit the applicability of these algorithms, in contrast to statistical learning, where empirical risk minimization (ERM) learns efficiently in the agnostic setting without any knowledge of the data distribution. We show that neither assumption is necessary, giving the first oracle-efficient algorithm that achieves sublinear regret in the agnostic setting without knowledge of μμ. Our algorithm, based on Gaussian Follow-The-Perturbed-Leader, is parameter-free: it requires no knowledge of μμ, the smoothing parameter σσ, or the horizon TT, and it achieves regret O~(dT/σ)\widetilde O(d\sqrt{T/σ}) for binary classes of VC dimension dd with a single call to an ERM oracle per round, which is optimal up to a d\sqrt{d} factor. En route to establishing the regret bound, we introduce several new techniques that may be of independent interest.


Source: arXiv:2610.10499v1 - http://arxiv.org/abs/2610.10499v1 PDF: https://arxiv.org/pdf/2610.10499v1 Original Link: http://arxiv.org/abs/2610.10499v1

Please sign in to join the discussion.

No comments yet. Be the first to share your thoughts!

Access Paper
View Source PDF
Submission Info
Date:
Oct 8, 2026
Topic:
Data Science
Area:
Statistics
Comments:
0
Bookmark