Papers
Topics
Authors
Recent
Search
2000 character limit reached

Thinking Outside the Ball: Optimal Learning with Gradient Descent for Generalized Linear Stochastic Convex Optimization

Published 27 Feb 2022 in cs.LG and math.OC | (2202.13328v2)

Abstract: We consider linear prediction with a convex Lipschitz loss, or more generally, stochastic convex optimization problems of generalized linear form, i.e.~where each instantaneous loss is a scalar convex function of a linear function. We show that in this setting, early stopped Gradient Descent (GD), without any explicit regularization or projection, ensures excess error at most ϵ\epsilon (compared to the best possible with unit Euclidean norm) with an optimal, up to logarithmic factors, sample complexity of O~(1/ϵ<sup>2)\tilde{O}(1/\epsilon<sup>2) and only O~(1/ϵ<sup>2)\tilde{O}(1/\epsilon<sup>2) iterations. This contrasts with general stochastic convex optimization, where Ω(1/ϵ<sup>4)\Omega(1/\epsilon<sup>4) iterations are needed Amir et al. [2021b]. The lower iteration complexity is ensured by leveraging uniform convergence rather than stability. But instead of uniform convergence in a norm ball, which we show can guarantee suboptimal learning using Θ(1/ϵ<sup>4)\Theta(1/\epsilon<sup>4) samples, we rely on uniform convergence in a distribution-dependent ball.

Authors (3)
Citations (5)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.