Papers
Topics
Authors
Recent
Search
2000 character limit reached

Learning Halfspaces and Neural Networks with Random Initialization

Published 25 Nov 2015 in cs.LG | (1511.07948v1)

Abstract: We study non-convex empirical risk minimization for learning halfspaces and neural networks. For loss functions that are LL-Lipschitz continuous, we present algorithms to learn halfspaces and multi-layer neural networks that achieve arbitrarily small excess risk $\epsilon&gt;0$. The time complexity is polynomial in the input dimension dd and the sample size nn, but exponential in the quantity (L/ϵ<sup>2)log(L/ϵ)(L/\epsilon<sup>2)\log(L/\epsilon). These algorithms run multiple rounds of random initialization followed by arbitrary optimization steps. We further show that if the data is separable by some neural network with constant margin $\gamma&gt;0$, then there is a polynomial-time algorithm for learning a neural network that separates the training data with margin Ω(γ)\Omega(\gamma). As a consequence, the algorithm achieves arbitrary generalization error $\epsilon&gt;0$ with poly(d,1/ϵ){\rm poly}(d,1/\epsilon) sample and time complexity. We establish the same learnability result when the labels are randomly flipped with probability $\eta&lt;1/2$.

Citations (35)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.