Papers
Topics
Authors
Recent
Search
2000 character limit reached

Sub-sampled Newton Methods with Non-uniform Sampling

Published 2 Jul 2016 in math.OC and stat.ML | (1607.00559v2)

Abstract: We consider the problem of finding the minimizer of a convex function F:R<sup>d</sup>RF: \mathbb R<sup>d</sup> \rightarrow \mathbb R of the form F(w):=i=1<sup>n</sup>fi(w)+R(w)F(w) := \sum_{i=1}<sup>n</sup> f_i(w) + R(w) where a low-rank factorization of <sup>2</sup>fi(w)\nabla<sup>2</sup> f_i(w) is readily available. We consider the regime where ndn \gg d. As second-order methods prove to be effective in finding the minimizer to a high-precision, in this work, we propose randomized Newton-type algorithms that exploit \textit{non-uniform} sub-sampling of <sup>2</sup>fi(w)<em>i=1<sup>n{\nabla<sup>2</sup> f_i(w)}<em>{i=1}<sup>{n}, as well as inexact updates, as means to reduce the computational complexity. Two non-uniform sampling distributions based on {\it block norm squares} and {\it block partial leverage scores} are considered in order to capture important terms among <sup>2</sup>fi(w)</em>i=1<sup>n{\nabla<sup>2</sup> f_i(w)}</em>{i=1}<sup>{n}. We show that at each iteration non-uniformly sampling at most O(dlogd)\mathcal O(d \log d) terms from <sup>2</sup>fi(w)i=1<sup>n{\nabla<sup>2</sup> f_i(w)}_{i=1}<sup>{n} is sufficient to achieve a linear-quadratic convergence rate in ww when a suitable initial point is provided. In addition, we show that our algorithms achieve a lower computational complexity and exhibit more robustness and better dependence on problem specific quantities, such as the condition number, compared to similar existing methods, especially the ones based on uniform sampling. Finally, we empirically demonstrate that our methods are at least twice as fast as Newton's methods with ridge logistic regression on several real datasets.

Citations (113)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.