Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Escaping Saddle Points with Stochastically Controlled Stochastic Gradient Methods

Mar 13, 2021

Guannan Liang, Qianqian Tong, Chunjiang Zhu, Jinbo Bi

Figure 1 for Escaping Saddle Points with Stochastically Controlled Stochastic Gradient Methods

Figure 2 for Escaping Saddle Points with Stochastically Controlled Stochastic Gradient Methods

Figure 3 for Escaping Saddle Points with Stochastically Controlled Stochastic Gradient Methods

Share this with someone who'll enjoy it:

Abstract:Stochastically controlled stochastic gradient (SCSG) methods have been proved to converge efficiently to first-order stationary points which, however, can be saddle points in nonconvex optimization. It has been observed that a stochastic gradient descent (SGD) step introduces anistropic noise around saddle points for deep learning and non-convex half space learning problems, which indicates that SGD satisfies the correlated negative curvature (CNC) condition for these problems. Therefore, we propose to use a separate SGD step to help the SCSG method escape from strict saddle points, resulting in the CNC-SCSG method. The SGD step plays a role similar to noise injection but is more stable. We prove that the resultant algorithm converges to a second-order stationary point with a convergence rate of $\tilde{O}( \epsilon^{-2} log( 1/\epsilon))$ where $\epsilon$ is the pre-specified error tolerance. This convergence rate is independent of the problem dimension, and is faster than that of CNC-SGD. A more general framework is further designed to incorporate the proposed CNC-SCSG into any first-order method for the method to escape saddle points. Simulation studies illustrate that the proposed algorithm can escape saddle points in much fewer epochs than the gradient descent methods perturbed by either noise injection or a SGD step.

View paper on

Share this with someone who'll enjoy it:

Title:Escaping Saddle Points with Stochastically Controlled Stochastic Gradient Methods

Paper and Code