Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Global Convergence and Stability of Stochastic Gradient Descent

Oct 04, 2021

Vivak Patel, Bowen Tian, Shushu Zhang

Figure 1 for Global Convergence and Stability of Stochastic Gradient Descent

Figure 2 for Global Convergence and Stability of Stochastic Gradient Descent

Figure 3 for Global Convergence and Stability of Stochastic Gradient Descent

Share this with someone who'll enjoy it:

Abstract:In machine learning, stochastic gradient descent (SGD) is widely deployed to train models using highly non-convex objectives with equally complex noise models. Unfortunately, SGD theory often makes restrictive assumptions that fail to capture the non-convexity of real problems, and almost entirely ignore the complex noise models that exist in practice. In this work, we make substantial progress on this shortcoming. First, we establish that SGD's iterates will either globally converge to a stationary point or diverge under nearly arbitrary nonconvexity and noise models. Under a slightly more restrictive assumption on the joint behavior of the non-convexity and noise model that generalizes current assumptions in the literature, we show that the objective function cannot diverge, even if the iterates diverge. As a consequence of our results, SGD can be applied to a greater range of stochastic optimization problems with confidence about its global convergence behavior and stability.

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:Global Convergence and Stability of Stochastic Gradient Descent

Paper and Code