article

Natasha 2: Faster Non-Convex Optimization Than SGD

  • Neural Information Processing Systems
Research footprint

At a glance

Citations
92
References
0
Comments
0
Paper overview

Abstract

(this is a theory paper) We design a stochastic algorithm to find e -approximate local minima of any smooth nonconvex function in rate O(e−3.25) , with only oracle access to stochastic gradients. The best result was essentially O(e−4) by stochastic gradient descent (SGD).

Record transparency

Publication details

OpenAlex
W2963926425
Document type
article
Language
EN
Source
Neural Information Processing Systems
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.