article

Natasha 2: Faster Non-Convex Optimization Than SGD

  • Neural Information Processing Systems
Research footprint

At a glance

الاستشهادات
92
المراجع
0
Comments
0
Paper overview

Abstract

(this is a theory paper) We design a stochastic algorithm to find e -approximate local minima of any smooth nonconvex function in rate O(e−3.25) , with only oracle access to stochastic gradients. The best result was essentially O(e−4) by stochastic gradient descent (SGD).

Record transparency

Publication details

OpenAlex
W2963926425
Document type
article
Language
EN
Source
Neural Information Processing Systems
Last metadata update
المجتمع

Comments

تسجيل الدخول للانضمام إلى النقاش.

  1. لا توجد تعليقات بعد. ابدأ النقاش.