Faster Stochastic Quasi-Newton Methods.

Qingsong Zhang,Feihu Huang,Cheng Deng,Heng Huang

doi:10.1109/tnnls.2021.3056947

Abstract

Stochastic optimization methods have become a class of popular optimization tools in machine learning. Especially, stochastic gradient descent (SGD) has been widely used for machine learning problems, such as training neural networks, due to low per-iteration computational complexity. In fact, the Newton or quasi-newton (QN) methods leveraging the second-order information are able to achieve a better solution than the first-order methods. Thus, stochastic QN (SQN) methods have been developed to achieve a better solution efficiently than the stochastic first-order methods by utilizing approximate second-order information. However, the existing SQN methods still do not reach the best known stochastic first-order oracle (SFO) complexity. To fill this gap, we propose a novel faster stochastic QN method (SpiderSQN) based on the variance reduced technique of SIPDER. We prove that our SpiderSQN method reaches the best known SFO complexity of O(n+n1/2ϵ-2) in the finite-sum setting to obtain an ϵ -first-order stationary point. To further improve its practical performance, we incorporate SpiderSQN with different momentum schemes. Moreover, the proposed algorithms are generalized to the online setting, and the corresponding SFO complexity of O(ϵ-3) is developed, which also matches the existing best result. Extensive experiments on benchmark data sets demonstrate that our new algorithms outperform state-of-the-art approaches for nonconvex optimization.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Faster Stochastic Quasi-Newton Methods.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems

Lead the way for us

Journal: IEEE Transactions on Neural Networks and Learning Systems	Publication Date: Sep 1, 2022
Citations: 8

Similar Papers

Asynchronous parallel stochastic Quasi-Newton methods
Qianqian Tong ... Jinbo Bi
Parallel Computing | VOL. 101
Qianqian Tong, et. al.Qianqian Tong ... Jinbo Bi
04 Nov 2020
Parallel Computing | VOL. 101

Quasi-Newton Optimization Methods for Deep Learning Applications
Jacob Rafati ... Roummel F Marica
-
Jacob Rafati, et. al.Jacob Rafati ... Roummel F Marica
01 Jan 2020
01 Jan 2020

Deep Neural Networks Training by Stochastic Quasi-Newton Trust-Region Methods
Mahsa Yousefi ... Ángeles Martínez
Algorithms | VOL. 16
Mahsa Yousefi, et. al.Mahsa Yousefi ... Ángeles Martínez
20 Oct 2023
Algorithms | VOL. 16

An automatic learning rate decay strategy for stochastic gradient descent optimization methods in neural networks
Kang Wang ... Peng Qiao
International Journal of Intelligent Systems | VOL. 37
Kang Wang, et. al.Kang Wang ... Peng Qiao
31 Mar 2022
International Journal of Intelligent Systems | VOL. 37

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Faster Stochastic Quasi-Newton Methods.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems