arXiv · 2211.15596
A survey of deep learning optimizers -- first and second order methods
Abstract
Deep Learning optimization involves minimizing a high-dimensional loss function in the weight space which is often perceived as difficult due to its inherent difficulties such as saddle points, local minima, ill-conditioning of the Hessian and limited compute resources. In this paper, we provide a comprehensive review of $14$ standard optimization methods successfully used in deep learning research and a theoretical assessment of the difficulties in numerical optimization from the optimization literature.
Explore related subjects
Keep this discovery
Rohan Kashyap. 2022-11-28. A survey of deep learning optimizers -- first and second order methods. https://arxiv.org/abs/2211.15596
Cite the original work for its findings. Save a collection to share your selection of sources.