arXiv · 2306.06613
Parameter-free version of Adaptive Gradient Methods for Strongly-Convex Functions
Abstract
The optimal learning rate for adaptive gradient methods applied to {\lambda}-strongly convex functions relies on the parameters {\lambda} and learning rate {\eta}. In this paper, we adapt a universal algorithm along the lines of Metagrad, to get rid of this dependence on {\lambda} and {\eta}. The main idea is to concurrently run multiple experts and combine their predictions to a master algorithm. This master enjoys O(d log T) regret bounds.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Deepak Gouda, Hassan Naveed, Salil Kamath. 2023-06-11. Parameter-free version of Adaptive Gradient Methods for Strongly-Convex Functions. https://arxiv.org/abs/2306.06613
Cite the original work for its findings. Save a collection to share your selection of sources.