arXiv · 2609.09152
Silver Rate Is (Almost) Optimal for Gradient Descent
Abstract
We study how far gradient descent (GD) can be accelerated by predetermined stepsizes in smooth convex optimization. Writing $p_{\mathrm{sil}}=\log_2(1+\sqrt{2})$, we prove an $\Omega\left(n^{-p_{\mathrm{sil}}-O(\sqrt{\log\log n/\log n})}\right)$ non-anytime lower bound. In the anytime setting, every infinite schedule has infinitely many horizons with error $\Omega\left(n^{-\frac{2p_{\mathrm{sil}}}{1+p_{\mathrm{sil}}}-O(\sqrt{\log\log n/\log n})}\right)$. Together with the silver-schedule upper bound [Altschuler and Parrilo, 2025] and the anytime upper bound [Zhang et al., 2025], our results determine the optimal polynomial convergence exponents in both settings.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yuhan Ye, Kaizhao Liu. 2026-09-08. Silver Rate Is (Almost) Optimal for Gradient Descent. https://arxiv.org/abs/2609.09152
Cite the original work for its findings. Save a collection to share your selection of sources.