arXiv · 2604.04090
Fine-grained Analysis of Stability and Generalization for Stochastic Bilevel Optimization
Abstract
Stochastic bilevel optimization (SBO) has been integrated into many machine learning paradigms recently, including hyperparameter optimization, meta learning, and reinforcement learning. Along with the wide range of applications, there have been numerous studies on the computational behavior of SBO. However, the generalization guarantees of SBO methods are far less understood from the lens of statistical learning theory. In this paper, we provide a systematic generalization analysis of the first-order gradient-based bilevel optimization methods. Firstly, we establish the quantitative connections between the on-average argument stability and the generalization gap of SBO methods. Then, we derive the upper bounds of on-average argument stability for single-timescale stochastic gradient descent (SGD) and two-timescale SGD, where three settings (nonconvex-nonconvex (NC-NC), convex-convex (C-C), and strongly-convex-strongly-convex (SC-SC)) are considered respectively. Experimental analysis validates our theoretical findings. Compared with the previous algorithmic stability analysis, our results do not require reinitializing the inner-level parameters at each iteration and are applicable to more general objective functions.
Explore related subjects
Keep this discovery
Xuelin Zhang, Hong Chen, Bin Gu, Tieliang Gong, Feng Zheng. 2026-04-05. Fine-grained Analysis of Stability and Generalization for Stochastic Bilevel Optimization. https://doi.org/10.24963/ijcai.2024/609
Cite the original work for its findings. Save a collection to share your selection of sources.