arXiv ScienceSearch

arXiv · 2410.02626

Online Learning Guided Quasi-Newton Methods with Global Non-Asymptotic Convergence

Abstract

In this paper, we propose a quasi-Newton method for solving smooth and monotone nonlinear equations, including unconstrained minimization and minimax optimization as special cases. For the strongly monotone setting, we establish two global convergence bounds: (i) a linear convergence rate that matches the rate of the celebrated extragradient method, and (ii) an explicit global superlinear convergence rate that provably surpasses the linear convergence rate after at most ${O}(d)$ iterations, where $d$ is the problem's dimension. In addition, for the case where the operator is only monotone, we prove a global convergence rate of ${O}(\min\{{1}/{k},{\sqrt{d}}/{k^{1.25}}\})$ in terms of the duality gap. This matches the rate of the extragradient method when $k = {O}(d^2)$ and is faster when $k = Ω(d^2)$. These results are the first global convergence results to demonstrate a provable advantage of a quasi-Newton method over the extragradient method, without querying the Jacobian of the operator. Unlike classical quasi-Newton methods, we achieve this by using the hybrid proximal extragradient framework and a novel online learning approach for updating the Jacobian approximation matrices. Specifically, guided by the convergence analysis, we formulate the Jacobian approximation update as an online convex optimization problem over non-symmetric matrices, relating the regret of the online problem to the convergence rate of our method. To facilitate efficient implementation, we further develop a tailored online learning algorithm based on an approximate separation oracle, which preserves structures such as symmetry and sparsity in the Jacobian matrices.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ruichen Jiang, Aryan Mokhtari. 2024-10-03. Online Learning Guided Quasi-Newton Methods with Global Non-Asymptotic Convergence. https://arxiv.org/abs/2410.02626

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Policy Iteration for Stationary Discounted Hamilton--Jacobi--Bellman Equations: A Viscosity Approach

We study policy iteration (PI) for deterministic infinite-horizon discounted control problems characterized by stationary Hamilton--Jacobi--Bellman equations. For general viscosity solutions, the classical gradient-based policy improvement step need not be defined pointwise. We introduce a semi-discrete formulation with centered difference quotients at scale $h$ and a separate artificial-viscosity term of order $O(h)$. The resulting stencil is monotone, and the positive discount yields a resolvent contraction. Under bounded Lipschitz data and a globally Lipschitz minimizing policy map, we prove monotone and geometric convergence of the value iterates for each fixed $h>0$, together with a local quadratic estimate whose constant is of order $h^{-2}$. Under the additional condition $λ>\Lip_x(f)$, we establish $\|V^h-V\|_\infty\le C\sqrt h$ and combine the discretization and iteration errors into a quantitative bound. A bounded Lipschitz example shows that the $\sqrt h$ exponent is sharp for this scheme. The combined estimate gives a sufficient iteration count of order $h^{-1}\log(1/h)$ to attain an error of order $\sqrt h$. In bounded-domain experiments, the smooth one-dimensional benchmark exhibits the predicted discretization plateau, while a nonlinear two-dimensional manufactured benchmark isolates convergence to the discrete solution. Exact policy evaluation gives substantially faster local convergence than the global geometric bound. A neural evaluation diagnostic illustrates the importance of controlling boundary errors as well as interior residuals.

math.OC

Inverse Problems for Costs and Controls in LQG MFGs via Mean Field Trajectories

This paper investigates inverse problems for Linear-Quadratic-Gaussian (LQG) Mean Field Games (MFGs) based entirely on the observation of mean-covariance trajectories. We address three sequential challenges: identifying the optimal control for observed initializations, determining the control for arbitrary initializations, and recovering consistent cost parameters. After establishing the existence and uniqueness of the forward Nash equilibrium under mild hypotheses, we analyze the injectivity of the parameter-to-trajectory mapping, demonstrating that it is inherently non-injective and providing sufficient conditions for parameter equivalence. We prove that while the optimal control is locally identifiable for observed initializations under minimal assumptions, global identifiability requires a deeper structural recovery of the game's costs. To bridge this gap, we propose a constructive semidefinite programming method to infer cost parameters that are strictly consistent with the observed population dynamics. Numerical experiments illustrate this method.

math.OC

Adaptive Schauder Stochastic Mirror Descent in Banach Spaces

In this paper, we introduce an adaptive regularization strategy for stochastic mirror descent (SMD) to solve a class of risk functional minimization problems in infinite-dimensional Banach spaces. This regularization strategy centers on using a Schauder basis to construct a nested family of finite-dimensional subspaces, with the dimension chosen adaptively according to the sample size $n$. We then restrict each SMD subproblem to the corresponding subspace and project the stochastic gradient onto its dual space. This yields closed-form solutions to the SMD subproblems and coordinate-wise updates of the basis coefficients, enabling an implementation with low computational and storage complexity. The subspace dimension also serves as a regularization parameter that balances approximation and optimization errors. For risk functional minimization in $\mathcal{L}^p$ spaces with $1<p<\infty$, we construct Bregman distances adapted to the geometry of the underlying Banach spaces using $\max\{2,p\}$-convex functionals induced by their uniform convexity. At the non-uniformly convex $\mathcal{L}^1$ endpoint, we instead construct a locally strongly convex functional based on the entropy function. By developing a new analytical framework, we establish a convergence rate of $\mathcal O\left(n^{-\min\{\frac12,\frac1p\}}\right)$, up to logarithmic factors. In the misspecified setting, where the minimizer satisfies only weaker regularity conditions, we prove that the risk functional still converges to its minimum value. Finally, we apply the method to statistical inverse problems and illustrate its empirical performance through numerical experiments in both settings.

math.OC