arXiv Science⌕ Search

arXiv · 2610.09220

Regret-Optimal Linear-Quadratic Control under Distributional Drift

Abstract

We study finite-horizon LQR with independent disturbances whose laws may drift over time. We address uncertainty in this evolution through ex-ante distributionally robust regret optimization (DRRO), which minimizes worst-case excess expected cost relative to the optimal causal controller that knows the law sequence. To specify the admissible law sequences, we introduce a temporally coupled moment ambiguity set with a Gelbrich kinetic-energy budget and a second budget limiting departures from reference moments and persistent cumulative mismatch. Despite the coupling of moments across time, affine disturbance-feedback synthesis reduces exactly to a semidefinite program. An optimal policy uses past disturbances to predict the correction to nominal LQR determined by current and future disturbance means. The SDP solution also yields worst-case moment paths and corresponding laws. For standard distributionally robust optimization (DRO), the analogous SDP relaxation can have a duality gap. Portfolio liquidation experiments show substantially lower worst-case regret over the drifting ambiguity set than common-law DRRO and nominal control.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Lukas-Benedikt Fiechtner, Jose Blanchet. 2026-10-06. Regret-Optimal Linear-Quadratic Control under Distributional Drift. https://arxiv.org/abs/2610.09220

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Constrained portfolio game with heterogeneous agents

We investigate stochastic utility maximization games under relative performance concerns in both finite-agent and infinite-agent (graphon) settings. An incomplete market model is considered where agents with power (CRRA) utility functions trade in a common risk-free bond and individual stocks driven by both common and idiosyncratic noise. The Nash equilibrium for both settings is characterized by forward-backward stochastic differential equations (FBSDEs) with a quadratic growth generator, where the solution of the graphon game leads to a novel form of infinite-dimensional McKean-Vlasov FBSDEs. Under mild conditions, we prove the existence of Nash equilibrium for both the graphon game and the $n$-agent game without common noise. Furthermore, we establish a convergence result showing that, with modest assumptions on the sensitivity matrix, as the number of agents increases, the Nash equilibrium and associated equilibrium value of the finite-agent game converge to those of the graphon game.

math.OC↗

A Model-Based Derivative-Free Optimization Algorithm for Partially Separable Problems

We propose UPOQA, a derivative-free optimization algorithm for partially separable unconstrained problems, leveraging quadratic interpolation and a structured trust-region framework. By decomposing the objective into element functions, UPOQA constructs underdetermined element models and solves subproblems efficiently via a modified projected gradient method. Innovations include an approximate projection operator for structured trust regions, improved management of elemental radii and models, a starting point search mechanism, and support for hybrid black-white-box optimization, etc. Numerical experiments on 85 CUTEst problems demonstrate that \texttt{UPOQA} can significantly reduce the number of function evaluations. To quantify the impact of exploiting partial separability, we introduce the speed-up profile to further evaluate the acceleration effect. Results show that the speed-up of UPOQA over baselines is less significant in low-precision scenarios but becomes more pronounced in high-precision scenarios. Applications to quantum variational problems further validate its practical utility.

math.OC↗

Convergence Analysis of Noisy Distributed Gradient Descent for Non-convex Optimization -- Saddle Point Escape

This paper studies noisy distributed gradient descent (\textbf{NDGD}) for smooth non-convex finite-sum optimization over networks. Random perturbations enable saddle-point escape while preserving distributed implementation and consensus. Under suitable regularity conditions, \textbf{NDGD} converges with high probability to a neighborhood of a common local minimizer. Its convergence complexity is comparable to centralized first-order saddle-point escape methods, reducing exponential dependence on problem dimension to polynomial dependence. Numerical experiments demonstrate improved saddle-point escape over standard \textbf{DGD}.

math.OC↗