arXiv · 2403.08553
Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG
Abstract
Recent advancement in online optimization and control has provided novel tools to study online linear quadratic regulator (LQR) problems, where cost matrices are time-varying and unknown in advance. In this work, we study the online linear quadratic Gaussian (LQG) problem over the manifold of stabilizing controllers that are linearly constrained to impose physical conditions such as sparsity. By adopting a Riemannian perspective, we propose the online Newton on manifold (ONM) algorithm, which generates an online controller on-the-fly based on the second-order information of the cost function sequence. To quantify the algorithm performance, we use the notion of regret, defined as the sub-optimality of the algorithm cumulative cost against a (locally) minimizing controller sequence. We establish a regret bound in terms of the path-length of the benchmark minimizer sequence, and we further verify the effectiveness of ONM via simulations.
Explore related subjects
Keep this discovery
Ting-Jui Chang, Shahin Shahrampour. 2024-03-13. Regret Analysis of Policy Optimization over Submanifolds for Linearly Constrained Online LQG. https://arxiv.org/abs/2403.08553
Cite the original work for its findings. Save a collection to share your selection of sources.