arXiv ScienceSearch

arXiv subjects

Mengmou Li

Publications and source records attributed to Mengmou Li.

17 recordsLinked to original sources

Interaction-Limited Safe Continuous-Time RL for Dynamical Medical Treatment

Dynamic medical treatment requires deciding treatment intensity and intervention timing, while patient states evolve continuously and adverse events may occur between clinical interactions. Most existing treatment learning methods assume fixed schedules or enforce safety only at discrete decision points. We propose Interaction-Limited Safe Continuous-Time Reinforcement Learning, a framework that jointly optimizes treatment administration and clinical interaction timing under trajectory-level safety constraints. Our key idea is to reformulate the continuous time treatment problem as an option-based semi-Markov decision process, where each option specifies a continuous-time treatment policy and its duration. We develop a safety-tightening mechanism showing that suitably constructed constraints at interaction times guarantee safety over the full continuous-time trajectory with high probability. We further establish finite-sample guarantees for policy learning from logged treatment trajectories and introduce a practical data-driven conservative surrogate. Experiments show that the proposed adaptive interaction-timing mechanism improves both safety and treatment effectiveness over equidistant interaction schemes across different safe policy optimization methods.

cs.LG

A Common Lyapunov Matrix Approach to the Exponential Stability of Augmented Primal-Dual Gradient Flow as LPV Systems

We show that a common Lyapunov matrix exists for the convex combination of two Hurwitz matrices if and only if the intersection of the set of strict Lyapunov matrices for one matrix and the set of non-strict Lyapunov matrices for the other is nonempty. This simple relaxation is useful for the convergence analysis of the augmented primal-dual gradient flow for constrained optimization problems with affine inequality constraints, which can be viewed as a polytopic linear parameter-varying (LPV) system driven by the active-constraint selector. Under a relaxed strong convexity condition, exponential convergence is proved for the LPV system. The analysis can further be extended to the integral quadratic constraints (IQCs) framework for LPV systems to facilitate numerical search of the convergence rate.

eess.SY

Finite-Time Optimization via Scaled Gradient-Momentum Flows

In this paper, we develop a scaled gradient-momentum framework for continuous-time optimization that achieves global finite-time convergence. A state-dependent scaling mechanism is introduced to enable classical dynamics, such as Heavy-Ball-type and proportional-integral (PI)-type flows, to attain finite-time convergence. We establish explicit conditions that bridge the gradient-dominance property of the objective function and finite-time stability of the proposed scaled dynamics. Numerical experiments validate the theoretical results.

math.OC

A Canonical Structure for Constructing Projected First-Order Algorithms With Delayed Feedback

This work introduces a canonical structure for a broad class of unconstrained first-order algorithms that admit a Lur'e representation, including systems with relative degree greater than one, e.g., systems with delayed gradient feedback. The proposed canonical structure is obtained through a simple linear transformation. It enables a direct extension from unconstrained optimization algorithms to set-constrained ones through projection in a Lyapunov-induced norm. The resulting projected algorithms attain the optimal solution while preserving the convergence rates of their unconstrained counterparts.

math.OC

First-Order Projected Algorithms With the Same Linear Convergence Rate Bounds as Their Unconstrained Counterparts

In this paper, we propose a systematic approach for extending first-order optimization algorithms, originally designed for unconstrained strongly convex problems, to handle closed convex set constraints. We show that the resulting projected algorithms retain the same linear convergence rate bounds, provided that the underlying unconstrained optimization algorithms admit a quadratic Lyapunov function obtained from integral quadratic constraint (IQC) analysis. The projected algorithms are constructed by applying a projection in the norm induced by the Lyapunov matrix, ensuring both constraint satisfaction and optimality at the fixed point. Furthermore, under a linear transformation associated with this matrix, the projection becomes non-expansive in the Euclidean norm, thereby preserving the convergence rate bounds under the composition of the linearly convergent algorithmic operator and the projection. Our results indicate that, when analyzing worst-case convergence rates or when synthesizing first-order optimization algorithms with potentially higher-order dynamics, it suffices to focus solely on the unconstrained dynamics, since the same parameters or stepsizes can be employed without retuning.

math.OC

Exponential Convergence of Augmented Primal-dual Gradient Algorithms for Partially Strongly Convex Functions

We show that the augmented primal-dual gradient algorithms can achieve global exponential convergence with partially strongly convex functions. In particular, the objective function only needs to be strongly convex in the subspace satisfying the equality constraint and can be generally convex elsewhere, provided the global Lipschitz condition for the gradient is satisfied. This condition implies that states outside the equality subspace will converge towards it exponentially fast. The analysis is then applied to distributed optimization, where the partially strong convexity can be relaxed to the restricted secant inequality condition, which is not necessarily convex. This work unifies global exponential convergence results for some existing centralized and distributed algorithms.

math.OC

Small-Gain Theorem Based Distributed Prescribed-Time Convex Optimization For Networked Euler-Lagrange Systems

In this paper, we address the distributed prescribed-time convex optimization (DPTCO) for a class of networked Euler-Lagrange systems under undirected connected graphs. By utilizing position-dependent measured gradient value of local objective function and local information interactions among neighboring agents, a set of auxiliary systems is constructed to cooperatively seek the optimal solution. The DPTCO problem is then converted to the prescribed-time stabilization problem of an interconnected error system. A prescribed-time small-gain criterion is proposed to characterize prescribed-time stabilization of the system, offering a novel approach that enhances the effectiveness beyond existing asymptotic or finite-time stabilization of an interconnected system. Under the criterion and auxiliary systems, innovative adaptive prescribed-time local tracking controllers are designed for subsystems. The prescribed-time convergence lies in the introduction of time-varying gains which increase to infinity as time tends to the prescribed time. Lyapunov function together with prescribed-time mapping are used to prove the prescribed-time stability of closed-loop system as well as the boundedness of internal signals. Finally, theoretical results are verified by one numerical example.

math.OC

Net-Zero Energy House-oriented Linear Programming for the Sizing Problem of Photovoltaic Panels and Batteries

The global drive towards carbon neutrality has led to a significant increase in the number of power plants based on renewable energy sources (RES). Concurrently, numerous households are adopting RES to generate their own energy, aiming to decrease both electricity costs and carbon footprints. To support these users, many papers have been devoted to developing optimal investment strategies for residential energy systems. However, there is still a significant gap as these studies often neglect important aspects like carbon neutrality. For this reason, in this paper, we explore the concept of net-zero energy houses (ZEHs) -- houses designed to have an annual net energy consumption around zero -- by presenting a constrained optimization problem to find the optimal number of photovoltaic panels and the optimal size of the battery system for home integration. Solving this constrained optimization problem is difficult due to its nonconvex constraints. Nevertheless, by applying a series of transformations, we reveal that it is possible to find an equivalent linear programming (LP) problem which is computationally tractable. The attainment of ZEH can be tackled by introducing a single constraint in the optimization problem. Additionally, we propose a sharing economy approach to the investment problem, offering a strategy that could potentially reduce investment costs and facilitate the attainment of ZEH more efficiently. Finally, we apply the proposed frameworks to a neighborhood in Japan as a case study, demonstrating the potential for long-term ZEH attainment. The results show that, under the right incentive, users can achieve ZEH, reduce their electricity costs and have a minimal impact on the main grid.

math.OC

Stochastic Optimal Investment Strategy for Net-Zero Energy Houses

In this research, we investigate Net-Zero Energy Houses (ZEH), which harness regionally produced electricity from photovoltaic(PV) panels and fuel cells, integrating them into a local power system in pursuit of achieving carbon neutrality. This paper examines the impact of electricity sharing among users who are working towards attaining ZEH status through the integration of PV panels and battery storage devices. We propose two potential scenarios: the first assumes that all users individually invest in storage devices, hence minimizing their costs on a local level without energy sharing; the second envisions cost minimization through the collective use of a shared storage device, managed by a central manager. These two scenarios are formulated as a stochastic convex optimization and a cooperative game, respectively. To tackle the stochastic challenges posed by multiple random variables, we apply the Monte Carlo sample average approximation (SAA) to the problems. To demonstrate the practical applicability of these models, we implement the proposed scenarios in the Jono neighborhood in Kitakyushu, Japan.

econ.GN

Convergence Rate Bounds for the Mirror Descent Method: IQCs, Popov Criterion and Bregman Divergence

This paper presents a comprehensive convergence analysis for the mirror descent (MD) method, a widely used algorithm in convex optimization. The key feature of this algorithm is that it provides a generalization of classical gradient-based methods via the use of generalized distance-like functions, which are formulated using the Bregman divergence. Establishing convergence rate bounds for this algorithm is in general a non-trivial problem due to the lack of monotonicity properties in the composite nonlinearities involved. In this paper, we show that the Bregman divergence from the optimal solution, which is commonly used as a Lyapunov function for this algorithm, is a special case of Lyapunov functions that follow when the Popov criterion is applied to an appropriate reformulation of the MD dynamics. This is then used as a basis to construct an integral quadratic constraint (IQC) framework through which convergence rate bounds with reduced conservatism can be deduced. We also illustrate via examples that the convergence rate bounds derived can be tight.

math.OC

Distributed Optimal Secondary Frequency Control in Power Networks with Delay Independent Stability

Distributed secondary frequency control for power systems, is a problem that has been extensively studied in the literature, and one of its key features is that an additional communication network is required to achieve optimal power allocation. Therefore, being able to provide stability guarantees in the presence of communication delays is an important requirement. Primal-dual and distributed averaging proportional-integral (DAPI) protocols, respectively, are two main control schemes that have been proposed in the literature. Each has its own relative merits, with the former allowing to incorporate general cost functions and additional operational constraints, and the latter being more straightforward in its implementation. Although delays have been addressed in DAPI schemes, there are currently no theoretical guarantees for the stability of primal-dual schemes for frequency control, when these are subject to communication delays. In fact, simulations illustrate that even small delays can destabilize such schemes. In this paper, we show how a novel formulation of prima-dual schemes allows to construct a distributed algorithm with delay independent stability guarantees. We also show that this algorithm can incorporate many of the key features of these schemes such as tie-line power flow requirements, generation constraints, and the relaxation of demand measurements with an observer layer. Finally, we illustrate our results through simulations on a 5-bus example and on the IEEE-39 test system.

math.OC

Convergence Rate Bounds for the Mirror Descent Method: IQCs and the Bregman Divergence

This paper is concerned with convergence analysis for the mirror descent (MD) method, a well-known algorithm in convex optimization. An analysis framework via integral quadratic constraints (IQCs) is constructed to analyze the convergence rate of the MD method with strongly convex objective functions in both continuous-time and discrete-time. We formulate the problem of finding convergence rates of the MD algorithms into feasibility problems of linear matrix inequalities (LMIs) in both schemes. In particular, in continuous-time, we show that the Bregman divergence function, which is commonly used as a Lyapunov function for this algorithm, is a special case of the class of Lyapunov functions associated with the Popov criterion, when the latter is applied to an appropriate reformulation of the problem. Thus, applying the Popov criterion and its combination with other IQCs, can lead to convergence rate bounds with reduced conservatism. We also illustrate via examples that the convergence rate bounds derived can be tight.

math.OC

Parallel Feedforward Compensation for Output Synchronization: Fully Distributed Control and Indefinite Laplacian

This work is associated with the use of parallel feedforward compensators (PFCs) for the problem of output synchronization over heterogeneous agents and the benefits this approach can provide. Specifically, it addresses the addition of stable PFCs on agents that interact with each other using diffusive couplings. The value in the application of such PFC is twofold. Firstly, it has been an issue that output synchronization among passivity-short systems requires global information for the design of controllers in the cases when initial conditions need to be taken into account, such as average consensus and distributed optimization. We show that a stable PFC can be designed to passivate a passivity-short system while its output asymptotically vanishes as its input tends to zero. As a result, output synchronization is achieved among these systems by fully distributed controls without altering the original consensus results. Secondly, in the literature of output synchronization over signed weighted graphs, it is generally required that the graph Laplacian be positive semidefinite, i.e., $L \geq 0$ for undirected graphs or $L + L^T \geq 0$ for balanced directed graphs. We show that the PFC serves as output feedback to the communication graph to enhance the robustness against negative weight edges. As a result, output synchronization is achieved over a signed weighted and balanced graph, even if the corresponding Laplacian is not positive semidefinite.

eess.SY

Distributed Optimization With Event-triggered Communication via Input Feedforward Passivity

In this work, we address the distributed optimization problem with event-triggered communication by the notion of input feedforward passivity (IFP). First, we analyze the distributed continuous-time algorithm over uniformly jointly strongly connected balanced digraphs in an IFP-based framework. Then, we propose a distributed event-triggered communication mechanism for this algorithm. Next, we discretize the continuous-time algorithm by the forward Euler method with a constant stepsize irrelevant to network size, and show that the discretization can be seen as a stepsize-dependent passivity degradation of the input feedforward passivity. Thus, the discretized system preserves the IFP property and enables the same event-triggered communication mechanism but without Zeno behavior due to the discrete-time nature. Finally, a numerical example is presented to illustrate our results.

math.OC

Smooth Dynamics for Distributed Constrained Optimization with Heterogeneous Delays

This work investigates the distributed constrained optimization problem under inter-agent communication delays from the perspective of passivity. First, we propose a continuous-time algorithm for distributed constrained optimization with general convex objective functions. The asymptotic stability under general convexity is guaranteed by the phase lead compensation. The inequality constraints are handled by adopting a projection-free generalized Lagrangian, whose primal-dual gradient dynamics preserves passivity and smoothness, enabling the application of the LaSalle's invariance principle in the presence of delays. Then, we incorporate the scattering transformation into the proposed algorithm to enhance the robustness against unknown and heterogeneous communication delays. Finally, a numerical example of a matching problem is provided to illustrate the results.

math.OC

Distributed Resource Allocation over Time-varying Balanced Digraphs with Discrete-time Communication

This work is concerned with the problem of distributed resource allocation in continuous-time setting but with discrete-time communication over infinitely jointly connected and balanced digraphs. We provide a passivity-based perspective for the continuous-time algorithm, based on which an intermittent communication scheme is developed. Particularly, a periodic communication scheme is first derived through analyzing the passivity degradation over output sampling of the distributed dynamics at each node. Then, an asynchronous distributed event-triggered scheme is further developed. The sampled-based event-triggered communication scheme is exempt from Zeno behavior as the minimum inter-event time is lower bounded by the sampling period. The parameters in the proposed algorithm rely only on local information of each individual nodes, which can be designed in a truly distributed fashion

cs.MA

Input-Feedforward-Passivity-Based Distributed Optimization Over Jointly Connected Balanced Digraphs

In this paper, a distributed optimization problem is investigated via input feedforward passivity. First, an input-feedforward-passivity-based continuous-time distributed algorithm is proposed. It is shown that the error system of the proposed algorithm can be decomposed into a group of individual input feedforward passive (IFP) systems that interact with each other using output feedback information. Based on this IFP framework, convergence conditions of a suitable coupling gain are derived over weight-balanced and uniformly jointly strongly connected (UJSC) topologies. It is also shown that the IFP-based algorithm converges exponentially when the topology is strongly connected. Second, a novel distributed derivative feedback algorithm is proposed based on the passivation of IFP systems. While most works on directed topologies require knowledge of eigenvalues of the graph Laplacian, the derivative feedback algorithm is fully distributed, namely, it is robust against randomly changing weight-balanced digraphs with any positive coupling gain and without knowing any global information. Finally, numerical examples are presented to illustrate the proposed distributed algorithms.

math.OC