arXiv ScienceSearch

arXiv subjects

Salim Msaad

Publications and source records attributed to Salim Msaad.

3 recordsLinked to original sources

Economic Model Predictive Control with Policy-Guided Terminal Ingredients

Conventional designs for model predictive control typically rely on terminal costs and constraints derived from a steady state to guarantee closed-loop stability and performance. However, this dependence on a steady-state assumption limits the applicability of this control method to systems in which such a fixed operating point is either not available or not desirable. This work introduces a novel framework, termed policy-guided MPC, to address this limitation. Our approach constructs terminal costs and constraints using a known sub-optimal control policy. Specifically, the terminal region is defined around a center determined by a rollout of the policy, and a penalty on deviation from this center is used to define the terminal cost. This method obviates the need for a steady state or reference trajectory. Closed-loop performance guarantees are established relative to the guiding policy, for both finite and infinite horizon problems. The effectiveness of the proposed framework is demonstrated through numerical simulations on an energy management example.

eess.SY

Improving greenhouse fruit-production control by integrating reinforcement learning into short-horizon model predictive control

Greenhouse fruit-production control aims to maximize the economic performance (fruit revenue minus operating costs) while operating within system constraints under external weather disturbances. Control methods need to balance the delayed economic benefit of fruit yield with current operating costs. For such problems, model predictive control (MPC) can explicitly handle system constraints under future weather disturbances, but can become computationally demanding when using sufficiently long prediction horizons for (relatively large) nonlinear greenhouse fruit production models. In contrast, reinforcement learning (RL) can learn control policies offline while considering longer-term economic performance, but struggles to enforce system constraints, and performance may degrade under unseen weather trajectories. This work proposes trajectory-selection RL-MPC, a framework that incorporates longer-term economic information of fruit yield into a short-horizon MPC optimization problem. The framework uses an RL rollout trajectory to define a terminal region constraint and terminal cost. Next, a nonlinear MPC solves a short-horizon optimization problem with these terminal ingredients to find a local optimum. Finally, the framework selects and executes the first input from the trajectory with the better objective value, either from the MPC-predicted or the RL rollout trajectory. The method is applied to GreenLight, a large-scale greenhouse tomato production model that exhibits stiff dynamics. The simulation results show that trajectory-selection RL-MPC with a one-hour prediction horizon matches the closed-loop performance of a high-performing guiding policy while significantly improving over standalone MPC with the same horizon.

math.OC

RL-Guided MPC for Autonomous Greenhouse Control

The efficient operation of greenhouses is essential for enhancing crop yield while minimizing energy costs. This paper investigates a control strategy that integrates Reinforcement Learning (RL) and Model Predictive Control (MPC) to optimize economic benefits in autonomous greenhouses. Previous research has explored the use of RL and MPC for greenhouse control individually, or by using MPC as the function approximator for the RL agent. This study introduces the RL-Guided MPC framework, where a RL policy is trained and then used to construct a terminal cost and terminal region constraint for the MPC optimization problem. This approach leverages the ability to handle uncertainties of RL with MPC's online optimization to improve overall control performance. The RL-Guided MPC framework is compared with both MPC and RL via numerical simulations. Two scenarios are considered: a deterministic environment and an uncertain environment. Simulation results demonstrate that, in both environments, RL-Guided MPC outperforms both RL and MPC with shorter prediction horizons.

eess.SY