arXiv ScienceSearch

arXiv · 2506.10506

On a mean-field Pontryagin minimum principle for stochastic optimal control

Abstract

This paper outlines a novel extension of the classical Pontryagin minimum (maximum) principle to stochastic optimal control problems. Contrary to the well-known stochastic Pontryagin minimum principle involving forward-backward stochastic differential equations, the proposed formulation is deterministic and of mean-field type. We denote it by the McKean-Pontryagin minimum principle. The Hamiltonian structure of the proposed McKean-Pontryagin minimum principle is achieved via the introduction of a pair of auxiliary functions. A gauge freedom in the choice of one of these two functions can be used to decouple the forward and reverse time equations; hence simplifying the solution of the underlying boundary value problem. We also consider infinite horizon discounted cost optimal control problems. In this case, the mean-field formulation allows one to convert the computation of the desired optimal control law into solving a pair of forward mean-field ordinary differential equations. The McKean-Pontryagin minimum principle is tested numerically for a controlled diffusion process in a double well, a controlled inverted pendulum, a controlled Lorenz-63 system, and a controlled Lorenz-96 system. Although the focus is on linear-quadratic control problems, the proposed methodology is extendable to more general problems including mean-field type control formulations.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Manfred Opper, Sebastian Reich. 2026-09-17. On a mean-field Pontryagin minimum principle for stochastic optimal control. https://arxiv.org/abs/2506.10506

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Strong Duality in Risk-Constrained Nonconvex Functional Programming

We show that a wide class of risk-constrained nonconvex functional optimization problems exhibit strong duality, regardless of nonconvexity. We develop two novel results under distinct sets of assumptions, establishing strong duality over both decomposable policies (matching and extending prior work in the risk neutral case) and nondecomposable policies with structure (e.g., continuity or smoothness), including certain universal finite-dimensional (fixed depth/width) neural network parametrizations as special cases (improving established results in the risk-neutral setting as well). Our results hold for constrained models featuring arbitrary convex risk functionals on $L_p,p\in[1,\infty]$. We further discuss reductions and generalizations of our base model, and establish its necessity/minimality by presenting explicitly constructed counterexamples, rigorously refuting the manifestation of strong duality for certain, potentially simpler optimization models. Lastly, we discuss applications in wireless systems resource allocation, and supervised constrained learning. Our core proof technique appears to be new and relies on a non-trivial application of the Kingman-Robertson generalization of Lyapunov's convexity theorem for vector measures taking values in infinite-dimensional topological spaces.

math.OC

A bin packing formulation for the joint batching, routing and sequencing problem

Warehouses are the scene of complex logistic problems integrating different decision layers. This paper addresses the Joint Order Batching, Picker Routing and Sequencing Problem with Deadlines (JOBPRSP-D) in rectangular warehouses. To tackle the problem, an exponential linear programming formulation related to bin packing is proposed. It is solved with a column generation heuristic able to provide lower and upper bounds on the optimal value. We start by showing that the JOBPRSP-D is related to a bin packing problem as opposed to a scheduling problem. We take advantage of this aspect to derive a number of valid inequalities that enhance the resolution of the master problem. The proposed algorithm is evaluated on publicly available data-sets. It is able to optimally solve instances with up to 18 orders in a few minutes. It is also able to prove optimality or to provide high-quality lower bounds on larger instances with 100 orders as long as the picker's capacity is small enough. Finally, new upper bounds are found on all the instances except two of the benchmark considered. To the best of our knowledge this is the first article that provides optimality guarantee on large size instances for the JOBPRSP-D; the results can therefore be used to assert the quality of heuristics proposed for the same problem.

math.OC

The Maximum Singularity Degree for Linear and Semidefinite Programming

Facial reduction (FR) is an important tool in linear and semidefinite programming, providing both algorithmic and theoretical insights into these problems. The maximum length of an FR sequence for a convex set is referred to as the maximum singularity degree (MSD). The MSD gives a choice-robust worst-case bound on the number of nontrivial FR steps. It also yields a sufficient rank for a low-rank formulation of the SDP exposing-vector search, while upper bounds on the MSD can serve as proof devices for bounding the singularity degree in structured problem classes. These concrete roles motivate our study of its fundamental properties. In this work, we show that if an FR sequence has the longest length, then it satisfies a certain minimal property. For linear programming (LP), we prove that every minimal FR sequence forms a basis of a fixed vector space. This yields a direct characterization of the longest FR sequences. To study the MSD for semidefinite programming (SDP), we provide several useful tools including simplification and upper-bounding techniques. By leveraging these tools and the characterization for LP problems, we prove that finding a longest FR sequence for SDP problems is NP-hard. This complexity result highlights a striking difference between the shortest and the longest FR sequences for SDP problems.

math.OC