arXiv ScienceSearch

arXiv subjects

Peter Benner

Publications and source records attributed to Peter Benner.

At least 19 recordsLinked to original sources

Tensor-based Approximation of Molecular Kinetics: Generator Learning, Reaction Coordinates and Incremental Updating

We present an approach to analyze long-timescale kinetics of molecular dynamics simulations - meta-stable states and transition timescales - by a tensor-based approximation of the infinitesimal generator. We start from the variational approximation of low-lying generator eigenvalues, and discretize the associated eigenvalue problem using a tensor-product basis. We derive analytical expressions for the resulting tensor operators in tensor train (TT) format, and provide efficient algorithms to solve the corresponding linear problems. We also treat the case of a state-dependent diffusion field, which is relevant when working in reaction coordinates. We show how the number of reaction coordinates can be gradually increased without having to solve the variational problem from scratch. Using MD simulations of fast-folding proteins and TICA coordinates as reaction coordinates, we demonstrate the effectiveness of the proposed method at identifying long-timescale transitions and meta-stable states.

physics.chem-ph

Boundary control and periodic trajectories of semilinear Euler equations with application to hydrogen transport

A mathematical model of gas flow in a pipeline controlled by the inlet pressure and the outlet mass flow flux is considered in the form of the isothermal Euler equations with an appropriate equation of state. Within the framework of boundary control systems, this model is transformed into a nonlinear abstract differential equation using an appropriate lifting operator. For this abstract equation, a representation in terms of Fourier coefficients is derived analytically. Conditions for the existence of periodic solutions to a broad class of nonlinear control systems with continuously differentiable controls are established in abstract spaces. This framework is applied to a realistic hydrogen transport model to evaluate periodic operating regimes under periodic fluctuations in supply and demand.

math.OC

Partial Observation of Linear Systems with the Mori-Zwanzig Formalism

The Mori-Zwanzig formalism provides a systematic framework for deriving reduced-order model of dynamical systems when only part of the state is observed, but its practical use is often limited by the complexity of the resulting computations. This paper develops an explicit formulation of the Mori-Zwanzig equation for linear time-invariant systems under partially observed observables. By expressing the dynamics in terms of observables, the Koopman generator, and projections onto resolved and unresolved components, we derive closed-form representations of the Markovian, noise, and memory contributions that arise in the Mori-Zwanzig identity. For the linear setting, the resulting formulas recover the reduced dynamics obtained from the variation-of-constants formula while retaining the operator-based structure of the Mori-Zwanzig approach. This makes the derivation a transparent reference case for reduced-order modelling with memory and clarifies how unresolved variables influence the observed dynamics through history-dependent terms. The analysis also identifies the ingredients needed for extensions to nonlinear systems and more general projections, including spectral filtering and data-driven approximations of memory effects. Analytical and numerical examples involving the harmonic oscillator and wave equations illustrate the construction and demonstrate how the formalism can be used to obtain interpretable reduced-order models for partially observed systems.

math.DS

Low Precision Fortran -- Enabling Low Precision Floating Point Arithmetic in Modern Fortran

Although Fortran is almost 70 years old, the language continues to evolve in order to keep pace with developments in computer science. In particular, a flexible type system was introduced that allows developers to specify the sizes of floating-point numbers and integers. In the latest revisions of the Fortran standard, portable type variants for IEEE 754 binary64 (double precision, real64) and binary32 (single precision, real32) were added. However, the rapid development of AI toolkits and accelerator hardware has created a strong focus on floating-point types of lower precision and lower memory usage than binary32. While the IEEE 754-2019 standard defines the binary16 type for representing half-precision numbers, the Fortran standard does not provide the real16 variant in the type system. In contrast, most C compilers support such a data type. In numerical linear algebra, there is strong interest in exploiting the high performance of accelerator devices for core algorithms like matrix decompositions or iterative solvers. Especially when the performance ratio between double, single, and half precision is on the order of 1:2:20, as on current NVidia H100 accelerators, it becomes highly beneficial to use lower-precision types. Yet, before performance can be targeted, correctness and accuracy must be verified when operating below single precision. In this article, we present our Low Precision Fortran (LPF) library that enables the use of low-precision types -- binary16, bfloat16, fp8_e4m3, and fp8_e5m2 -- just like any other floating-point type in Fortran. Furthermore, we introduce extensions that support BLAS operations in low precision and show how easily existing routines can be rewritten to use these data types.

cs.MS

An approach to encode divergence-free stress fields in neural approximations based on stress potentials

The purpose of the current work is the development of an approach to account for quasi-static mechanical equilibrium in empirical (i.e., data-based) models for the stress field employing neural approximations (NAs), which include neural networks (NNs) and neural operators (NOs), in particular Fourier NOs (FNOs). Rather than including such constraints from physics in the loss function as done in the (now standard) physics-informed approach, the current approach incorporates or "encodes" such constraints directly into the architecture of the NA. As a result, both NA training and output are physically constrained in the physics-encoded approach, in contrast to the physics-informed approach, in which only training is physically constrained. For the current constraint of divergence-free stress, a novel encoding approach based on a stress potential is proposed. As a "proof-of-concept" example application of the current approach, a physics-encoded FNO (PeFNO) is developed for a heterogeneous polycrystalline material consisting of isotropic elastic grains and subject to uniaxial extension. Stress field data for this purpose are obtained from the numerical solution of corresponding boundary-value problems for quasi-static mechanical equilibrium. For comparison with the PeFNO, this data is also employed to develop an analogous physics-guided FNO (PgFNO) and physics-informed FNO (PiFNO). As expected theoretically, and confirmed by this computational comparison, for comparable accuracy of the stress field itself as compared to the data, the stress field output by the trained and tested PeFNO is significantly more accurate in satisfying mechanical equilibrium than the output of either the PgFNO or the PiFNO.

cs.CE

Exponential Decay for a Boundary-Controlled Nonlinear Parabolic Reactor Model

We study an axial dispersion tubular reactor model governed by a nonlinear parabolic equation with Robin-type boundary conditions and boundary feedback control. We derive sufficient conditions for the exponential stability of the steady-state solution of the closed-loop system and provide an explicit estimate of the decay rate. In addition, numerical simulations are presented to illustrate the sharpness of the obtained decay rate for different choices of the feedback gain parameter.

math.AP

Periodic solutions of nonlinear control systems with switching: a Lie-algebraic and contraction approach

This paper is devoted to the analysis of periodic solutions of nonlinear control-affine systems with bang-bang controls. Such problems naturally arise in periodic optimal control with constrained inputs, which have, in particular, important applications in the performance optimization of chemical reactions. We reduce the problem of constructing a periodic solution to that of finding a fixed point of a composition of exponential maps. The latter problem is then addressed using the Baker-Campbell-Hausdorff-Dynkin (BCHD) formula. We establish the equivalence between periodic solutions of the original control system and those of an associated autonomous system involving iterated Lie brackets. Applying incremental stability arguments allows us to further simplify the problem to finding the equilibria of this autonomous system. The developed theory is then applied to nonlinear chemical reaction models with constrained controls.

math.OC

Reduced rank extrapolation for multi-term Sylvester equations

We investigate the acceleration of stationary iterations for multi-term Sylvester equation by means of reduced rank extrapolation (RRE). Theoretical convergence results and implementations are provided for both small and large-scale problems. For the large-scale problems, an inexact non-stationary iteration is discussed, which makes use of low-rank matrix approximations. Numerical experiments illustrate the potential of the RRE acceleration which often leads to a substantial gain in convergence speed and therefore reducing the consumption of storage and computing time.

math.NA

Structure-preserving Krylov Subspace Approximations for the Matrix Exponential of Hamiltonian Matrices: A Comparative Study

We study structure-preserving Krylov subspace methods for approximating the matrix-vector products f(H)b, where H is a large Hamiltonian matrix and f denotes either the matrix exponential or the related phi-function. Such computations are central to exponential integrators for Hamiltonian systems. Standard Krylov methods generally destroy the Hamiltonian structure under projection, motivating the use of Krylov bases with J-orthogonal columns that yield Hamiltonian projected matrices and symplectic reduced exponentials. We compare several such structure-preserving Krylov methods on representative Hamiltonian test problems, focusing on accuracy, efficiency, and structure preservation, and briefly discuss adaptive strategies for selecting the Krylov subspace dimension.

math.NA

A Local Discontinuous Galerkin Method for Dirichlet Boundary Control Problems

In this paper, we consider control constrained $L^2-$Dirichlet boundary control of a convection-diffusion equation on a two dimensional convex polygonal domain. We discretize the control problem based on the local discontinuous Galerkin method with piecewise linear ansatz functions for the flux and potential. We derive a priori error estimates for the full as well as for the variational discrete control approximation. We present a selection of numerical results to demonstrate the performance of our approach and to underpin the theoretical findings.

math.OC

Reduced-Order Inference with Structure-Preserving Parametrization for Bending and Rotating Systems

Mechanical systems are often characterized only by their response to certain loads known from experiments or simulations. The obtained data can be used for various purposes: system analysis, design of mathematical models, or construction of reduced-order models for further simulations under different loading conditions. The use of data for reduced-order modeling is an important developing research direction, especially when the high-dimensional system operators are unknown and their low-dimensional approximation is required for accurate but fast simulations. Our goal is to obtain the low-dimensional surrogate model from the available input signal and deformation trajectory data, capturing the correct system behavior for the basic deformation cases, namely bending and rotation, which are present in almost every complex mechanical system. In this work, we propose a methodology to infer the system operators by solving a nonlinear unconstrained optimization problem. The methodology is based on the operator inference approach for second-order systems and includes a parametrization of the unknown operators that preserves their symmetric positive definite or skew-symmetric structure. We demonstrate the performance of the novel approach for three numerical examples that are used to simulate basic bending and rotating.

math.DS

Computation of structured stability radii for Dissipative-Hamiltonian systems

We study linear time-invariant Dissipative Hamiltonian (DH) systems arising in energy-based modeling of dynamical systems. An advantage of DH systems is that they are always stable due to the structure of their coefficient matrices, and, under further weak conditions, even asymptotically stable. In this paper, we discuss the computation of the stability radii for a given asymptotically stable DH system; i.e., the smallest structured perturbation that puts a DH system on the boundary of the region of asymptotic stability, so that it has purely imaginary eigenvalues. We obtain explicit computable formulas for various structured stability radii. For this, the problem of computing stability radii is reformulated in terms of minimizing the Rayleigh quotient of a Hermitian matrix or the sum of two generalized Rayleigh quotients of Hermitian semidefinite matrices. This reformulation results in the problem of minimizing the largest eigenvalue of an eigenvector-dependent Hermitian matrix or minimizing the smallest eigenvalue of a Hermitian matrix which depends on the eigenvector. It is also demonstrated (via numerical experiments) that, under structure-preserving perturbations, the asymptotic stability of a DH system is much more robust than under general perturbations, since the distance to instability is typically much larger when structure-preserving perturbations are considered. Finally, similar results are obtained for optimally robust representations of stable systems.

math.OC

Fast tensor-based electrostatic energy calculations in the perspective of protein-ligand docking problem

We propose and justify a new approach for fast calculation of the electrostatic interaction energy of clusters of charged particles in constrained energy minimization in the framework of rigid protein-ligand docking. Our ``blind search'' docking technique is based on the low-rank range-separated (RS) tensor-based representation of the free-space electrostatic potential of the biomolecule represented on large $n\times n\times n$ 3D grid. We show that both the collective electrostatic potential of a complex protein-ligand system and the respective electrostatic interaction energy can be calculated by tensor techniques in $O(n)$-complexity, such that the numerical cost for energy calculation only mildly (logarithmically) depends on the number of particles in the system. Moreover, tensor representation of the electrostatic potential enables usage of large 3D Cartesian grids (of the order of $n^3 \sim 10^{12}$), which could allow the accurate modeling of complexes with several large proteins. In our approach selection of the correct geometric pose predictions in the localized posing process is based on the control of van der Waals distance between the target molecular clusters. Here, we confine ourselves by constrained minimization of the energy functional by using only fast tensor-based free-space electrostatic energy recalculation for various rotations and translations of both clusters. Numerical tests of the electrostatic energy-based ``protein-ligand docking'' algorithm applied to synthetic and realistic input data present a proof of concept for rather complex particle configurations. The method may be used in the framework of the traditional stochastic or deterministic posing/docking techniques.

math.NA

Mixed-precision iterative refinement for low-rank Lyapunov equations

We develop a mixed-precision iterative refinement framework for solving low-rank Lyapunov matrix equations $AX + XA^T + W =0$, where $W=LL^T$ or $W=LSL^T$. Via rounding error analysis of the algorithms we derive sufficient conditions for the attainable normwise residuals in different precision settings and show how the algorithmic parameters should be chosen. These conditions are independent of the choice of inner solver, provided that the prescribed residual accuracy is attained in the inner solves. Using the sign-function Newton iteration as the solver, we demonstrate that reduced precisions, such as half precision with unit roundoff $u_s$, can be used efficiently for Lyapunov equations with condition numbers of order $1/u_s$ without compromising the attainable solution quality. This provides an algorithmic framework towards exploiting native low-precision hardware to accelerate Lyapunov solvers without sacrificing accuracy.

math.NA

Time-adaptive H\'enonNets for separable Hamiltonian systems

Measurement data is often sampled irregularly, i.e., not on equidistant time grids. This is also true for Hamiltonian systems. However, existing machine learning methods, which learn symplectic integrators, such as SympNets [1] and H\'enonNets [2] still require training data generated by fixed step sizes. To learn time-adaptive symplectic integrators, an extension to SympNets called TSympNets is introduced in [3]. The aim of this work is to do a similar extension for H\'enonNets. We propose a novel neural network architecture called T-H\'enonNets, which is symplectic by design and can handle adaptive time steps. We also extend the T-H\'enonNet architecture to non-autonomous Hamiltonian systems. Additionally, we provide universal approximation theorems for both new architectures for separable Hamiltonian systems and discuss why it is difficult to handle non-separable Hamiltonian systems with the proposed methods. To investigate these theoretical approximation capabilities, we perform different numerical experiments.

cs.LG

Time-adaptive SympNets for separable Hamiltonian systems

Measurement data is often sampled irregularly i.e. not on equidistant time grids. This is also true for Hamiltonian systems. However, existing machine learning methods, which learn symplectic integrators, such as SympNets [20] and H\'enonNets [4] still require training data generated by fixed step sizes. To learn time-adaptive symplectic integrators, an extension to SympNets, which we call TSympNets, was introduced in [20]. We adapt the architecture of TSympNets and extend them to non-autonomous Hamiltonian systems. So far the approximation qualities of TSympNets were unknown. We close this gap by providing a universal approximation theorem for separable Hamiltonian systems and show that it is not possible to extend it to non-separable Hamiltonian systems. To investigate these theoretical approximation capabilities, we perform different numerical experiments. Furthermore we fix a mistake in a proof of a substantial theorem [25, Theorem 2] for the approximation of symplectic maps in general, but specifically for symplectic machine learning methods.

cs.LG

Symplectic convolutional neural networks

We propose a new symplectic convolutional neural network (CNN) architecture by leveraging symplectic neural networks, proper symplectic decomposition, and tensor techniques. Specifically, we first introduce a mathematically equivalent form of the convolution layer and then, using symplectic neural networks, we demonstrate a way to parameterize the layers of the CNN to ensure that the convolution layer remains symplectic. To construct a complete autoencoder, we introduce a symplectic pooling layer. We demonstrate the performance of the proposed neural network on three examples: the wave equation, the nonlinear Schr\"odinger (NLS) equation, and the sine-Gordon equation. The numerical results indicate that the symplectic CNN outperforms the linear symplectic autoencoder obtained via proper symplectic decomposition.

cs.LG