arXiv ScienceSearch

arXiv subjects

Michael Hecht

Publications and source records attributed to Michael Hecht.

At least 19 recordsLinked to original sources

Analysis of the Ill-Conditioning of the Discrete Inverse Laplace transform in Monte Carlo Simulations of Quantum Many-Body Systems

Analytic continuation of imaginary-time quantum Monte Carlo data to real-frequency spectra requires the inversion of a severely ill-posed two-sided Laplace transform and arises naturally in quantum many-body calculations of dynamic properties. In this work, we distinguish the intrinsic ill-posedness of the continuous inverse two-sided Laplace problem from the conditioning of its finite-dimensional discretization. For equidistant sampling and reconstruction grids, we express the discrete problem via a diagonally scaled monomial Vandermonde matrix with exponentially distributed nodes. Imposing the physical detailed-balance symmetry transforms the discretization into a diagonally scaled Chebyshev-Vandermonde system. Exploiting these structures, we derive explicit lower and upper bounds on the condition numbers in terms of the physical and discretization parameters. For the unconstrained discretization, our bounds reveal super-exponential growth of the condition number with the reconstruction dimension, which cannot be removed by increasing the number of imaginary-time samples alone. Detailed-balance substantially improves the conditioning, especially in the practically relevant pre-asymptotic regime, although the asymptotic super-exponential dependence remains. In both cases, our bounds identify a low-dimensional regime in which the super-exponential contribution is suppressed and stable reconstruction remains feasible. These results provide a mathematical explanation for the effectiveness of low-dimensional spectral representations and detailed-balance in analytic continuation. While they do not remove the fundamental ill-posedness, they substantially mitigate the ill-conditioning of its finite-dimensional discretization, delay the onset of its catastrophic super-exponential growth and thereby enlarge the range of numerically accessible reconstructions.

math.NA

On Adversarial Vulnerability of Vision-Language Models through the Lens of Intermediate Spectral Subspaces

Adversarial vulnerability in deep neural networks (DNNs) has been studied from the perspectives of decision-boundary geometry, feature robustness, input-output Jacobians, and the instability of inverse problems. Here, we focus on the spectral structure of intermediate linear transformations that propagate information through modern DNNs, an unexplored mechanism of adversarial vulnerability. Specifically, we investigate transformer-based vision-language models, whose linear layers admit interpretable spectral decompositions and whose widespread adoption makes understanding their robustness increasingly important. We propose a white-box spectral-subspace-guided attack (SSGRA) that aligns intermediate representations with the subspace spanned by the bottom right singular vectors. Our experiments show improved attack effectiveness over existing baselines. In addition, SSGRA offers a spectral interpretation of adversarial vulnerability in VLMs, providing insights for improving their robustness.

cs.LG

Mitigating Numerical Stiffness in Least-Squares Formulations of Elliptic PDEs for Physics-Informed Neural Networks

We present theoretical insights into $H^{-1}$ residual loss formulations of physics-informed neural networks (PINNs) for learning solutions of partial differential equations (PDEs). Standard PINN formulations use a multi-term loss functional consisting of interior and boundary loss terms that are based on $L^2$-residuals and discretized as mean square errors (MSE). Imbalanced magnitudes of these terms cause numerical stiffness phenomena, resulting in ill-conditioning and slow convergence. In this work, we analyze discretizations of the $H^{-1}$-norm that are used in the context of elliptic PDEs with arbitrary, nonzero Dirichlet boundary conditions. We prove that these $H^{-1}$ discretizations rebalance the PDE loss, improve conditioning, and mitigate stiffness effects compared with the standard MSE discretization. We validate our theoretic results through operator-level experiments with randomly sampled residuals and PINN experiments for the Poisson and stationary incompressible Navier-Stokes equations. These experiments confirm the numerical effectiveness of the proposed rebalancing for elliptic PDEs and, more broadly, for problems with elliptic behavior.

math.NA

Extinction and persistence criteria in non-local Klausmeier model of vegetation dynamics on flat landscapes

This paper investigates the dynamics of vegetation patterns in water-limited ecosystems using a generalized Klausmeier model that incorporates non-local plant dispersal within a finite habitat. We establish the well-posedness of the system and provide a rigorous analysis of the conditions required for vegetation survival. Our results identify a critical patch size governed by the trade-off between local growth and boundary losses; habitats smaller than this threshold lead to inevitable extinction. Furthermore, we derive a critical maximal biomass density below which the population collapses to a desert state, regardless of the domain size. We determine stability criteria for stationary solutions and describe the emergence of stable, non-trivial biomass distributions. Numerical experiments comparing sub-Gaussian and super-Gaussian kernels confirm that non-local dispersal mechanisms, particularly those with fat tails, enhance ecosystem resilience by allowing vegetation to persist in smaller, fragmented habitats than predicted by classical local diffusion models.

math.AP

Interpolation in Polynomial Spaces of p-Degree

We recently introduced the Fast Newton Transform (FNT), an hierarchical algorithm for performing multivariate Newton interpolation in arbitrary downward closed polynomial spaces of spatial dimension $m$. Here, we analyze the FNT in the context of a specific family of downward closed sets $A_{m,n,p}$, defined as all multi-indices with $\ell^p$ norm less than $n$ with $p \in [0,\infty]$. The FNT performs with time complexity $\mathcal{O}(|A_{m,n,p}|mn)$ on the induced downward closed polynomial spaces $\Pi_{m,n,p}$. We show that the $\Pi_{m,n,p}$ choice compared to the tensor product spaces $\Pi_{m,n,\infty}$, reduces time complexity by a factor of $\rho_{m,n,p}$, decaying super exponentially with spatial dimension when $m \lesssim n^p$. We showcase the efficiency of the FNT by computing activity scores in sensitivity analysis.

math.NA

Accelerating Multivariate Newton Interpolation in Downward Closed Polynomial Spaces

We introduce the fast Newton transform (FNT), a multivariate Newton interpolation algorithm for downward closed polynomial spaces in quasi-tensorial grids. The FNT computes the Newton coefficients directly, without relying on embeddings into enclosing tensor-product spaces. For a downward closed index set $A \subset \mathbb N_0^m$, the FNT achieves a time complexity of $\mathcal O(m \overline n |A|)$, where $\overline n$ is the mean of the coordinate-wise maximal polynomial degrees $n_1, \ldots, n_m$ across the $m$ spatial dimensions. In the univariate case, the FNT renders the classic Newton divided difference scheme (DDS). In the multivariate case, however, it improves on the quadratic complexity $\mathcal O(|A|^2)$ of the DDS and on the cost $\mathcal{O}(m \overline n (n_1+1) \cdots (n_m+1))$ of the tensorial interpolation. For sufficiently regular functions, Newton interpolation in Euclidean-degree downward closed polynomial spaces is known to deliver approximation rates equal to those of the tensor product interpolation. Thus, the acceleration power of the FNT comes from requiring substantially fewer degrees of freedom while reaching the same approximation quality as the tensorial interpolation. The inverse transformation has the same time complexity and enables fast evaluation and differentiation of Newton interpolants in quasi-tensorial grids.

math.NA

Second roton feature in the strongly coupled electron liquid

We present extensive \emph{ab initio} path integral Monte Carlo (PIMC) results for the dynamic properties of the finite temperature uniform electron gas (UEG) over a broad range of densities, $2\leq r_s\leq300$. We demonstrate that the direct analysis of the imaginary-time density--density correlation function (ITCF) allows for a rigorous assessment of the density and temperature dependence of the previously reported roton-type feature [T.~Dornheim, \emph{Phys.~Rev.~Lett.}~\textbf{121}, 255001 (2018)] at intermediate wavenumbers. We clearly resolve the emergence of a second roton at the second harmonic of the original feature for $r_s\gtrsim100$, which we identify as an incipient phonon dispersion. Finally, we use our highly accurate PIMC results for the ITCF as the basis for an analytic continuation to compute the dynamic structure factor, which additionally substantiates the existence of the second roton in the strongly coupled electron liquid. Our investigation further elucidates the complex interplay between quantum delocalization and Coulomb coupling in the UEG. All PIMC results are freely available online and provide valuable benchmarks for other theoretical methodologies and approximations.

physics.chem-ph

PyLIT: Reformulation and implementation of the analytic continuation problem using kernel representation methods

Path integral Monte Carlo (PIMC) simulations are a cornerstone for studying quantum many-body systems. The analytic continuation (AC) needed to estimate dynamic quantities from these simulations is an inverse Laplace transform, which is ill-conditioned. If this inversion were surmounted, then dynamical observables (e.g. dynamic structure factor (DSF) $S(q,\omega)$) could be extracted from the imaginary-time correlation functions estimates. Although of important, the AC problem remains challenging due to its ill-posedness. To address this challenge, we express the DSF as a linear combination of kernel functions with known Laplace transforms that have been tailored to satisfy its physical constraints. We use least-squares optimization regularized with a Bayesian prior to determine the coefficients of this linear combination. We explore various regularization term, such as the commonly used entropic regularizer, as well as the Wasserstein distance and $L^2$-distance as well as techniques for setting the regularization weight. A key outcome is the open-source package PyLIT (\textbf{Py}thon \textbf{L}aplace \textbf{I}nverse \textbf{T}ransform), which leverages Numba and unifies the presented formulations. PyLIT's core functionality is kernel construction and optimization. In our applications, we find PyLIT's DSF estimates share qualitative features with other more established methods. We identify three key findings. Firstly, independent of the regularization choice, utilizing non-uniform grid point distributions reduced the number of unknowns and thus reduced our space of possible solutions. Secondly, the Wasserstein distance, a previously unexplored regularizer, performs as good as the entropic regularizer while benefiting from its linear gradient. Thirdly, future work can meaningfully combine regularized and stochastic optimization. (text cut for char. limit)

physics.comp-ph

Multivariate Newton Interpolation in Downward Closed Spaces Reaches the Optimal Geometric Approximation Rates for Bos--Levenberg--Trefethen Functions

We extend the univariate Newton interpolation algorithm to arbitrary spatial dimensions and for any choice of downward-closed polynomial space, while preserving its quadratic runtime and linear storage cost. The generalisation supports any choice of the provided notion of non-tensorial unisolvent interpolation nodes, whose number coincides with the dimension of the chosen-downward closed space. Specifically, we prove that by selecting Leja-ordered Chebyshev-Lobatto or Leja nodes, the optimal geometric approximation rates for a class of analytic functions -- termed Bos--Levenberg--Trefethen functions -- are achieved and extend to the derivatives of the interpolants. In particular, choosing Euclidean degree results in downward-closed spaces whose dimension only grows sub-exponentially with spatial dimension, while delivering approximation rates close to, or even matching those of the tensorial maximum-degree case, mitigating the curse of dimensionality. Several numerical experiments demonstrate the performance of the resulting multivariate Newton interpolation compared to state-of-the-art alternatives and validate our theoretical results.

math.NA

Hybrid Surrogate Models: Circumventing Gibbs Phenomenon for Partial Differential Equations with Finite Shock-Type Discontinuities

We introduce the concept of Hybrid Surrogate Models (HSMs) -- combining multivariate polynomials with Heavyside functions -- as approximates of functions with finitely many jump discontinuities. We exploit the HSMs for formulating a variational optimization approach, solving non-regular partial differential equations (PDEs) with non-continuous shock-type solutions. The HSM technique simultaneously obtains a parametrization of the position and the height of the shocks as well as the solution of the PDE. We show that the HSM technique circumvents the notorious Gibbs phenomenon, which limits the accuracy that classic numerical methods reach. Numerical experiments, addressing linear and non-linearly propagating shocks, demonstrate the strong approximation power of the HSM technique.

math.NA

High-order numerical integration on regular embedded surfaces

We present a high-order surface quadrature (HOSQ) for accurately approximating regular surface integrals on closed surfaces. The initial step of our approach rests on exploiting square-squeezing--a homeomorphic bilinear square-simplex transformation, re-parametrizing any surface triangulation to a quadrilateral mesh. For each resulting quadrilateral domain we interpolate the geometry by tensor polynomials in Chebyshev--Lobatto grids. Posterior the tensor-product Clenshaw-Curtis quadrature is applied to compute the resulting integral. We demonstrate efficiency, fast runtime performance, high-order accuracy, and robustness for complex geometries.

math.NA

PMBO: Enhancing Black-Box Optimization through Multivariate Polynomial Surrogates

We introduce a surrogate-based black-box optimization method, termed Polynomial-model-based optimization (PMBO). The algorithm alternates polynomial approximation with Bayesian optimization steps, using Gaussian processes to model the error between the objective and its polynomial fit. We describe the algorithmic design of PMBO and compare the results of the performance of PMBO with several optimization methods for a set of analytic test functions. The results show that PMBO outperforms the classic Bayesian optimization and is robust with respect to the choice of its correlation function family and its hyper-parameter setting, which, on the contrary, need to be carefully tuned in classic Bayesian optimization. Remarkably, PMBO performs comparably with state-of-the-art evolutionary algorithms such as the Covariance Matrix Adaptation -- Evolution Strategy (CMA-ES). This finding suggests that PMBO emerges as the pivotal choice among surrogate-based optimization methods when addressing low-dimensional optimization problems. Hereby, the simple nature of polynomials opens the opportunity for interpretation and analysis of the inferred surrogate model, providing a macroscopic perspective on the landscape of the objective function.

math.OC

High-order integration on regular triangulated manifolds reaches super-algebraic approximation rates through cubical re-parameterizations

We present a novel methodology for deriving high-order volume elements (HOVE) designed for the integration of scalar functions over regular embedded manifolds. For constructing HOVE we introduce square-squeezing --a homeomorphic multilinear hypercube-simplex transformation reparametrizing an initial flat triangulation of the manifold to a cubical mesh. By employing square-squeezing, we approximate the integrand and the volume element for each hypercube domain of the reparameterized mesh through interpolation in Chebyshev-Lobatto grids. This strategy circumvents the Runge phenomenon, replacing the initial integral with a closed-form expression that can be precisely computed by high-order quadratures. We prove novel bounds of the integration error in terms of the $r^\text{th}$-order total variation of the integrand and the surface parameterization, predicting high algebraic approximation rates that scale solely with the interpolation degree and not, as is common, with the average simplex size. For smooth integrals whose total variation is constantly bounded with increasing $r$, the estimates prove the integration error to decrease even exponentially, while mesh refinements are limited to achieve algebraic rates. The resulting approximation power is demonstrated in several numerical experiments, particularly showcasing $p$-refinements to overcome the limitations of $h$-refinements for highly varying smooth integrals.

math.NA

Ensuring Topological Data-Structure Preservation under Autoencoder Compression due to Latent Space Regularization in Gauss--Legendre nodes

We formulate a data independent latent space regularisation constraint for general unsupervised autoencoders. The regularisation rests on sampling the autoencoder Jacobian in Legendre nodes, being the centre of the Gauss-Legendre quadrature. Revisiting this classic enables to prove that regularised autoencoders ensure a one-to-one re-embedding of the initial data manifold to its latent representation. Demonstrations show that prior proposed regularisation strategies, such as contractive autoencoding, cause topological defects already for simple examples, and so do convolutional based (variational) autoencoders. In contrast, topological preservation is ensured already by standard multilayer perceptron neural networks when being regularised due to our contribution. This observation extends through the classic FashionMNIST dataset up to real world encoding problems for MRI brain scans, suggesting that, across disciplines, reliable low dimensional representations of complex high-dimensional datasets can be delivered due to this regularisation technique.

cs.LG

Polynomial-Model-Based Optimization for Blackbox Objectives

For a wide range of applications the structure of systems like Neural Networks or complex simulations, is unknown and approximation is costly or even impossible. Black-box optimization seeks to find optimal (hyper-) parameters for these systems such that a pre-defined objective function is minimized. Polynomial-Model-Based Optimization (PMBO) is a novel blackbox optimizer that finds the minimum by fitting a polynomial surrogate to the objective function. Motivated by Bayesian optimization the model is iteratively updated according to the acquisition function Expected Improvement, thus balancing the exploitation and exploration rate and providing an uncertainty estimate of the model. PMBO is benchmarked against other state-of-the-art algorithms for a given set of artificial, analytical functions. PMBO competes successfully with those algorithms and even outperforms all of them in some cases. As the results suggest, we believe PMBO is the pivotal choice for solving blackbox optimization tasks occurring in a wide range of disciplines.

cs.LG

Roadmap on Deep Learning for Microscopy

Through digital imaging, microscopy has evolved from primarily being a means for visual observation of life at the micro- and nano-scale, to a quantitative tool with ever-increasing resolution and throughput. Artificial intelligence, deep neural networks, and machine learning are all niche terms describing computational methods that have gained a pivotal role in microscopy-based research over the past decade. This Roadmap is written collectively by prominent researchers and encompasses selected aspects of how machine learning is applied to microscopy image data, with the aim of gaining scientific knowledge by improved image quality, automated detection, segmentation, classification and tracking of objects, and efficient merging of information from multiple imaging modalities. We aim to give the reader an overview of the key developments and an understanding of possibilities and limitations of machine learning for microscopy. It will be of interest to a wide cross-disciplinary audience in the physical sciences and life sciences.

physics.optics

Extraction of the frequency moments of spectral densities from imaginary-time correlation function data

We introduce an exact framework to compute the positive frequency moments $M^{(\alpha)}(\mathbf{q})=\braket{\omega^\alpha}$ of different dynamic properties from imaginary-time quantum Monte Carlo data. As a practical example, we obtain the first five moments of the dynamic structure factor $S(\mathbf{q},\omega)$ of the uniform electron gas at the electronic Fermi temperature based on \emph{ab initio} path integral Monte Carlo simulations. We find excellent agreement with known sum rules for $\alpha=1,3$, and, to our knowledge, present the first results for $\alpha=2,4,5$. Our idea can be straightforwardly generalized to other dynamic properties such as the single-particle spectral function $A(\mathbf{q},\omega)$, and will be useful for a number of applications, including the study of ultracold atoms, exotic warm dense matter, and condensed matter systems.

cond-mat.quant-gas

Learning Partial Differential Equations by Spectral Approximates of General Sobolev Spaces

We introduce a novel spectral, finite-dimensional approximation of general Sobolev spaces in terms of Chebyshev polynomials. Based on this polynomial surrogate model (PSM), we realise a variational formulation, solving a vast class of linear and non-linear partial differential equations (PDEs). The PSMs are as flexible as the physics-informed neural nets (PINNs) and provide an alternative for addressing inverse PDE problems, such as PDE-parameter inference. In contrast to PINNs, the PSMs result in a convex optimisation problem for a vast class of PDEs, including all linear ones, in which case the PSM-approximate is efficiently computable due to the exponential convergence rate of the underlying variational gradient descent. As a practical consequence prominent PDE problems were resolved by the PSMs without High Performance Computing (HPC) on a local machine. This gain in efficiency is complemented by an increase of approximation power, outperforming PINN alternatives in both accuracy and runtime. Beyond the empirical evidence we give here, the translation of classic PDE theory in terms of the Sobolev space approximates suggests the PSMs to be universally applicable to well-posed, regular forward and inverse PDE problems.

math.NA