arXiv ScienceSearch

SEARCH · arXiv Science

Results for “math.OA”

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,313 records · Page 7Linked to original sources

Accelerated High-Accuracy Sampling from a Warm Start via the Proximal Bouncy Particle Sampler

We study the problem of sampling from $μ(\mathrm{d}x)\propto e^{-V(x)}\,\mathrm{d}x$ on $\mathbb{R}^d$, where $V$ is $α$-strongly convex and $β$-smooth, and write $κ:=β/α$. We design and analyze the Proximal Bouncy Particle Sampler (Proximal BPS), a new sampler that combines ideas from the proximal sampler and the bouncy particle sampler. From a warm start initialization with $ O(1) $ Rényi divergence w.r.t. $μ$, Proximal BPS returns a sample whose law is $\varepsilon$-close to $μ$ in total variation distance using $\widetilde O(\sqrtκ\,d^{1/4} \,\mathrm{polylog}(1/\varepsilon))$ gradient queries in expectation.

math.ST

Tensor network representations of discrete maximum entropy distributions via mean polytopes

We present tensor network representations for discrete maximum entropy distributions under expectation constraints. To this end, we introduce Computation-Activation Networks (CompActNets), a tensor network architecture that subsumes exponential families. By leveraging the geometry of the convex polytope of realizable expectation vectors, we represent any maximum entropy distribution in the same architecture. We exploit the fact that proper faces of this polytope correspond to the boundary closure of exponential families, which restricts the distribution's support. We then derive explicit representations for the support within the CompActNet architecture. The proposed framework suggests tensor network ranks as complexity measures for faces. Finally, a case study on Boolean statistics links the geometry of 0/1-polytopes directly to propositional formulas.

math.ST

Coexact completion of profinite Heyting algebras and uniform interpolation

This paper shows that the sheaf representation of finitely generated free Heyting algebras constructed by Ghilardi and Zawadowski can be factored as the profinite completion of Heyting algebras, followed by identifying the dual category of profinite Heyting algebras as a full subcategory of a sheaf topos. We show that the dual category of profinite Heyting algebras is an infinitary extensive regular category, and its ex/reg-completion is exactly the aforementioned sheaf topos, which we refer to as the K-topos. We show how certain properties of uniform interpolation can be generalised to the context of arbitrary profinite Heyting algebras, and that they are consequences of the internal logic of the K-topos. Along the way we also establish various topos-theoretic properties of the K-topos.

math.LO

The Minimum Number of Measurements for Almost-Everywhere Complex Phase Retrieval

Let $d\geq 2$ and let $\bf{f}_1,\ldots,\bf{f}_m\in\mathbb C^d$. We prove that if $m\leq 2d-1$, then the intensity measurement map \[ \bf{x}\longmapsto \bigl( |\langle \bf{x},\bf{f}_1\rangle|^2, \ldots, |\langle \bf{x},\bf{f}_m\rangle|^2 \bigr) \] fails to recover almost every signal in $\mathbb C^d$ uniquely up to a global phase factor. Combined with the known generic sufficiency of $2d$ measurements, our result establishes that the minimum number of measurements required for almost-everywhere phase retrieval in $\mathbb C^d$ is exactly $2d$. This resolves an open problem in phase retrieval by determining the exact measurement threshold for almost-everywhere phase retrieval in ${\mathbb C}^d$.

cs.IT

Frequency-explicit convergence analysis of a multiscale finite element method for highly heterogeneous scattering problems

We analyze the numerical approximation of time-harmonic scattering by highly heterogeneous penetrable obstacles. These problems are especially challenging in the high-frequency regime, where the size of the scatterer $L$ is much larger than the wavelength, i.e., the wavenumber $k$ is such that $kL \gg 1$. Here, we further consider the situation where the scatterer contains different materials, with a characteristic size $\varepsilon$ such that $k\varepsilon \ll 1$. We propose a high-order multiscale finite element method, and provide an error analysis that is explicit in both $k$ and $\varepsilon$. Crucially, our error estimates suggest that using a high-order method should reduce the computational cost for large frequencies, which is corroborated by numerical examples.

math.NA

Differential uniformity properties of some classes of permutation polynomials

The notion of $c$-differential uniformity has recently received a lot of attention since its proposal~\cite{Ellingsen}, and recently a characterization of perfect $c$-nonlinear functions in terms of difference sets in some quasigroups was obtained in~\cite{AMS22}. Independent of their applications as a measure for certain statistical biases, the construction of functions, especially permutations, with low $c$-differential uniformity is an interesting mathematical problem in this area, and recent work has focused heavily in this direction. We provide a few classes of permutation polynomials with low $c$-differential uniformity. The used technique involves handling various Weil sums, as well as analyzing some equations in finite fields, and we believe these can be of independent interest.

cs.IT

Quiver Semistability and Structured Kalman Decompositions for Networked Linear Dynamical Systems

We introduce new notions of controllability and observability for networked linear time-invariant (LTI) systems based on $σ$-semistability of quiver representations. Utilizing King's criterion for $σ$-semistability, we define a network generalization of the Kalman decomposition for networked LTI systems, which systematically decomposes the local and interconnection dynamics while respecting the underlying network structure. Furthermore, we present efficient algorithms for deciding the proposed controllability and observability of a given networked LTI system and for finding the Kalman-type decomposition. We also show efficient algorithms for deciding the $σ$-semistability of representations of acyclic quivers with self-loops if the weight $σ$ has the same sign for all vertices with self-loops. Such quiver representations and weights arise from networked LTI systems.

math.OC

Bellman-sufficient Information Complexity

We introduce Bellman-sufficient information complexity for minimax analysis of sequential decision problems. A Bellman-sufficient state retains enough of the history to close the controlled recursion, while an index $Y=χ(Ω)$ specifies the decision-relevant information being charged. The upper bound is a log-penalized Bellman program; the lower bound is a Bellman--Fano comparison along an algorithm-dependent reference trajectory. If the two values match at a common localization scale and the stated admissibility, calibration, and growth conditions hold, they form an information-risk sandwich. UCB, E2D, and AMS/EBO control or relax the upper Bellman bracket in different ways. For the main application, we give a negative answer to a widely studied form of the GP--UCB minimax-optimality question. For every $0<α<1/4$, we construct one bounded continuous kernel whose minimax regret is $Θ(T^{1-α})$ along an infinite sequence of horizons, while two globally calibrated GP--UCB rules incur linear regret under one fixed truth. An epochwise finite-marginal action-index AIR Bellman policy, implemented through robust AIR/AMS/EBO control, attains the minimax order. The construction separates realized information from the cost of uniform optimism: many low-value directions inflate the exploration multiplier and change the trajectory. Through the canonical RKHS feature map, it also yields a finite-horizon polynomial minimax separation for the specified maximal-information-calibrated LinUCB rule. A reproducible experiment illustrates the mechanism.

cs.LG

ENPINN: Energy-Norm-Guided Gradient-Enhanced PINNs for Generalized Transport Problems with Sharp Gradients

Physics-informed neural networks (PINNs) have emerged as a meshless alternative to conventional numerical methods for solving partial differential equations (PDEs). However, their limited ability to capture sharp gradients can lead to substantial errors when resolving boundary and interior layers. Here, we introduce an energy-norm-enhanced PINN (ENPINN) that incorporates gradient information and variational structure into the loss function to improve the resolution of layer-dominated solutions. We first examine two related formulations: weak-loss PINNs (WLPINNs), which incorporate test functions into the conventional PINN residual, and gradient-enhanced PINNs (gPINNs), which augment the loss with spatial derivatives of the PDE residual. By analyzing these formulations, we identify their limitations in resolving steep solution gradients and motivate the systematic construction of ENPINN. We establish theoretically how the energy-norm error depends on the ENPINN loss and show that a suitably modified residual-derivative term is essential for accurately capturing boundary layers. We further establish the existence of neural-network approximations with arbitrarily small energy error and derive corresponding derivative bounds, providing a theoretical foundation for the proposed framework. The performance of ENPINN is assessed through systematic comparisons with existing PINN variants for convection-diffusion-reaction problems exhibiting steep gradients. Numerical experiments include a combustion model, a coupled multi-scale system, a two-dimensional Burgers equation with an interior layer, and a three-dimensional time-dependent problem.

math.NA

Semi-discrete quadratic Wasserstein energy and state-dependent Langevin exploration

We study the semi-discrete quadratic Wasserstein energy. The energy is nonsmooth at collisions of sites. We prove local Lipschitz continuity on the full configuration space, together with global semiconcavity, coercivity, and dissipativity; show that every global minimizer is interior and collision free; and establish $C^2$ regularity on the collision-free configuration space. The gradient is expressed through the barycenters of the balanced Laguerre cells, while the Hessian is given by an explicit facet formula and satisfies a global one-sided bound. We also solve the one-dimensional problem explicitly in each ordering chamber and give a two-site example on the unit square with non-minimizing Lloyd fixed points. For $d\ge 2$, we then formulate an entropy-regularized relaxed control of the Langevin temperature. The controlled dynamics is strongly well posed, nonexplosive, and collision free. Its value function is a classical interior solution of the exploratory Hamilton-Jacobi-Bellman equation; the Laplacian of the value function is locally $C^1$, which yields a locally Lipschitz optimal temperature feedback. Independently of this optimal-control result, for every fixed Borel temperature rule bounded away from zero, and every sufficiently small step size, the associated Gaussian Euler chain is geometrically ergodic with a full-support invariant law. The raw iterates do not converge, whereas the best-so-far energy converges almost surely to the global minimum and the running record approaches the set of global minimizers.

math.NA

Compact implicit high-resolution numerical scheme for multicomponent transport problems with Langmuir sorption

We present a compact implicit high-resolution finite volume scheme for one-dimensional multicomponent transport problems with nonlinear Langmuir sorption. The method combines a second order compact implicit discretization with WENO reconstruction and a time limiter for the treatment of steep gradients and discontinuities. A special attention is given to the Courant number used in the limiter, where local characteristic Courant numbers based on the eigenvalues of the inverse retardation matrix are considered. The compact implicit schemes are compared with analogous explicit first and second order methods. Numerical experiments show second order convergence for smooth solutions and demonstrate that the implicit approach can be used with significantly larger Courant numbers. The method is also applied to a three-component displacement chromatography problem with sharp concentration fronts.

math.NA

Eigenvalue stability and new perturbation bounds for the extremal eigenvalues of a matrix

Let $A$ be a full ranked $ n\times n$ matrix, with singular values $σ_1 (A) \ge \dots \ge σ_n (A) >0$. The condition number $κ(A):= σ_1(A)/σ_n(A)=\|A\|\cdot \|A\|^{-1}$ is a key parameter in the analysis of algorithms taking $A$ as input. In practice, matrices (representing real data) are often perturbed by noise. Technically speaking, the real input would be a noisy variant $\tilde A =A +E$ of $A$, where $E$ represents the noise. The condition number $κ(\tilde A)$ will be used instead of $κ(A)$. Thus, it is of importance to measure the impact of noise on the condition number. In this paper, we focus on the case when the noise is random. We introduce the notion of regional stability, via which we design a new framework to estimate the perturbation of the extremal singular values and the condition number of a matrix. Our framework allows us to bound the perturbation of singular values through the perturbation of singular spaces. We then bound the latter using a novel contour analysis argument, which, as a co-product, provides an improved version of the classical Davis-Kahan theorem in many settings. Our new estimates concerning the least singular value $σ_n(A)$ complement well-known results in this area, and are more favorable in the case when the ground matrix $A$ is large compared to the noise matrix $E$.

math.NA

Thermal Recurrence Orders of the Potts Model Partition Function in Grid Graphs

We study the linear recurrence order of the Potts model partition function on thermal 2D grid graphs. By restricting the transfer matrix (TM) to the real coupling axis, the physical operator maintains diagonalizability under the Spectral Theorem. We show that this thermal regularization allows the Krylov subspace to saturate the unconstrained planar state capacity, locking the recurrence order to the Dyck path up to reversal (OEIS A007123) for $q \ge 4$. Furthermore, the recurrence order collapses to height-restricted Dyck paths up to reversal (OEIS A001998) for $q=3$ due to finite-index Jones-Wenzl projections, and to the zero-magnetization conservation sector (OEIS A001405) for $q=2$. This framework bridges graph-theoretic combinatorics with the representation theory of physical loop gas models.

cs.DM

No Equivariant Architecture Covers All Equivariant Attention

We give a complete characterization of equivariant multi-head self-attention (MHSA): if an MHSA layer is equivariant to a symmetry group $G$, then $G$ can only act by permuting head-clusters, with QK and OV matrices satisfying an equivariance constraint tied to the group action. As a consequence, we prove that any fixed MHSA architecture that achieves exact equivariance by polynomially parameterizing unconstrained MHSA parameters inevitably leads to expressivity loss within the class of equivariant maps: the equivariance locus of unconstrained MHSA forms a union of extremely many Zariski-irreducible components in a reduced parameter space, and any single architecture covers at most one. For $G=D_4$ acting on $C$ copies of the regular representation as the token feature space, we show that there are $Ω(C^{64})$ components for eight attention heads.

cs.LG

Rethinking quantum smooth entropies: Tight one-shot analysis of quantum privacy amplification

We introduce an improved one-shot characterisation of randomness extraction against quantum side information (privacy amplification), strengthening known one-shot bounds and providing a unified derivation of the tightest known asymptotic constraints. Our main tool is a new class of smooth conditional entropies defined by lifting classical smooth divergences through measurements. A key role is played by the measured smooth Rényi relative entropy of order 2, which we show to admit an equivalent variational form: it can be understood as allowing for smoothing over not only states, but also non-positive Hermitian operators. Building on this, we establish a tightened leftover hash lemma, significantly improving over all known smooth min-entropy bounds on extractable randomness and recovering the sharpest classical achievability results. We extend these methods to decoupling, the coherent analogue of privacy amplification, obtaining a corresponding improved one-shot bound. Relaxing our smooth entropy bounds leads to one-shot achievability results in terms of measured Rényi divergences, tightening the bounds of [Dupuis, arXiv:2105.05342] and recovering state-of-the-art asymptotic i.i.d. error exponents. We show an approximate optimality of our results by giving a matching one-shot converse bound up to additive logarithmic terms. This yields an optimal second-order asymptotic expansion of privacy amplification under trace distance, establishing a significantly tighter one-shot achievability result than previously shown in [Shen et al., arXiv:2202.11590] and proving its optimality for all hash functions.

quant-ph

DOFFO_TR: a Decentralized Objective Function-Free Optimization method with Trust-Region

In this paper, we propose a novel objective function-free trust-region method designed to solve optimization problems over decentralized networks. Unlike traditional approaches that often rely on stepsize tuning, our framework employs a function-free trust-region procedure that enables adaptive selection of the step length. Our approach accommodates first- and second-order models and eliminates the need to share local function values and gradients among agents, thereby enhancing privacy and computational efficiency. On the theoretical side, we establish provable iteration complexity guarantees that, for some variants, match those established for classical centralized trust-region methods. Numerical evaluations demonstrate that our approach achieves a favorable trade-off between performance and efficiency, requiring only moderate communication overhead compared to state-of-the-art methods in the literature.

math.OC

Fully discrete stochastic maximal regularity and $H^\infty$-calculus for second-order elliptic operators

This paper establishes the fully discrete stochastic maximal $L^p$-regularity and the accompanying sharp maximal estimate for numerical approximations of parabolic stochastic partial differential equations. We consider the spatial finite element discretization $A_h$ of a general second-order elliptic operator $A=-\nabla \cdot a\nabla +b\cdot \nabla +c$ with Dirichlet boundary conditions on a smooth, bounded, convex domain in $\mathbb{R}^3$, coupled with a broad class of temporal schemes, including rational approximations and the exponential Euler method. To obtain these optimal discrete regularity results, we establish a bounded $H^\infty$-calculus for the discrete spatial operator $A_h$, uniformly in the mesh size $h$. As a direct byproduct, we also establish the discrete-in-space stochastic maximal regularity for the corresponding spatial semi-discretizations.

math.AP

A Complete Resolution of Forsythe's Conjecture for Restarted Conjugate Gradients

Forsythe's conjecture, published in 1968, asserts that for each restart length $s$, every exact-arithmetic restarted conjugate-gradient iteration on a real symmetric positive definite problem either terminates or has normalised residuals that converge separately along the even and odd restart subsequences. Apart from the classical steepest-descent case, this asymptotic question remained unresolved in full generality for nearly six decades. We give a complete classification by restart length in the original finite-dimensional setting and identify a sharp threshold. For $s=2$ and $s=3$, every problem either terminates or has convergent even and odd residual directions. For every $s\ge4$, there is a diagonal positive definite counterexample of dimension $s+4$ which never terminates and whose even residual directions do not converge. Together with Akaike's theorem for $s=1$, this shows that the conjectured universal conclusion is true precisely for $s\in\{1,2,3\}$ and false for every $s\ge4$. The positive results follow from a degree-independent double-orthogonality identity and an analysis of the low-degree fixed-point sets. At restart length four, rational interval arithmetic and Sturm sequences certify a transverse Hopf point of the leading vector field of the rescaled squared-weight map. Analytic periodic-orbit and shadowing arguments yield the counterexample at restart length four, and degree elevation extends the construction to every larger restart length. The classification for all $s\ge2$ is also formally verified in Lean.

math.NA