arXiv ScienceSearch

SEARCH · arXiv Science

Results for “math.OA”

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

929 records · Page 5Linked to original sources

Feasible approximation of matching equilibria for large-scale matching for teams problems

We propose a numerical algorithm for computing feasible and approximately optimal solutions of the matching for teams problem. Specifically, we introduce the notion of approximate matching equilibrium as a feasible approximation of a matching equilibrium with relaxed rationality, and we show that a true equilibrium is recovered in the limit of a sequence of approximate matching equilibria with sub-optimality approaching 0. In our approximation scheme, we parametrize the so-called transfer functions, and we show that tackling the resulting parametric primal and dual optimization problems yields two approximate matching equilibria as well as provable and computable lower and upper bounds for the optimal social welfare. Under a flexible Euclidean setting, we show that the approximation error of our scheme can be controlled to be arbitrarily close to 0, we derive an explicit computational complexity bound, and we develop an algorithm for computing approximate matching equilibria that is efficient for large-scale problems involving a large number of agent populations. We study three problems in our numerical experiments: a retail business problem, the Wasserstein barycenter problem, and a large-scale problem involving up to 1000 agent populations. We show that the proposed algorithm can produce nearly optimal approximate matching equilibria to provide quantitative managerial insights for policymakers, and that the computed sub-optimality estimates are much less conservative than theoretical estimates.

math.OC

Group-averaged Markov chains II: tuning of group action in finite state space

We study group-averaged Markov chains obtained by augmenting a $π$-stationary kernel $P$ with orbit kernels induced by a group action. We analyse the Gibbs ($G$), Metropolis--Hastings ($M$), and Barker ($B$) kernels, their sandwiches $QPQ$, and mixtures $\tfrac{1}{2}(P+Q)$, where $Q\in\{G,M,B\}$. Under suitable conditions, $M^t$ and $B^t$ converge blockwise to $G$. The projection chains of $GPG$ and $P$ coincide, while every sandwich $QPQ$ has absolute spectral gap no smaller than that of reversible $P$. For $GPG$, we derive an additive asymptotic-variance bound, prove monotonicity for $G$-invariant observables, and identify it as the Kullback--Leibler (KL) information projection of $P$ onto the $G$-invariant kernels. For a fixed orbit partition, the spectral and KL properties of $GPG$ reduce to those of a lower-dimensional orbit-space chain. Among Gibbs projections with a prescribed number of orbits, we identify the partition minimizing KL divergence to stationarity and characterize exact stationarity. Finally, alternating group projections converge at a rate determined by singular values of an overlap matrix and, in structured cases, can yield exact sampling with logarithmically many group actions. These results motivate tuning heuristics and yield polynomial mixing for a Curie--Weiss example in a regime where Glauber dynamics is exponentially slow.

math.PR

An intrinsic expansion approach to the Galerkin approximations for the Navier-Stokes equations (with an appendix by Chengzhang Fu)

We study the Galerkin approximation of the three-dimensional Navier-Stokes equations. In particular, we examine the convergence of these solutions in a sequence of finite dimensional spaces as the dimension goes to infinity. For any sequence of steady state or, respectively, time dependent Galerkin solutions that converges to a solution of the Navier-Stokes equations, we obtain a subsequence with an intrinsic asymptotic expansion in appropriate nested function spaces. Consequently, an induced asymptotic expansion is obtained in a more standard spatial Sobolev or, respectively, spatiotemporal Sobolev-Lebesgue space. In the case of steady states, we establish certain relations among leading terms of this expansion.

math.AP

Four-Entropic Matroids Are Quaternary

For an integer $q\ge2$, a matroid is $q$-entropic if its rank function, multiplied by $\log q$, is the joint-entropy function of random variables on a $q$-element alphabet. We prove that a matroid is $4$-entropic if and only if it is representable over $\F_4$. The corresponding statements for alphabet sizes two and three were known. The proof combines minor closure and the excluded-minor characterization of quaternary matroids with structural properties of quasigroups of order four. Thus arbitrary four-symbol partition representations yield no matroids beyond the quaternary ones. As an application, every access structure admitting an ideal perfect scheme with a uniform four-symbol secret and four-symbol active shares also admits an ideal $\F_4$-linear scheme.

math.CO

Blind Random Search with Noisy Loss Measurements: Averaging, Thresholding, and Almost Sure Convergence

Blind random search repeatedly draws a candidate point and replaces the current estimate whenever the candidate has a lower loss. In the absence of noise, the true loss is observed directly. It decreases strictly at every accepted update and is monotone nonincreasing over all iterations. Measurement noise can make a worse candidate appear better and thereby break this monotonicity. To recover almost sure convergence under noise, we incorporate averaging and thresholding into the original decision criterion. These two classical tools are coupled. As the sample sizes grow, the positive threshold shrinks at a matched rate. These modifications allow blind random search to recover eventual monotonicity of the true loss under noisy measurements and to converge almost surely.

math.OC

Deep learning based numerical approximation algorithms for stochastic partial differential equations

In this article, we introduce a deep learning based approximation algorithm for SPDEs. Our approach employs neural networks to approximate the solutions of SPDEs along given realizations of the driving noise process. If applied to a set of simulated noise trajectories, it yields empirical distributions of SPDE solutions, from which functionals like the mean and variance can be estimated. We test the performance of the method on stochastic heat equations with additive and multiplicative noise as well as stochastic Black-Scholes equations with multiplicative noise and Zakai equations from nonlinear filtering theory. In all cases, the proposed algorithm yields accurate results with short runtimes in up to 100 space dimensions.

math.NA

SafeMath: Safe Solutions for Unsafe Math Word Problems

Recent research points toward LLMs being manipulated through adversarial and seemingly benign inputs, resulting in harmful, biased, or policy-violating outputs. In this paper, we study an underexplored issue concerning harmful and toxic mathematical word problems. We show that math questions, particularly those framed as natural language narratives, can serve as a subtle medium for propagating biased, unethical, or psychologically harmful content, with heightened risks in educational settings involving children. To support a systematic study of this phenomenon, we introduce ToxicGSM, a dataset of 1.9k arithmetic problems in which harmful or sensitive context is embedded while preserving mathematically well-defined reasoning tasks. Using this dataset, we audit the behaviour of existing LLMs and analyse the trade-offs between safety enforcement and mathematical correctness. We further propose SafeMath -- a safety alignment technique that reduces harmful outputs while maintaining, and in some cases improving, mathematical reasoning performance. Our results highlight the importance of disentangling linguistic harm from math reasoning and demonstrate that effective safety alignment need not come at the cost of accuracy.

cs.CL

Coexact completion of profinite Heyting algebras and uniform interpolation

This paper shows that the sheaf representation of finitely generated free Heyting algebras constructed by Ghilardi and Zawadowski can be factored as the profinite completion of Heyting algebras, followed by identifying the dual category of profinite Heyting algebras as a full subcategory of a sheaf topos. We show that the dual category of profinite Heyting algebras is an infinitary extensive regular category, and its ex/reg-completion is exactly the aforementioned sheaf topos, which we refer to as the K-topos. We show how certain properties of uniform interpolation can be generalised to the context of arbitrary profinite Heyting algebras, and that they are consequences of the internal logic of the K-topos. Along the way we also establish various topos-theoretic properties of the K-topos.

math.LO

The Minimum Number of Measurements for Almost-Everywhere Complex Phase Retrieval

Let $d\geq 2$ and let $\bf{f}_1,\ldots,\bf{f}_m\in\mathbb C^d$. We prove that if $m\leq 2d-1$, then the intensity measurement map \[ \bf{x}\longmapsto \bigl( |\langle \bf{x},\bf{f}_1\rangle|^2, \ldots, |\langle \bf{x},\bf{f}_m\rangle|^2 \bigr) \] fails to recover almost every signal in $\mathbb C^d$ uniquely up to a global phase factor. Combined with the known generic sufficiency of $2d$ measurements, our result establishes that the minimum number of measurements required for almost-everywhere phase retrieval in $\mathbb C^d$ is exactly $2d$. This resolves an open problem in phase retrieval by determining the exact measurement threshold for almost-everywhere phase retrieval in ${\mathbb C}^d$.

cs.IT

Frequency-explicit convergence analysis of a multiscale finite element method for highly heterogeneous scattering problems

We analyze the numerical approximation of time-harmonic scattering by highly heterogeneous penetrable obstacles. These problems are especially challenging in the high-frequency regime, where the size of the scatterer $L$ is much larger than the wavelength, i.e., the wavenumber $k$ is such that $kL \gg 1$. Here, we further consider the situation where the scatterer contains different materials, with a characteristic size $\varepsilon$ such that $k\varepsilon \ll 1$. We propose a high-order multiscale finite element method, and provide an error analysis that is explicit in both $k$ and $\varepsilon$. Crucially, our error estimates suggest that using a high-order method should reduce the computational cost for large frequencies, which is corroborated by numerical examples.

math.NA

Differential uniformity properties of some classes of permutation polynomials

The notion of $c$-differential uniformity has recently received a lot of attention since its proposal~\cite{Ellingsen}, and recently a characterization of perfect $c$-nonlinear functions in terms of difference sets in some quasigroups was obtained in~\cite{AMS22}. Independent of their applications as a measure for certain statistical biases, the construction of functions, especially permutations, with low $c$-differential uniformity is an interesting mathematical problem in this area, and recent work has focused heavily in this direction. We provide a few classes of permutation polynomials with low $c$-differential uniformity. The used technique involves handling various Weil sums, as well as analyzing some equations in finite fields, and we believe these can be of independent interest.

cs.IT

Quiver Semistability and Structured Kalman Decompositions for Networked Linear Dynamical Systems

We introduce new notions of controllability and observability for networked linear time-invariant (LTI) systems based on $σ$-semistability of quiver representations. Utilizing King's criterion for $σ$-semistability, we define a network generalization of the Kalman decomposition for networked LTI systems, which systematically decomposes the local and interconnection dynamics while respecting the underlying network structure. Furthermore, we present efficient algorithms for deciding the proposed controllability and observability of a given networked LTI system and for finding the Kalman-type decomposition. We also show efficient algorithms for deciding the $σ$-semistability of representations of acyclic quivers with self-loops if the weight $σ$ has the same sign for all vertices with self-loops. Such quiver representations and weights arise from networked LTI systems.

math.OC

Bellman-sufficient Information Complexity

We introduce Bellman-sufficient information complexity for minimax analysis of sequential decision problems. A Bellman-sufficient state retains enough of the history to close the controlled recursion, while an index $Y=χ(Ω)$ specifies the decision-relevant information being charged. The upper bound is a log-penalized Bellman program; the lower bound is a Bellman--Fano comparison along an algorithm-dependent reference trajectory. If the two values match at a common localization scale and the stated admissibility, calibration, and growth conditions hold, they form an information-risk sandwich. UCB, E2D, and AMS/EBO control or relax the upper Bellman bracket in different ways. For the main application, we give a negative answer to a widely studied form of the GP--UCB minimax-optimality question. For every $0<α<1/4$, we construct one bounded continuous kernel whose minimax regret is $Θ(T^{1-α})$ along an infinite sequence of horizons, while two globally calibrated GP--UCB rules incur linear regret under one fixed truth. An epochwise finite-marginal action-index AIR Bellman policy, implemented through robust AIR/AMS/EBO control, attains the minimax order. The construction separates realized information from the cost of uniform optimism: many low-value directions inflate the exploration multiplier and change the trajectory. Through the canonical RKHS feature map, it also yields a finite-horizon polynomial minimax separation for the specified maximal-information-calibrated LinUCB rule. A reproducible experiment illustrates the mechanism.

cs.LG

ENPINN: Energy-Norm-Guided Gradient-Enhanced PINNs for Generalized Transport Problems with Sharp Gradients

Physics-informed neural networks (PINNs) have emerged as a meshless alternative to conventional numerical methods for solving partial differential equations (PDEs). However, their limited ability to capture sharp gradients can lead to substantial errors when resolving boundary and interior layers. Here, we introduce an energy-norm-enhanced PINN (ENPINN) that incorporates gradient information and variational structure into the loss function to improve the resolution of layer-dominated solutions. We first examine two related formulations: weak-loss PINNs (WLPINNs), which incorporate test functions into the conventional PINN residual, and gradient-enhanced PINNs (gPINNs), which augment the loss with spatial derivatives of the PDE residual. By analyzing these formulations, we identify their limitations in resolving steep solution gradients and motivate the systematic construction of ENPINN. We establish theoretically how the energy-norm error depends on the ENPINN loss and show that a suitably modified residual-derivative term is essential for accurately capturing boundary layers. We further establish the existence of neural-network approximations with arbitrarily small energy error and derive corresponding derivative bounds, providing a theoretical foundation for the proposed framework. The performance of ENPINN is assessed through systematic comparisons with existing PINN variants for convection-diffusion-reaction problems exhibiting steep gradients. Numerical experiments include a combustion model, a coupled multi-scale system, a two-dimensional Burgers equation with an interior layer, and a three-dimensional time-dependent problem.

math.NA

Semi-discrete quadratic Wasserstein energy and state-dependent Langevin exploration

We study the semi-discrete quadratic Wasserstein energy. The energy is nonsmooth at collisions of sites. We prove local Lipschitz continuity on the full configuration space, together with global semiconcavity, coercivity, and dissipativity; show that every global minimizer is interior and collision free; and establish $C^2$ regularity on the collision-free configuration space. The gradient is expressed through the barycenters of the balanced Laguerre cells, while the Hessian is given by an explicit facet formula and satisfies a global one-sided bound. We also solve the one-dimensional problem explicitly in each ordering chamber and give a two-site example on the unit square with non-minimizing Lloyd fixed points. For $d\ge 2$, we then formulate an entropy-regularized relaxed control of the Langevin temperature. The controlled dynamics is strongly well posed, nonexplosive, and collision free. Its value function is a classical interior solution of the exploratory Hamilton-Jacobi-Bellman equation; the Laplacian of the value function is locally $C^1$, which yields a locally Lipschitz optimal temperature feedback. Independently of this optimal-control result, for every fixed Borel temperature rule bounded away from zero, and every sufficiently small step size, the associated Gaussian Euler chain is geometrically ergodic with a full-support invariant law. The raw iterates do not converge, whereas the best-so-far energy converges almost surely to the global minimum and the running record approaches the set of global minimizers.

math.NA

Thermal Recurrence Orders of the Potts Model Partition Function in Grid Graphs

We study the linear recurrence order of the Potts model partition function on thermal 2D grid graphs. By restricting the transfer matrix (TM) to the real coupling axis, the physical operator maintains diagonalizability under the Spectral Theorem. We show that this thermal regularization allows the Krylov subspace to saturate the unconstrained planar state capacity, locking the recurrence order to the Dyck path up to reversal (OEIS A007123) for $q \ge 4$. Furthermore, the recurrence order collapses to height-restricted Dyck paths up to reversal (OEIS A001998) for $q=3$ due to finite-index Jones-Wenzl projections, and to the zero-magnetization conservation sector (OEIS A001405) for $q=2$. This framework bridges graph-theoretic combinatorics with the representation theory of physical loop gas models.

cs.DM

No Equivariant Architecture Covers All Equivariant Attention

We give a complete characterization of equivariant multi-head self-attention (MHSA): if an MHSA layer is equivariant to a symmetry group $G$, then $G$ can only act by permuting head-clusters, with QK and OV matrices satisfying an equivariance constraint tied to the group action. As a consequence, we prove that any fixed MHSA architecture that achieves exact equivariance by polynomially parameterizing unconstrained MHSA parameters inevitably leads to expressivity loss within the class of equivariant maps: the equivariance locus of unconstrained MHSA forms a union of extremely many Zariski-irreducible components in a reduced parameter space, and any single architecture covers at most one. For $G=D_4$ acting on $C$ copies of the regular representation as the token feature space, we show that there are $Ω(C^{64})$ components for eight attention heads.

cs.LG

Rethinking quantum smooth entropies: Tight one-shot analysis of quantum privacy amplification

We introduce an improved one-shot characterisation of randomness extraction against quantum side information (privacy amplification), strengthening known one-shot bounds and providing a unified derivation of the tightest known asymptotic constraints. Our main tool is a new class of smooth conditional entropies defined by lifting classical smooth divergences through measurements. A key role is played by the measured smooth Rényi relative entropy of order 2, which we show to admit an equivalent variational form: it can be understood as allowing for smoothing over not only states, but also non-positive Hermitian operators. Building on this, we establish a tightened leftover hash lemma, significantly improving over all known smooth min-entropy bounds on extractable randomness and recovering the sharpest classical achievability results. We extend these methods to decoupling, the coherent analogue of privacy amplification, obtaining a corresponding improved one-shot bound. Relaxing our smooth entropy bounds leads to one-shot achievability results in terms of measured Rényi divergences, tightening the bounds of [Dupuis, arXiv:2105.05342] and recovering state-of-the-art asymptotic i.i.d. error exponents. We show an approximate optimality of our results by giving a matching one-shot converse bound up to additive logarithmic terms. This yields an optimal second-order asymptotic expansion of privacy amplification under trace distance, establishing a significantly tighter one-shot achievability result than previously shown in [Shen et al., arXiv:2202.11590] and proving its optimality for all hash functions.

quant-ph