arXiv ScienceSearch

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 recordsLinked to original sources

Finite-Modal Realization and Operator-Norm Convergence of a Source-to-Observation Electromagnetic Scattering Green Operator

Source-to-observation operators provide reusable environment-level descriptions for multi-query electromagnetic (EM) prediction and communication-mode analysis. However, in practical multiple-scattering models, these operators are represented with finitely many angular modes, and agreement for selected excitations or between successive truncation orders does not establish uniform accuracy of the full map or reliability of its singular channels. To close this gap, we formulate the environment-induced response as a scattering Green operator on fixed continuous source and observation spaces and derive an exact trace-space factorization that reconstructs the Maxwell scattered field. For fixed, pairwise-disjoint enclosing trace spheres and a well-posed collective problem, nested vector spherical wave function (VSWF) realizations converge in operator norm. A structural bound separates external modal tails from collective-resolvent sensitivity, and operator-norm convergence guarantees uniform convergence of the singular values. We further construct a finite metric core that preserves the nonzero singular values of each finite-order operator and reconstructs matched orthonormal source--field channels without introducing external-support discretization degrees of freedom (DoF) into the spectral problem. Full-wave benchmarks verify the finite-order implementation. A controlled near-resonant two-sphere study shows that adjacent-order agreement can precede resolution of the dominant high-order collective direction. It further shows that only part of the internal amplification appears in externally accessible gains and that resonance promotes a distinct high-order channel pair above an otherwise preserved low-order family. The resulting framework provides a convergent, metric-consistent finite-modal representation of multiple-scattering source-to-observation operators and their accessible channels.

cs.IT

No information transmission through quantum channels above capacity

We show that the capacity of a quantum channel demarcates a phase transition: while reliable transmission below capacity is always possible, any attempt to transmit information above it fails catastrophically. Specifically, we prove exponential strong converse theorems for unassisted quantum and classical communication over arbitrary finite-dimensional memoryless quantum channels. At rates beyond the respective capacity, the entanglement-generation fidelity and the success probability for classical communication decay exponentially with the number of channel uses. This rules out transmission above capacity even when one tolerates arbitrarily large errors. Our proof follows the classical Arimoto strategy, augmented by a crucial new ingredient: integral representations of R\'enyi information measures that lead to asymptotic continuity bounds for R\'enyi capacities.

quant-ph

Equivalence of Fixed-Rank and Rank-One Even-Order Symmetric Tensor Factorization

In the recent work of Barbier, Ko, and the second present author on sublinear-rank symmetric matrix factorization [Math. Stat. Learn. 9 (2026), 1-68], a key result is that, in the Bayes-optimal setting, the large-size limit of the free entropy of the finite-rank spiked Wigner model is the same as in the rank-one case when the signal has centered i.i.d. entries. In this paper, we show that this rank-one equivalence result extends to the case of finite-rank, even-order, symmetric tensor factorization. Moreover, we give a natural reformulation of a hypothesis that was stated in the aforementioned work to be necessary for this result. As in the matrix case, we use information-theoretic identities and replica symmetry to reduce a known multi-dimensional variational formula for the limiting free entropy to its one-dimensional analog. The novelty stems from the fact that said formula involves a replica symmetric potential containing Hadamard (entrywise) powers, rather than squares, of the matrix-valued variational parameter, so the eigenvalue-based approach used in the matrix case must be adjusted.

cs.IT

Structure-Preserving Data-Driven Identification of Port-Hamiltonian Differential-Algebraic Systems

We present a data-driven approach to identifying linear index-1 differential-algebraic pH systems (pH-DAEs) based on input-output measurements. In comparison to the identification of port-Hamiltonian (pH) systems, the algebraic constraint and the index condition pose additional challenges. First, we establish a structure-preserving formulation of the considered pH-DAE class and derive an implicit midpoint discretization that preserves the algebraic constraints and discrete dissipation inequality. We formulate the identification problem as a regularized least-squares minimization problem subject to the pH-DAE dynamics. Exploiting the index-1 structure, we reduce the constrained problem to an unconstrained optimization problem over the system parameters while preserving the port-Hamiltonian structure. Next, we derive an adjoint-based formulation to efficiently evaluate the gradient of the resulting reduced cost functional. This enables us to use gradient-based optimization methods for parameter estimation. Under suitable assumptions on the admissible parameter set, the existence of a minimizer is established. Numerical experiments demonstrate that the proposed approach can identify surrogate pH-DAE systems that accurately reproduce the input-output behavior of reference systems. Further investigations show the approach's potential for identifying reduced-order surrogate models. Cross-validation with independent input signals confirms the predictive capability of the identified models.

math.NA

Topology Obstructs Pure Foundation Neural Quantum States

Foundation models for ground states in spin-1/2 systems are a promising method for problems ranging from quantum chemistry to identifying new phase diagrams. Nearly all such models are currently pure-states that condition on the Hamiltonian's parameters, whose Monte Carlo samples give energy estimates according to the variational principle. In this contribution, we show that this representation is topologically obstructed. For any gapped Hamiltonian family whose ground-state bundle is non-trivial, every continuous normalized state-vector model has zero fidelity with the ground state at some parameter value in the Hamiltonian family. For that value, the energy is at least one spectral gap, $\Delta$, with an $O(\Delta)$ gap in an open-neighbourhood of that point. We show that this is a sufficient no-go also in the case of degenerate ground-state manifolds, time dynamics, and periodic systems with mixed space-time topology, demonstrating these obstructions on one- and two-qubit systems. We discuss how this causes a spike in the fidelity susceptibility, giving a numerical signature of a phase-transition where there is none. We then show that operator-valued models canonically avoid these obstructions and preserve topological information, implying a structural necessity in representation for foundation neural quantum states.

quant-ph

Local gradient neural operator

Field temporal prediction and source identification constitute canonical problems in dynamical systems. Conventional approaches to these problems depend on a thorough understanding of the governing partial differential equations (PDEs). Recently, deep learning, as represented by neural operators, has provided a data-driven paradigm for addressing such tasks. However, most existing global neural operators for PDEs require large training datasets and many learnable parameters, with limited interpretability and generalization. We propose the local gradient neural operator (LGNO) as a lightweight and interpretable alternative for field temporal evolution prediction and source identification in typical mechanical problems. The method builds on priors from nonlinear gradient discretization and uses multilayer perceptron convolutional layers to learn translation-invariant local kernels that resemble discrete stencils. A zero consistent stencil factorization separates coefficient learning from field reconstruction, rendering the learned operators more transparent. For problems with symmetries, network folding shares equivalent components and reduces parameter counts. We evaluate the method on PDE benchmarks covering linear and nonlinear, static and dynamic, and low and high dimensional cases. Results show that LGNO maintains accuracy, parameter efficiency, and rollout stability across these tasks, and further exhibits wide applicability to mechanical problems including diffusion, flow, and quantum phenomena.

cs.LG

Efficient computation of the asymptotics of extensive-rank HCIZ integrals

We study the high-dimensional asymptotics of Harish-Chandra-Itzykson-Zuber (HCIZ) integrals in the extensive-rank regime. The limit of these integrals is governed by a one-dimensional boundary-value hydrodynamical problem originally derived by Matytsin (1994) and rigorously proved by Guionnet and Zeitouni (2002). Despite its wide-ranging applications, explicit solutions to this problem are known only in a few specific cases. In this work, we introduce an efficient numerical scheme based on a particle discretization and prove its convergence to the continuous boundary-value problem for generic boundary densities. We validate our approach against known analytical solutions and apply it to generic densities, uncovering interesting dynamical phenomena. The high-dimensional limit of HCIZ integrals appears in various contexts, from the large deviations of random matrix spectra to the limiting free energy of disordered systems, high-dimensional statistics, and machine learning. As such, our contribution opens the way towards the numerical exploration of a wide range of high-dimensional models that were previously intractable.

math.PR

A counterexample to Kenig's conjecture for the Laplace double-layer operator

Layer potentials provide a classical approach to boundary value problems for Laplace's equation on Lipschitz domains. Kenig's 1994 spectral-radius conjecture for the double-layer operator would ensure operator-norm convergence of the associated Neumann series on mean-zero $L^2$ densities when the boundary is connected. We disprove this conjecture by constructing a bounded simply connected planar Lipschitz domain whose double-layer operator on arclength $L^2$ has essential spectral radius strictly greater than $1/2$. More precisely, for every $t>1/2$ sufficiently close to $1/2$, we obtain such a domain with $\pm i t$ in its Fredholm essential spectrum. The construction starts from smooth graphs whose shapes repeat under translation. In the limit of separated scales, refinement makes solutions of adjoint resolvent equations grow with fixed forcing. The graph slopes remain uniformly bounded. A computer-assisted certificate proves this growth through an inequality for Hermitian $2\times2$ matrices. Its strict margin at $- i/2$ persists at nearby spectral parameters. Normalisation and a Floquet transform then give compactly supported densities with small residuals on the full graphs. We insert rescaled segments of successive graphs into one bounded boundary, where these densities form a weakly null sequence of approximate eigenvectors. The same spectral conclusion holds on a single periodic Lipschitz graph. The certificate combines continuous estimates, exact rational arithmetic and rigorous interval enclosures.

math.AP

An Inverse Problem for Determining the Piston Speed from a Given Lipschitz Leading Shock

We analyze an inverse problem for determining the piston speed and the associated flow field from a prescribed leading shock and the initial data in a shock tube. The gas flow is described by the isentropic Euler equations (i.e., the $p$-system), while the trajectory of the leading shock is prescribed as a given Lipschitz curve. Under an Ole\u{i}nik-type entropy condition on the leading shock, we develop a modified wavefront tracking scheme to construct the flow field behind the shock. This construction enables us to determine the corresponding piston speed and the associated flow field.

math.AP

Convex optimization on moment polytopes: Hadamard mirror descent and efficient algorithms for quantum functionals and other tensor parameters

Convex optimization on polytopes arises in many areas of science. When the polytope is given implicitly or has exponentially many vertices and facets, standard methods may not apply or be ineffective. This is the case for moment polytopes, such as the entanglement polytopes, which play a foundational role in quantum information and algebraic complexity. They give rise to important entanglement measures and tensor parameters such as the quantum functionals, yet general effective methods for computing these quantities have been elusive. In this paper we address this challenge. We develop a first-order framework called Hadamard mirror descent to optimize suitable convex functions over moment polytopes and, more generally, the gradient sets of geodesically convex functions. It operates locally and does not rely on any explicit description of the polytope. Our framework extends mirror descent, an effective and widely used framework for convex optimization, from the Euclidean setting to Hadamard manifolds, and is motivated by a recent work by Hirai, which we interpret as a Hadamard version of mirror flow. Applying the framework to entanglement polytopes yields the first efficient first-order algorithms to compute the quantum functionals, the symmetric quantum functional, and the G-stable ranks, as well as a new direct algorithm for the non-commutative rank.

cs.CC

PPIM: Pennes Physics-Informed Mamba for Heat-Source-Conditioned 3D Bioheat Simulation

Three-dimensional bioheat simulation aims to predict transient temperature distributions in biological tissue and is commonly modeled using the Pennes bioheat equation, which combines thermal diffusion, perfusion-mediated heat loss, and external heat generation. In this study, we consider a controlled 3D Pennes bioheat simulation under a localized heat-source condition inspired by microwave ablation (MWA). To evaluate neural approximation performance, we compare three neural partial differential equation (PDE) solvers under the same controlled simulation: a spatial Fourier-feature physics-informed neural network (PINN), a generic PINNMamba temporal subsequence model, and Pennes Physics-Informed Mamba (PPIM). PPIM builds on the temporal subsequence model by incorporating conditioned heat-source input and Pennes-aware state-space model (SSM) decay initialization. All three neural models are trained under the same conditions with the same Pennes residual, and an explicit finite-difference method (FDM) solution is used only as the numerical reference. In a representative 600~s run, PPIM achieved the lowest MAE, relative $L_1$ error, and relative $L_2$ error among the evaluated neural solvers. Error maps further showed that the remaining PPIM errors were more concentrated near the heat-source region than across the rest of the domain. These results indicate that PPIM is effective for approximating the FDM reference final temperature field in this controlled simulation. The source code is available at https://github.com/muvYun/PPIM.

cs.LG

Variational Continuation for Double Pendulum Periodic Orbits

We present a Hessian-based approach to numerically continue periodic orbits in dynamical systems. A loop (periodic orbit candidate) is parametrized as a Fourier series; a loss function is defined based on the deviation of the loop from the physical differential equations. Unlike previous work relying on hand-derived Jacobians, our method automates the process by leveraging automatic differentiation, a common machine learning technique. The continuation direction can be determined by the flat directions of the loss landscapes (directions with zero eigenvalues), making the search of periodic orbits efficient and guided. Our method is integrator-free, precisely initializes oscillations around unstable fixed points, and efficiently detects orbit family intersections and subharmonic bifurcations. As a demonstration, we present full continuations of periodic double pendulum oscillations from fixed points, showing bifurcations along orbit families and categorizing branches of periodic orbits. In particular, we find periodic orbits where both pendulum masses are never simultaneously at rest, which to our knowledge has been missing in the literature.

cs.LG

Why Multi-Layer Message Passing Works: Completeness Theory for Graph Neural Network Interatomic Potentials

We prove that the Hypergraph Neural Network, an invariant architecture with 3-body message passing, is a universal approximator for potential energy surfaces. Our main contribution is a multi-layer completeness theory. We show that $L$ layers of message passing on sparse, cutoff-based graphs achieve the same representational power as having access to the full $L$-hop neighborhood, provided the configurations are generic, satisfy an overlap condition and a connectivity condition. This provides the first rigorous justification for the common practice of using multi-layer message passing with a per-layer cutoff smaller than the physical interaction range, the setting used by virtually all practical graph neural network based machine-learned interatomic potentials. As immediate consequences, we show that both DPA3 and CHGNet architectures inherit universal approximation.

cs.LG

On the Gram matrix of standard inner products of asymmetrically-weighted Hermite functions

Let A denote the infinite Gram matrix associated with the standard L2 inner product of asymmetrically-weighted (AW) Hermite functions. We derive an explicit representation of its entries and its Cholesky factorization. We further show that this factorization admits a natural interpretation on a scaled Bargmann-Fock basis. An explicit formula for the inverse of A is also obtained. We then consider the corresponding finite Gram matrix and analyze its asymptotic property, as well as that of its Schur complement. The analysis is motivated by numerical methods for plasma physics, in particular Galerkin spectral methods applied to the Vlasov-Poisson (VP) system. As an application, we demonstrate how the derived Gram matrix formulas and asymptotic results can be exploited in the analysis and implementation of a Galerkin spectral method for the VP system.

math.NA

Families of relative periodic orbits in the planar three-body problem via consecutive alignments

Relative periodic orbits (RPOs) are solutions of the three-body problem that are periodic in a uniformly rotating reference frame and, in general, quasi-periodic in inertial coordinates. We present a numerical procedure for computing and continuing one-parameter families of RPOs of the planar Newtonian three-body problem. The method exploits consecutive syzygies, understood here as configurations in which the three bodies are aligned and their velocities satisfy the corresponding symmetry conditions. Matching the positions and momenta at two consecutive alignments reduces the computation of RPOs to a low-dimensional nonlinear problem. Its solutions are then numerically continued, and linear stability is determined from the nontrivial eigenvalues of the rotated monodromy matrix after removing the neutral directions associated with conserved quantities and continuous symmetries. The procedure is applied to several mass distributions and initial configurations, producing families of Poincar\'e, Hill, and binary-type solutions. These families exhibit transitions from nearly circular to highly eccentric motion, changes of stability near resonances and turning points, and absolute periodic solutions when the rotation angle is a rational multiple of 2{\pi}. In the Hill families, the continuation connects satellite configurations with circumstellar motion as the smallest body loses its gravitational binding to the intermediate body. Circumbinary and circumstellar configurations are also obtained in the binary regime. The results illustrate the dynamical diversity of RPOs and provide coherent three-body motions that can be used as prescribed trajectories in restricted four-body models.

math.DS

A Human-AI Theorem Connecting Spontaneous and Field-Induced Mechanisms of Collective Behavior in One Dimension

Can an artificial intelligence (AI) generate a scientific hypothesis outside a human collaborator's active hypothesis space (AHS), and can human-AI research be organized to make such breakthroughs more likely? We document such a case while proving a theorem that connects two basic organizing mechanisms of statistical physics: collective behavior arising in zero field from competing interactions and that induced or controlled by an external field. A zero-field $O(n)$-vector open chain with arbitrary inhomogeneous nearest- and next-nearest-neighbor interaction functions $U_i(S_i\cdot{S}_{i+1})$ and $V_i(S_i\cdot{S}_{i+2})$ is microscopically, via a temperature-independent mapping at the Hamiltonian level, equivalent to a simpler $O(n)$ open chain with nearest-neighbor interaction $V_i( \sigma_i\cdot \sigma_{i+1})$ and axial single-spin potential $U_i(\sigma_i^z)$ for every integer $n\ge1$ and every system size $L\ge1$. The homogeneous linear specialization maps the foundational frustrated $J_1$-$J_2$ model onto the canonical $J$-$h$ field model---with $n=1,2,3$ being the Ising, XY, and Heisenberg classical spin models, respectively. An analogous theorem holds when the continuous $O(n)$ spins are replaced by the $q$-state Potts spins with the standard Potts interaction, implying a closed-form exact solution of the $J_1$-$J_2$ Potts open chain for every $q\ge2$ and every $L\ge1$. The emergence of the theorems from sustained human-AI collaboration suggests that involving AI throughout a systematic research program may incubate autonomous scientific breakthroughs.

cond-mat.stat-mech

ENPINN: Energy-Norm-Guided Gradient-Enhanced PINNs for Generalized Transport Problems with Sharp Gradients

Physics-informed neural networks (PINNs) have emerged as a meshless alternative to conventional numerical methods for solving partial differential equations (PDEs). However, their limited ability to capture sharp gradients can lead to substantial errors when resolving boundary and interior layers. Here, we introduce an energy-norm-enhanced PINN (ENPINN) that incorporates gradient information and variational structure into the loss function to improve the resolution of layer-dominated solutions. We first examine two related formulations: weak-loss PINNs (WLPINNs), which incorporate test functions into the conventional PINN residual, and gradient-enhanced PINNs (gPINNs), which augment the loss with spatial derivatives of the PDE residual. By analyzing these formulations, we identify their limitations in resolving steep solution gradients and motivate the systematic construction of ENPINN. We establish theoretically how the energy-norm error depends on the ENPINN loss and show that a suitably modified residual-derivative term is essential for accurately capturing boundary layers. We further establish the existence of neural-network approximations with arbitrarily small energy error and derive corresponding derivative bounds, providing a theoretical foundation for the proposed framework. The performance of ENPINN is assessed through systematic comparisons with existing PINN variants for convection-diffusion-reaction problems exhibiting steep gradients. Numerical experiments include a combustion model, a coupled multi-scale system, a two-dimensional Burgers equation with an interior layer, and a three-dimensional time-dependent problem.

math.NA

Algorithmic threshold for high-dimensional projection pursuit I: general theory

We study a null model of high-dimensional projection pursuit: we are given $M$ points sampled i.i.d. from a standard gaussian in $N$ dimensions, where $M,N\to\infty$ with $M/N\to\alpha\in(0,\infty)$. Our goal is to characterize the possible empirical distributions of these points' projections along a data-dependent direction $x$, which ranges over either the sphere $S_N=\sqrt{N}\mathbb{S}^{N-1}$ or cube $\Sigma_N=\{-1,+1\}^N$. We consider this problem in an algorithmic setting, where $x$ must be the output of an algorithm with dimension-free Lipschitz dependence on the input; this class of algorithms includes general gradient-based methods such as Langevin dynamics and approximate message passing (AMP). Our main result exactly characterizes the set of empirical distributions attainable by this class in terms of a one-dimensional stochastic control problem. As a consequence of our main result, we obtain exact algorithmic thresholds for optimizing the Hamiltonian of a spherical or Ising perceptron model with general bounded continuous activation. For the spherical problem, independent work of Montanari and Zhou (2024) characterized the empirical distributions attainable by a related two-stage AMP algorithm, also in terms of stochastic control. Our proof of hardness builds on the branching overlap gap property introduced in earlier work by the first two authors. Our main innovation is to develop stochastic control theory within the branching OGP framework, significantly expanding the settings in which it locates an exact algorithmic threshold. Notably, our methods apply even though the non-algorithmic problem of characterizing all feasible projections remains a major outstanding challenge. For the matching algorithmic result, we construct a new incremental AMP algorithm that acts on a Brownian-bridge revelation of the gaussian disorder and simulates the same family of controlled SDEs.

math.PR