arXiv ScienceSearch

SEARCH · arXiv Science

Results for “math.AP”

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

990 records · Page 2Linked to original sources

Sharp Mixed Spectral Barron Regularity of Coulombic Many-Electron Wave Functions

We establish sharp mixed spectral Barron regularity for eigenfunctions of molecular Coulomb Hamiltonians. The mixed norm is a Fourier $L^1$ norm with one isotropic weight and coordinate-product weights, and therefore detects regularity invisible to the isotropic Barron scale. For a nonempty set $I$ of electron indices on which the wave function is antisymmetric, we derive an explicit admissible region for the isotropic order $s$ and the coordinate orders $α,β$. This region is optimal as a uniform statement over the class of clamped-nuclei Coulomb Hamiltonians. For fixed-spin components with two occupied spin blocks, it reduces to $s+α+β<1$; in the fully spin-polarized class it reduces to $s+α<1$. In particular, if $\mathcal I_σ$ denotes the family of occupied same-spin blocks determined by $σ$, then every fixed-spin spatial component $ψ_σ$ satisfies, for every $0\leqα<1$, \[ \left(\sum_{I\in\mathcal I_σ}\prod_{i\in I}\langleξ_i\rangle^α\right)\widehat{ψ_σ}\in L^1(\mathbb{R}^{3N}). \] For a fully spin-polarized state, $\mathcal I_σ=\{\{1,\ldots,N\}\}$.

math.AP

On the Abundance of Critical Points of the t-SNE Energy

This paper considers the energy landscape of the t-SNE algorithm. While this algorithm has enjoyed broad adoption, the non-convexity of the associated energy has made it difficult to rigorously understand what the algorithm captures in many settings. In particular, a number of well-known numerical examples, several of which are reproduced in this article, suggest a complicated energy landscape with many local minimizers that do not respect the topology or clustering structure of the underlying data. This work seeks to provide first steps towards a rigorous explanation of these phenomena. Specifically, for a general family of energies, which include both the original t-SNE algorithm and recently identified large data limits, and for densities in feature space which obey a continuous symmetry, we construct infinite families of distinct critical points. These critical points are based upon identifying pairs of discrete symmetries, one in the original feature space and the other in the target embedding space, which are preserved under gradient dynamics. These critical configurations exhibit many characteristics, such as topology breaking and spurious clustering, which are often observed empirically. Finally, numerical and analytical examples are given throughout as a means of illustrating the approach.

cs.LG

Inverse Source Problem for a Time-Fractional Diffusion-Wave Equation with a Singular Inverse-Square Potential

This paper investigates an inverse source problem for a time-fractional diffusion-wave equation with a singular inverse-square potential. The source term is assumed to consist of a known temporal factor and an unknown spatial component, which is to be recovered from terminal-state measurements. The well-posedness and regularity of the forward problem are established within an appropriate energy framework by exploiting Hardy-type inequalities and the spectral properties of the associated singular elliptic operator. The terminal observation operator is then shown to be compact, and uniqueness of the spatial source is established under a suitable nondegeneracy condition on the temporal factor. To stabilize the resulting ill-posed inverse problem, a Tikhonov regularization approach is introduced. The gradient of the regularized functional is derived through an adjoint problem involving a right-sided fractional derivative, leading to an adjoint-based conjugate gradient method with an exact line search for the numerical reconstruction of the unknown source. Numerical experiments are conducted on both one-and two-dimensional spatial domains, using both exact and noisy terminal data, to demonstrate the effectiveness and stability of the proposed source reconstruction method.

math.NA

Yield Trajectory Tracking for Hyperbolic Age-Structured Population Systems

For population systems modeled by age-structured hyperbolic partial differential equations (PDEs) that are bilinear in the input and evolve with a positive-valued infinite-dimensional state, global stabilization of constant yield set points was achieved in prior work. Seasonal demands in biotechnological production processes give rise to time-varying yield references. For the proposed control objective aiming at a global attractivity of desired yield trajectories, multiple non-standard features have to be considered: a non-local boundary condition, a PDE state restricted to the positive orthant of the function space and arbitrary restrictive but physically meaningful input constraints. Moreover, we provide Control Lyapunov Functionals ensuring an exponentially fast attraction of adequate reference trajectories. To achieve this goal, we make use of the relation between first-order hyperbolic PDEs and integral delay equations leading to a decoupling of the input-dependent dynamics and the infinite-dimensional internal one. Furthermore, the dynamic control structure does not necessitate exact knowledge of the model parameters or online measurements of the age-profile. With a Galerkin-based numerical simulation scheme using the key ideas of the Karhunen-Loève-decomposition, we demonstrate the controller's performance.

math.OC

Computational Oncology of Chemotaxis-Driven Tumour--Immune Spatial Patterning and Stability

We develop a reaction--diffusion--chemotaxis model for spatial tumour--immune--chemokine dynamics that couples logistic tumour growth, immune-mediated killing, chemokine-dependent immune recruitment, chemotactic migration, and signal production. For the nondimensional system, we establish local classical solvability, nonnegativity, a uniform tumour-density bound, and global mass estimates for the immune and chemokine components. The tumour-free equilibrium is stable precisely when the baseline immune-control index satisfies \(σ_0/δ>1\), whereas positive homogeneous coexistence is characterized by a scalar nonlinear equation. Linearization in the Neumann Laplacian eigenbasis yields a mode-dependent cubic dispersion relation, showing that chemotaxis does not alter the tumour-invasion threshold but can destabilize homogeneous coexistence through a finite-wavelength oscillatory instability above a critical sensitivity \(ξ_c\). A conservative finite-volume discretization with upwind chemotactic fluxes and implicit backward differentiation formula time integration is used to test these predictions. Numerical experiments recover the analytical equilibria and growth rates, identify the dominant unstable mode, reproduce the transition to spatial heterogeneity, and quantify the effects of immune recruitment, decay, and diffusion on the stability boundary. Grid-refinement, mass-balance, residual, and nonnegativity diagnostics support the computational reliability of the results.

math.AP

The $α$-Limit Problem: Convergence of a Linear Degenerate Interface Transmission Problem

We study the singular limit of a family of linear degenerate interface transmission problems arising from a regularization procedure in the newly proposed Two-Parameter Diffuse Domain Method (DDM2p). For $α>0$, the regularized problem admits a strictly convex variational formulation on $H^{1}(Ω)$. In the limit $α\to0$, the problem degenerates to a weakly coupled interface system with a nonstandard energy structure. To characterize the limit, we introduce a closed Hilbert subspace $\mathcal{H}\subset H^{1}(Ω)$, defined through an auxiliary Helmholtz problem on an annular subdomain $Ω_2\subset Ω$, and identify the limiting energy functional $\mathcal{E}_{0}$ on $\mathcal{H}$. We prove that the regularized energies $\mathcal{E}_α$ $Γ$-converge to $\mathcal{E}_{0}$ in the strong $L^{2}(Ω)$ topology, using the standard framework. Consequently, minimizers of $\mathcal{E}_α$ converge to the unique minimizer of $\mathcal{E}_{0}$, which is shown to be equivalent to the solution of the limiting interface problem. We further prove strong convergence $u_α\to u_{0}$ in $H^{1}(Ω)$ and establish an $O(α)$ convergence rate. Numerical experiments in one spatial dimension confirm the predicted first-order convergence rate and suggest that this rate is sharp.

math.AP

A Temperature-Coupled Cahn-Hilliard-Stokes-Heat Model for Thermally Driven Phase Separation

We study a diffuse-interface model for thermally driven phase separation in viscous incompressible mixtures. The system couples a convective Cahn-Hilliard equation for the order parameter with a Stokes subsystem for the velocity-pressure field and a heat equation for the temperature. Temperature enters the bulk free energy through a Landau-type coefficient, while the phase field affects the flow through concentration-dependent density and viscosity. The model serves as a proxy for temperature-triggered condensation-like phase separation; humidity, latent heat, vapor pressure, and capillary forcing are absorbed into the choice of the threshold temperature $Θ_S$. We motivate the chemical potential through a temperature-dependent Landau free energy and use a regularized auxiliary formulation to prove local-in-time existence of weak solutions. For the numerical analysis, we employ a first-order sequential finite-element discretization of a simplified quasi-static formulation. The heat equation is advanced by implicit diffusion, the variable-coefficient Stokes problem is treated by a Taylor-Hood discretization, and the Cahn-Hilliard bulk derivative is evaluated at the previous time level, so each algebraic subproblem is linear. An isothermal diffusive test confirms mass conservation to roundoff and exhibits monotone discrete-energy decay for the tested parameters. Time-step and mesh-refinement studies show first-order temporal and approximately second-order spatial behavior. The remaining computations provide qualitative, parameter-specific illustrations; no global discrete energy law is claimed for the non-isothermal sequential scheme.

math.AP

A Tensor Neural Network Method for High-Order Homogenization of Locally Periodic Elliptic Problems

We develop a high-order tensor neural network (TNN) method for locally periodic elliptic multiscale problems of the form $-\nabla\cdot(A(x,x/\varepsilon)\nabla u_\varepsilon)=f$. Because the coefficient depends on both the slow variable $x$ and the fast periodic variable $y=x/\varepsilon$, the high-order cell problems and macroscopic corrector equations are more involved than in the classical case $A=A(y)$, and the correctors depend parametrically on $x$. We derive a computable high-order two-scale expansion and prove an $H^1$ convergence estimate for the partial expansion in boundary-layer-free settings, including periodic domains and ideal boundary-matching configurations. The proof uses the recursive compatibility structure of the corrector hierarchy and a zero-mean oscillation estimate in $H^{-1}$. We then construct a TNN framework for the high-dimensional corrector problems. Its tensor-product structure permits deterministic one-dimensional quadrature for the cell problems, homogenized coefficients, macroscopic source terms, and loss functions, avoiding Monte Carlo integration error. The numerical realization assumes that the coefficient entries and assembled data admit finite or controlled tensor-product representations; this computational assumption is separate from the general matrix-valued coefficient class used in the analysis. Experiments with scalar locally periodic coefficients show accurate high-order correctors. The $H^1$ semi-norm errors are consistent with the proved estimate, while point-normalized $L^2$ errors display the nominal high-order behavior predicted by the formal expansion.

math.NA

Analysis of Moment Closures Using $φ$-Divergences for Rarefied Dynamics with Binary Collisions and Their Galerkin Discretizations

This work introduces a robust deterministic framework for approximating solutions of the Boltzmann equation with binary collisions by discretizing their dependence on time, position, and velocity using Galerkin methods. By employing a family of parametric Galerkin closures based on $φ$-divergences in velocity space, we derive rigorous hierarchies of moment equations that govern fluid dynamic variables. Addressing the limitation that these closures alone do not guarantee dissipation of a $φ$-divergence entropy for the true binary collision operator, we restore this property by formulating a compatible approximate collision operator tailored to each closure. This constructed operator intrinsically retains fundamental physical properties essential for high-fidelity flow simulations, including Galilean invariance, exact conservation of mass, momentum, and energy, and strict dissipation of a $φ$-divergence entropy. Furthermore, we show that the resulting closed moment systems are symmetric-dissipative, yielding Cauchy problems that are well-posed locally in time. To translate this mathematical foundation into an efficient computational tool, we discretize the position and time variables with an entropy-stable discontinuous Galerkin (DG) finite element method. The fully implicit, entropy-stable space-time approach enables time steps far beyond typical CFL-limited step sizes and the direct computation of steady states. The robustness and accuracy of the methodology are verified and validated through numerical simulations on the supersonic nozzle flow of argon, mass flow through a channel, and heat transfer between parallel walls, demonstrating agreement with analytical benchmarks, experimental measurements, and stochastic particle simulations.

math.NA

Assessing Nonlinear Elimination Preconditioning for Trust-Region Phase-Field Fracture

Each quasi-static load step of phase-field fracture is a bound-constrained minimization of a nonconvex, coupled displacement-damage energy under an irreversibility bound on the damage. Monolithic Newton stalls once the nonlinearity localizes at the advancing crack front, and staggered (alternate-minimization) schemes converge slowly there. We present an on-demand nonlinear-elimination preconditioned trust-region Newton method: an energy Steihaug-Toint trust region, a primal-dual active set for irreversibility, and a bound-constrained field-split sweep that eliminates an algebraically-identified "hard set" spanning both fields before each step. The elimination is applied on demand -- triggered by the coupled Newton's own stalling and otherwise skipped -- so the method reduces to monolithic Newton at no surcharge where the step is already healthy. We find the robustness to come from the energy trust region: with that globalization fixed, monolithic Newton already completes every loading history without cutbacks, where residual-merit Newton death-spirals, alternate minimization stalls, and the full-field sweep loses robustness. Against that well-globalized baseline, the on-demand elimination cuts outer nonlinear iterations by 19-25% (brittle) and 17% (ductile), with always-on elimination reaching 26-28% and about $39\%$ at the ductile nucleation step. Measured machine-independently, as a full-mesh-equivalent assembly-work proxy rather than wall-clock, it is competitive with -- not faster than -- monolithic Newton (within about 10%), whereas an always-on sweep adds up to 30%. Nonlinear elimination is thus an iteration-reduction mechanism whose overhead the on-demand gate bounds, with no demonstrated total-work advantage over well-globalized monolithic Newton.

cs.CE

Shape Holomorphy and Sparse Approximation of the Maxwell Electric Field Integral Operator

Uncertainty quantification for time-harmonic Maxwell scattering by obstacles of uncertain shape needs more than holomorphic dependence of the scattered field: for a boundary element method it is the boundary integral operator family itself that must depend holomorphically on the shape parameters. Two obstructions stand in the way. The natural energy space of the electric field integral equation, $\boldsymbol H^{-1/2}_{\mathrm{div}_Γ}(Γ)$, depends on the geometry, and the available operator-valued shape-holomorphy theory for weakly singular kernels is set in $L^2$, which does not reach it. We remove both. A surface contravariant Piola transformation identifies the geometry-dependent Maxwell trace spaces with a fixed reference space, and in the pulled-back variational formulation the surface Jacobians cancel exactly. The principal analytical ingredient is then a uniform fractional mapping theorem $H^{-1/2}\to H^{1/2}$ for the complex-deformed scalar single-layer family on uniformly $C^{1,1}$ surfaces, obtained by realizing the Laplace principal part as the trace of a complex-coefficient Newton problem on a fixed ambient space. The pulled-back operators are consequently $(\bm b,p,\eps)$-holomorphic for $\bm b\in\ell^p(\N)$, $0<p<1$, and pointwise exclusion of interior electric resonances over the compact real parameter set yields uniform invertibility. Legendre coefficients are therefore $\ell^p$ summable, so the operator family, the surface current and the far field all admit sparse polynomial approximations at dimension-independent best $N$-term rates. These statements are for the operator family itself in its energy-space operator norm, not only for individual solutions.

math.NA

Projection-based low-rank assembly in IgA

Isogeometric Analysis (IgA) uses the same spline functions to represent the computational domain and to approximate the solution. This allows exact geometry descriptions, but the resulting mass and stiffness matrices are expensive to assemble and to store, especially in three dimensions. We present a projection-based low-rank approach for assembling the mass and stiffness tensors of orientation-preserving tensor-product B-spline geometries. For the mass tensor, we exploit the polynomial structure of the determinant of the Jacobian of the geometry map and represent it in reduced spline product spaces by univariate coefficient transfer operators; with exact quadrature and without truncation, the resulting low-rank tensor is an exact reformulation of the standard Galerkin mass tensor. For the stiffness tensor, we split the rational weight function into a polynomial numerator, again represented in reduced spline product spaces, and the reciprocal determinant, which is in general not a spline function and is therefore approximated by an $L^2$-projection onto a tensor-product spline space. Both constructions are carried out entirely in the tensor-train (TT) format, with the projection system solved by the alternating minimal energy (AMEn) method, so that full high-order coefficient tensors are never formed and the multidimensional integrals reduce to univariate integrals and contracted products. The method is implemented in MATLAB using GeoPDEs and the TT-Toolbox. Numerical experiments show that it is competitive with full assembly and with the interpolation-based low-rank method, and that it applies in two situations in which interpolation is problematic: a singular interpolation system and nearly singular geometries. The construction is restricted to orientation-preserving tensor-product B-spline geometries and does not cover NURBS.

math.NA

UnifSrv: AP Selection for Achieving Uniformly Good Performance of CF-mMIMO in Realistic Urban Networks

Under the ideal assumption of uniform propagation, cell-free massive MIMO (CF-mMIMO) provides uniformly high throughput over the network by effectively surrounding each user with its serving access point (AP) set. However, in realistic non-uniform urban propagation environments, it is difficult to consistently select good limited serving AP sets, resulting in significantly degraded throughput, especially for the worst-served (formerly "cell-edge") users. To restore the uniformly good performance of scalable CF-mMIMO in realistic urban networks, we formulate a novel multi-objective optimization problem to jointly achieve high throughput by maximizing the sum data rate, uniform throughput by maximizing Jain's fairness index of the throughput per user, and scalability by minimizing the serving AP set size. We then propose the UnifSrv AP selection algorithms to solve this optimization problem, consisting of a deep reinforcement learning (DRL)-based algorithm UnifSrv-DRL and a heuristic algorithm UnifSrv-heu. We conduct a comprehensive performance evaluation of scalable CF-mMIMO under realistic urban network distributions, propagation, and mobility patterns. Our results show that UnifSrv significantly outperforms the prior benchmark AP selection schemes, and for the first time achieves uniformly high throughput of CF-mMIMO under non-uniform urban propagation. Importantly, our heuristic algorithm achieves equivalent throughput to our DRL one, but with orders of magnitude lower complexity. We thus for the first time propose a practical AP selection algorithm that makes CF-mMIMO viable in realistic urban networks.

eess.SY

Tri-Band Channel Measurement-Enabled Multi-Layer Digital Twin for Terahertz Wireless Data Centers

The rapid growth of AI computing has driven increasing demands for flexible and high-capacity data-center interconnections. Owing to its ultra-wide bandwidth and high spatial reuse capability, terahertz (THz) communication has emerged as a promising solution for future wireless data centers, while digital twins (DTs) enable efficient wireless planning and real-time optimization. In this work, a measurement-driven multi-layer DT framework is proposed for THz wireless data centers, where the physical, channel, evaluation, and manipulation layers are progressively constructed from bottom to top. First, extensive channel measurements are conducted at 140, 220, and 300 GHz to characterize frequency-dependent propagation behaviors. Based on the tri-band measurements, a measurement-calibrated physical twin is established by jointly optimizing the geometry, material, antenna, and hybrid propagation models. On top of the physical twin, a line-of-sight (LoS)-aware implicit neural field is developed to construct an AI channel twin for efficient channel reconstruction. The proposed AI twin learns location-dependent channel statistics from the calibrated twin, enabling real-time prediction of received power and LoS probability. Building upon the reconstructed channel field, a system-level evaluation layer is derived to analyze coverage and interference for both AP-to-rack and rack-to-rack communications. Experimental results show that the proposed AI twin achieves lower power reconstruction error than existing neural-field baselines while maintaining real-time inference capability. Moreover, the ceiling-mounted AP deployment achieves over 90% coverage under a 10 dB signal-to-interference-plus-noise ratio (SINR) threshold, demonstrating the effectiveness of the proposed DT framework for THz wireless data-center planning and optimization.

cs.LG

IndicQE-APE: A Benchmark for Quality Estimation and Automatic Post-Editing for Indic Languages

Indic quality estimation (QE) and automatic post-editing (APE) data is spread across separate releases, so no single resource supports training and evaluation across tasks and language pairs on one footing. We consolidate the WMT 2020--2024 shared-task lineage with an extended English--Malayalam resource into \indicqe: $126{,}754$ instances over nine directional pairs, with up to four label types aligned on the same segment, a direct assessment, a human post-edit, word-level OK/BAD tags and an error explanation, and a test set stratified over four difficulty axes. On it we benchmark six prompted LLMs and three COMET metrics on segment-level QE, and three systems on APE. Two of the axes are defined partly on the direct assessment and select a compressed slice of it, so each axis is compared against a control drawn from the same language pair with the same score distribution. Only one survives that control: segments whose holistic and token-level quality signals conflict are ranked worse than equally-scored segments of the same language, for all nine systems and all seven pairs that carry the axis. Annotator disagreement, which looks second-hardest without the control, has no effect with it. Few-shot prompting costs every model $\leq$ $3.4$B both correlation and output-format compliance. Within-language accuracy does not make scores comparable across pairs: of the three trained metrics, the one with the best within-language correlation loses most when the pairs are pooled. The benchmark (https://huggingface.co/datasets/surrey-nlp/IndicQE-APE) and code (https://github.com/surrey-nlp/IndicQE-APE) are released.

cs.CL

Measuring the Installed Base: Nordic Health Dataset Catalogues Against HealthDCAT-AP Release 7

The European Health Data Space requires member states to publish machine readable descriptions of the health datasets available for secondary use, and the European Commission publishes HealthDCAT-AP as the metadata profile those descriptions are meant to satisfy. The profile has been designed and validated against curated examples, never against the catalogues already live. We report that measurement for the Nordic region. On 25 August 2026 the 11 Nordic national catalogues harvested by the European data portal held 2,811 dataset descriptions carrying the EU health theme, and none satisfies all eight properties HealthDCAT-AP Release 7 makes mandatory on a dataset. Three of the eight are present on exactly zero records across five countries. Set beside the portal's own quality assessment, which validates DCAT-AP and never mentions the health profile, this is not a health extension skipped on top of sound generic practice: no Nordic catalogue reaches the assessment's top rating band, 5 of the 11 are reported at zero per cent DCAT-AP compliant, and the properties surviving in both layers are the ones a human types into a form, not the ones needing a value bound to a controlled vocabulary. Two further results follow. Finland contributes 2,259 descriptions to the European portal of which 1,146 carry a theme, and not one uses the EU theme authority vocabulary, so a European health filter returns no Finnish dataset at all. Separately, the authority namespace answers HTTP 200 with a well formed empty document for terms it never defined, letting 1,238 datasets across the wider portal carry theme IRIs that resolve to nothing while passing any status code check. We publish the vocabulary, the shapes and the harvesters, record every verdict as a dated observation rather than a property of the dataset, and report the five errors this discipline caught before publication.

cs.DL

Diffusion Distillation for Efficient Weather Ensembles

Diffusion models generate skillful weather ensembles but require costly iterative sampling. We introduce a supervised energy-distance distillation method that compresses a multi-step diffusion teacher into a single-step student by aligning student forecasts with teacher samples and ground-truth observations. Experiments on global forecasting and typhoon-track prediction show that our student outperforms existing distillation methods and preserves skill for extreme events. It matches or surpasses the teacher across key metrics using only one neural function evaluation per autoregressive step.

cs.LG

Network-Aware Forecasting on Wireless Access Points

Enterprise wireless access points (APs) are promising platforms for predictive machine learning (ML), but their primary responsibility remains providing wireless connectivity and network services. Predictive inference must therefore share an AP's CPU and memory with packet processing, Wi-Fi and IoT radio operations, and client management. This resource contention creates two risks: a model that performs well on proxy hardware may be too slow on the target AP, while a model that fits in isolation may still degrade network services under load. We define \textit{network-aware deployability} using two gates: qualification of the model and its execution path on the target AP, followed by validation of its execution profile under packet-service and forecasting constraints. Our benchmarks show that edge testbeds do not reliably capture target behavior. Across matched artifacts and serving settings, five model implementations run 6.1--19.1$\times$ slower on an AP than on a Raspberry Pi~5, while peak memory usage differs by up to 22\%. Moreover, two forecasting foundation models of similar size differ in AP latency by 19$\times$. When serving a smaller model across 13 parallel streams at a 30~s cadence under network saturation, default execution increases p99 round-trip time (RTT) by 76\% and reduces throughput by 7.06\%. Understanding these trade-offs is essential for live deployment if we aim to use APs for both networking and ML workloads.

cs.NI