arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,081 records · Page 60Linked to original sources

Contact-resolved deployment of the Contour Neurovascular System in patient-specific intracranial aneurysms

While intrasaccular flow disruptors are increasingly used to treat wide-necked intracranial aneurysms (IAs), many patient-specific computational workflows prescribe a pre-seated device geometry and omit deployment mechanics. This simplification is particularly restrictive for the Contour Neurovascular System (CNS), for which neck coverage, wall apposition, and post-contact motion depend on the deployment process. To address this, we present a contact-resolved finite-element framework that models CNS deployment within patient-specific IAs. We represent the device as a dual-layer interwoven Nitinol braid using geometrically exact beam models, and the aneurysm wall as a deformable hyperelastic shell. Frictional contact governs wire-wire and wire-wall interactions during staged release. The framework is demonstrated across three patient-specific anatomies. In the reference anatomy, the seated configuration is sensitive to the assumed device-wall tangential friction and to release depth relative to the aneurysm neck plane. Frictionless wall contact permits pronounced post-contact sliding, whereas finite tangential friction strongly suppresses residual pole motion. Release depth alters both the onset of wall engagement and the subsequent deployment path. Geometric placement approaches that prescribe the implanted configuration cannot recover this contact history or the associated wall-supported state. Our framework provides a mechanics-based route to deployed geometries for downstream hemodynamic, fluid-structure interaction, and mechanobiological analyses.

physics.comp-ph↗

Right q-vector calculus at integral superdimension: localized decompositions and resonance

We specialize the intrinsic right $q$-vector derivative on radial algebras to integral superdimension. The formal dimension is encoded by an independent coefficient $Q$, and formal radial superspace of superdimension $M=m-2n$ is obtained by the coefficient specialization $Q\mapsto q^M$. This gives a rigorous universal calculus and, on finite blocks containing at most $m$ abstract vectors, a faithful coordinate realization on $\mathbb R^{m|2n}$. The localized exterior result is formulated as a Green decomposition by complementary projector images. Whenever the specialized finite determinant is nonzero, full left multiplication yields a determinant-localized right-monogenic Fischer decomposition. Beyond this base-change theory, we determine the exceptional one-vector calculus completely: for $M=-2\ell$ there is one additional singular monomial and one missing image monomial, whereas all other integral superdimensions give a surjective derivative with constants as its kernel. We then prove that the degree-zero Fischer operator is exactly diagonal on exterior blades, obtain its determinant explicitly, classify all support-resonance values in $0<q<1$, and give an exact kernel-rank formula as a sum of support multiplicities. On a block with $N$ auxiliary vectors, an even support rank $p$ has pure multiplicity $\binom Np$ on the support truncation. At an odd-support root, every lower odd factor is nonzero and any simultaneous lower resonance is unique, even, and characterized by one strictly monotone scalar equation. These results distinguish persistent nonpositive-even-superdimension defects from isolated support-dependent $q$-resonances. An appendix records that constant scalar projection of two independent orthogonal right $q$-vector derivatives does not descend to the Hermitian quotient.

math.CV↗

Lattice Boltzmann Methods for Navier-Stokes Equations in General Orthogonal Coordinates for Efficient Flow Simulations using Nonuniform Clustered Grids

Resolving multiscale fluid flows or boundary layers effectively requires the use of nonuniform meshes with local grid clustering. The standard lattice Boltzmann method (LBM), a kinetic theory-based approach for computational fluid dynamics, however, is restricted to the use of uniform Cartesian grids. We present new and improved formulations of the LBM that accommodate continuously varying spatial grids via coordinate transformations to simulate the Navier-Stokes equations (NSE) in the general orthogonal coordinates (GOC). They are constructed using a Chapman-Enskog analysis to specify the equilibrium moments of the distribution functions and the geometric force terms used in the collision step to be dependent on the local metric factors and their spatial derivatives, along with the density, momentum and their fluxes, and some correction terms related to the normal velocity gradients so as to accurately represent the NSE in the GOC. The resulting GOC-LBM importantly maintains the simplicity of the collide-and-stream approach and is Galilean invariant that is free of the cubic velocity artifacts. Our GOC-LBM is general and modular in that it can be used with any collision model with appropriate modifications to the equilibria and forcing terms. We present its implementation details for a variety of collision models while the central moments-based model using multiple relaxation times was found to be the most robust in practical implementations. We validate the GOC-LBM through numerical simulations for various benchmark flow problems. Moreover, we demonstrate significant computational advantages of our approach for a case study on simulating boundary layer flows efficiently that involves coupling the GOC-LBM for the NSE with a new GOC-LB scheme for solving the magnetic induction equation for magnetohydrodynamics (MHD), and for another case study involving orthogonal curvilinear grids.

physics.flu-dyn↗

Lieb-Thirring bounds for Melik-Adamyan canonical Hamiltonians

We study a class of positive matrix Hamiltonians arising from the canonical differential expressions of Melik--Adamyan and appearing in the appendix of Alpay--Gohberg. Let $J$ and $B$ be self-adjoint involutions on $\mathbb C^{2n}$ satisfying $JB=-BJ$, and let $H>0$ satisfy $HJH=J$. For $m>0$ we consider $$ \mathcal A_{m,H}=H^{-1}\left(-iJ\frac{d}{dt}+mB\right) $$ in the weighted space $L^2_H$. A locally absolutely continuous $J$-unitary gauge $Θ$ representing $H$ reduces this expression to the free massive Dirac operator plus the Hermitian coefficient $$ P_{m,Θ}=-iΘ^*JΘ'+m(Θ^*BΘ-B). $$ Whenever this coefficient belongs to $L^2$, the corresponding self-adjoint realization, including its operator domain, is independent of the chosen representing gauge. Minimizing $\int\mathrm{Tr}|P_{m,Θ}|^2$ over the gauge fibre defines an intrinsic energy. A two-sided Birman--Schwinger decoupling, combined with a truncated pseudo-relativistic estimate proved here, gives a $3/2$-moment bound for all eigenvalues in the gap $(-m,m)$ in terms of this energy. The Dirac estimate applies to arbitrary Hermitian matrix coefficients in $L^2$ and requires no sign condition. On the half-line we treat every self-adjoint Lagrangian boundary condition. Two reflection-compatible conditions require no endpoint correction, while an arbitrary condition contributes at most $2nm^{3/2}$. At zero mass, the optimal-gauge energy is computed explicitly in terms of $H^{-1/2}H'H^{-1/2}$. For a scalar hyperbolic-rotation family the massive gauge minimization reduces exactly to a one-dimensional phase functional. We prove existence of a minimizer in the principal phase sector and give an explicit trial phase that strictly and quantitatively improves the positive lift whenever the corresponding first variation is nonzero.

math.SP↗

SechKAN: Kolmogorov-Arnold Networks with Hyperbolic Secant Functions

In recent years KolmogorovArnold Networks KANs have attracted increasing attention due to their effectiveness in machine learning and scientific computing offering a new paradigm for neural network design In this paper we present SechKAN a novel KAN based on hyperbolic secant sech functions The hyperbolic secant basis is adopted for its smooth bellshaped form localized responses and wellbehaved gradients We employ a 1D linear projection to reduce the number of parameters allowing SechKAN to maintain a model size comparable to that of multilayer perceptrons MLPs Experimental results show the effectiveness of SechKAN on function fitting PDE surrogate modeling and image classification benchmarks including MNIST FashionMNIST CIFAR10 and CIFAR100 On function fitting SechKAN achieves performance comparable to both MLPs and representative KAN variants On PDE surrogate modeling it outperforms MLPs and achieves competitive or better performance than representative KAN variants On image classification benchmarks SechKAN achieves the best performance among the evaluated KAN variants while remaining competitive with MLPs using a comparable number of parameters However SechKAN still incurs higher computational cost than MLPs and some KAN variants Our source code is publicly available at https://github.com/hoangthangta/All-KAN.

cs.LG↗

No Prime Tiling of an Isosceles Triangle

No isosceles triangle can be cut into a prime number (greater than three) of congruent triangles, and only an equilateral triangle can be cut into three congruent triangles.

math.MG↗

The quantitative non-unique-product landscape at the global minimum: the Nielsen-Soelberg groups

Nielsen and Soelberg proved that a finite subset $A$ of a torsion-free group with $A\cdot A$ having no unique product satisfies $|A|\ge 8$, and exhibited two groups, here $G_1$ and $G_2$, attaining the bound. Nothing quantitative was known about these extremal configurations. We construct exact, independently verified models of both groups and compute the first quantitative invariants at the global minimum. In $G_1$ no $8$-element symmetric witness lies in the radius-$6$ ball ($933$ elements, certified infeasible), while the Nielsen-Soelberg witness lies in the radius-$7$ ball: the global minimum is spread out. In $G_2$, with its natural eight-generator metric, the witness and its inverse are the only two non-UP $8$-sets in the radius-$1$ ball, and the unique-product staircase takes the value $0$ at $n=8$ but $1$ at $n=9$ -- the first known minimizer whose square has exactly one uniquely represented element, so the simultaneous failure of t.u.p. and u.p. seen in the Promislow group is not universal. No $(7,9)$ two-sided witness exists in the searched balls, so the Nielsen-Soelberg profile bound may not be sharp. Finally we treat the universal group $G_3$. Its structure is known -- Soelberg's thesis identifies an index-$8$ Heisenberg subgroup of step $8$ and proves torsion-freeness, and Gardam, studying the same group as an amalgam of Klein bottle groups, shows it to be virtually nilpotent but not virtually abelian -- and what we add is a model in search coordinates in which balls can be enumerated. In it we reproduce the Nielsen-Soelberg two-sided pair and exhibit a symmetric $15$-element witness whose trivial-coset singleton generates the centre of that Heisenberg subgroup. It is rigid and rare: within $B(5)$ the size $15$ is exactly minimal, the coset profile is forced, and exactly four such witnesses exist in $B(4)$, one orbit. Hence $m_1(G_3)\in[8,15]$ against $m_2(G_3)=16$.

math.GR↗

Li--Yorke Chaos Along Any Infinite Sequence: Relative Mixing, Sofic and Rokhlin Entropy versus Naive Entropy

Let $G$ be a countably infinite discrete group and let $ π:(X,μ,G)\to(Y,ν,G) $ be a nontrivial relatively mixing extension, where $X$ is a compact metrizable $G$-space. We prove that there exists a constant $δ>0$ such that, for every injective sequence $(s_i)_{i\geq 1}$ in $G$, there is a Cantor set $K_{(s_i)}\subseteq X$ whose distinct points $x,x'$ satisfy \[ \liminf_{i\to\infty}ρ(s_i x,s_i x')=0, \qquad \limsup_{i\to\infty}ρ(s_i x,s_i x')>δ. \] The method also yields higher-order scrambled Cantor sets. As a principal application, for a sofic group $G$, positive topological sofic entropy implies the preceding conclusion, answering a question of Huang, Li, and Ye. The same conclusion also holds for actions of arbitrary countably infinite discrete groups admitting an essentially free invariant measure of positive Rokhlin entropy. In contrast, we construct a hereditary subshift over a residually finite group with positive naive topological entropy for which all points converge to a common fixed point along a prescribed injective sequence. This answers a question of García-Ramos and Li negatively even for sofic groups.

math.DS↗

Exceptional supersphere integration and logarithmic Pizzetti kernels

We study orthosymplectically invariant supersphere integration at the exceptional superdimensions $M=-2u$, where the harmonic Fischer structure becomes nonsemisimple and the Pizzetti pairing degenerates. For the meromorphically continued homogeneous inverse kernels we obtain the generating function $$ \mathscr G_μ(ρ;x,y) =\frac{Γ(μ/2)}{2π^{μ/2}} \bigl(1+ρ\{x,y\}+ρ^2x^2y^2\bigr)^{-μ/2}. $$ At $μ=-2u$, its Laurent expansion has a polynomial residue and a logarithmic finite part. We prove that these coefficients recover the complete degreewise duality structure on a fixed superspace with nonzero bosonic dimension. In degrees $k\le u$, the residue inverts a canonical renormalized pairing on $\mathcal P_k$. In the collision range $u<k\le2u$, the ordinary pairing has radical $(x^2)^{k-u}\mathcal P_{2u-k}$; the finite part reproduces the quotient, while the residue reproduces the radical after transport from the reflected degree. For $k\ge2u+1$, the finite part is the ordinary inverse kernel. We also establish the nondegenerate head--socle pairing on the generalized harmonic modules. As an application, we derive covariant right--left radial $q$-monogenic zonal symbols and identify precise degree-one obstructions to transferring scalar Pizzetti reproduction through a one-sided $q$-Fischer projection.

math.CV↗

An Evaluation Framework for Structured Audio Captions Validated by Controlled Perturbations

Recent advances in automated audio captioning (AAC) are driving a shift from monolithic sentences toward structured formats that disentangle acoustic and semantic properties, such as timestamped captions for different sound events. Such representations can support faceted sound search for creators and richer access to auditory information for Deaf and Hard of Hearing people. Yet, it remains unclear how to meaningfully evaluate these hybrid, structured captions. We propose an evaluation framework for structured audio descriptions, spanning five complementary axes: tag sets, descriptions, reasoning, numeric measurements, and spectral profiles. The framework combines large language model (LLM) judges for semantic fields with deterministic metrics for temporal and acoustic attributes. To validate these metrics, we introduce controlled perturbations that apply typed, graded changes to ground-truth annotations. Results show that the proposed metrics remain robust to meaning-preserving paraphrases while responding to genuine semantic and acoustic corruptions, enabling more reliable evaluation of structured captions.

cs.CL↗

Disentangling mixed neutron fields: multi-source identification from few detected events

Identifying neutron-emitting materials is central to nuclear nonproliferation, safeguards, nuclear forensics, and emergency response, yet remains difficult when several sources contribute simultaneously: relevant fission, $(α,\text{n})$, and fusion sources emit broad, strongly overlapping energy distributions, and the associated spectral inversion is severely ill-conditioned. Here we demonstrate quantitative identification of mixed neutron fields directly from scatter-based (recoil) spectroscopy measurements, together with simultaneous estimation of the emission rate of each contributing source and a rigorous statistical confidence level for every candidate source combination. Using a compact $21.6\,\mathrm{cm}^3$ organic-glass scintillator spectrometer, we correctly identify Cf-252, a deuterium--deuterium (DD) neutron generator, and their mixture with decisive statistical support ($>\!4σ$), and further resolve a weak deuterium--tritium contaminant in the nominal DD generator field. High-fidelity Monte Carlo simulations spanning exhaustive single-, two-, and three-source mixtures show that identification requires remarkably little information: between $\mathcal{O}(10^1)$ and $\mathcal{O}(10^6)$ detected recoil events, set primarily by spectral similarity, mixture complexity, and emission-rate imbalance. For the compact spectrometer used here, this corresponds to acquisition times as short as a few minutes. These results substantially extend the operational reach of simple single-volume neutron spectrometers, enabling rapid, quantitative, and confidence-calibrated attribution of complex neutron fields in field-deployable instruments.

physics.ins-det↗

Identifying the Sign of Coherent Over-Rotations with Logarithmically Many Pauli Settings

Calibrating a quantum processor means estimating gate-error parameters from data, and a hard-to-estimate parameter is usually assumed to leave a weak signature more repetitions will resolve. Coherent over-rotations break that premise. For commuting single- and two-qubit transverse over-rotations with known support on a computational-basis input, the passive histogram is exactly invariant under a sign group acting on the coherent angles, of order $2^{n+1}$ for the complete family with $n\ge2$, so no estimator resolves the signs uniformly at any sample size. A calibration correction applied with the wrong sign doubles the rotation error it should remove. At zero angle the obstruction turns continuous, with a Fisher information singular along every coherent direction and an infinite Cramér-Rao bound. At generic angles the continuous defect clears for $5\le n\le8$, the range we compute, while the sign degeneracy persists. Covering the support removes both, and a design resolves the signed angles uniformly exactly when it covers. On $|η_R|<π/2$, $0\le p_R<1/2$ coverage is constructive, returning each angle in closed form from a ratio of two measured expectations without knowing the rates. A twirl-free code of $\lceil\log_2(n+1)\rceil$ added product-Pauli settings covers, and for the complete family none uses fewer. Conditioning then sets the Cramér-Rao cost, and we give its floor explicitly. Exact computation matches the theory, the passive fit stays at chance on the signs at every budget while the closed-form inverse recovers them, and angle magnitudes on two IBM Heron processors follow the conditioning ordering.

quant-ph↗

A Multi-level Information Integration Framework for Physically Verifiable Fault Diagnosis of Rotating Machinery

Integrating multi-level information, from physical models through data-driven diagnostics to natural language reasoning, into verifiable decision chains is a growing need in intelligent manufacturing. In bearing fault diagnosis, taken here as a representative testbed, the standard output is a class label and a confidence score derived from the classifier's own distribution, offering limited means of comparison against independent physical knowledge. Meanwhile, language models increasingly used for maintenance communication may introduce unsupported content. This work addresses both limitations from the output side. The proposed Diagnostic Evidence Network (DENet) is an encoder-agnostic multi-task framework that extends the output to a structured evidence record: the classification, a predicted characteristic frequency comparable against the theoretical value determined by bearing geometry and shaft speed, and a temporal localization of transient impulses inspectable on the raw waveform. Across four encoders and three public datasets, this evidence incurs no statistically significant accuracy cost, with a frequency error of about 6 Hz on 1,024-point segments. The deviation between predicted and theoretical frequency constitutes a label-free, inference-time validation signal. It detects misclassifications with AUROC of 0.970 and 0.871, and retains separation within the high-confidence subset. Finally, a QLoRA-adapted language model renders DENet's evidence into traceable maintenance reports without contributing diagnostic decisions, reducing unsupported-claim rates from 10-12% to 2% with no fabricated quantities observed.

cs.LG↗

No Free Lunch in Flow Surrogates under Time-Varying Boundary Conditions: A Two-Regime Study

We test whether an architecture that succeeds on a simple flow regime also succeeds on a richer one, with each trained separately on each regime. We explore two transient flows under time-varying boundary conditions: the three-dimensional slurry film in chemical-mechanical planarisation (CMP), central to semiconductor manufacturing, and the two-dimensional Kármán vortex street (KVS). Eight surrogate models on one shared pipeline differ in whether they learn the full field or a latent representation, and in whether they predict in one shot or step by step. No single architecture wins both regimes. On the film, a one-shot full-field model reconstructs the cumulative wall shear stress to 2.7% relative error. On the wake, a latent autoregressive DeepONet retains 90% of the shedding power that direct and one-shot models damp to almost zero. The treatment of time decides the outcome. The self-sustained wake calls for autoregressive feedback and the boundary-driven film for a direct map. Pointwise RMSE hides the damped oscillation on the wake, compresses the sixfold lead on the film's process target, and picks the damped model under wake extrapolation. The evaluation scores five physical questions. Trained surrogates answer queries 10^3 to 10^4 times faster than the finite-element solver and pay off from the first query beyond the training set on the film and from the third on the wake. Neither the winning architecture nor its validation holds across regimes. The choice of surrogate should follow the dynamical character of the target flow, and its validation should resolve the failure modes.

math.NA↗

CONSISTRE: A Unified Consistency-Aware Framework for Document-Level Relation Extraction with Large Language Models

Document-level relation extraction (DocRE) aims to extract relations among multiple entities across extended contexts while maintaining consistency across predicted triples. Although large language models (LLMs) show remarkable reasoning capabilities in information extraction, their predictions are typically generated independently for each candidate triple and may violate fundamental relational constraints such as transitivity, symmetry, and functional uniqueness, leading to contradictory and unreliable outputs. We propose CONSISTRE, a unified consistency-aware framework for DocRE that addresses this limitation through two complementary tracks. The first operates at inference time for black-box LLMs, combining constraint-aware prompting, constraint-based verification, and iterative self-reflection to refine predictions without task-specific fine-tuning. The second injects consistency knowledge into smaller open-source models via a knowledge distillation and reinforcement learning pipeline: reasoning traces from a powerful teacher are distilled into a student via supervised fine-tuning, followed by GRPO alignment using a composite reward that jointly optimizes extraction performance and relational consistency. Together, the two tracks cover both API-accessible and locally deployable scenarios under a unified consistency formulation. Experiments on DocRED show that both tracks outperform their baselines, with the inference-time track achieving competitive F1 using off-the-shelf black-box LLMs and the training-time track substantially narrowing the gap between 7--8B open-source models and state-of-the-art proprietary LLMs at a fraction of their inference cost. Ablation studies confirm that explicit consistency modeling mitigates relational contradictions and enhances the reliability of LLM-based DocRE across both deployment paradigms.

cs.CL↗

A Bogoliubov-ratio framework for quantum-information diagnostics of time-dependent two-mode Boson System

We formulate a compact dynamical representation of quantum-information diagnostics for time-dependent two-mode bosonic systems in terms of the Bogoliubov ratio $λ_k(η)=β_k(η)/α_k(η)$. For vacuum-evolved pure Gaussian pair states generated by a Hermitian quadratic Hamiltonian, $λ_k$ obeys a closed complex Riccati equation. After one partner mode is traced out, the reduced-state spectrum is determined by $q_k=|λ_k|^2$, from which the purity, linear entropy, Rényi-2 entropy, and von Neumann entropy follow directly. These Gaussian results are mathematically equivalent to those obtained from the standard density-matrix or covariance-matrix formalism; the advantage of the ratio representation is a compact separation between model-dependent dynamics and universal state diagnostics. Writing $λ_k=\sqrt{q_k}e^{iθ_k}$ further exposes how pair production drives radial growth, whereas frequency rotation and detuning act through the phase and can suppress coherent squeezing accumulation. We also establish an exact extension to non-Gaussian $SU(1,1)$ sectors. For a lowest-weight Fock seed with conserved number difference $d$, the same ratio equation governs the evolution, while the reduced spectrum becomes a negative-binomial distribution determined by $(q_k,d)$. The Gaussian formulas are recovered at $d=0$. Cosmological perturbations and a chirped-pulse optical parametric amplifier illustrate the common dynamical mechanism and its information-theoretic consequences.

quant-ph↗

Global Convergence of DGM and PINN Algorithms for Solving Nonlinear PDEs

The Deep Galerkin Method (DGM) and Physics Informed Neural Networks (PINNs) have become widely-used methods for solving partial differential equations (PDEs) in the rapidly growing field of scientific machine learning. In these methods, a neural network is trained to approximate the PDE solution by using (stochastic) gradient descent to minimize the PDE residual of the neural network. Due to the non-convexity of the PDE residual objective function, the trained neural network may, in principle, only converge to a local minimizer of the objective function (which would not be a solution of the PDE). Therefore, there is a longstanding question regarding the mathematical foundations of these algorithms, and it is highly valuable to establish that the trained neural network will converge to the PDE solution. In this paper, we consider a class of semilinear PDEs with nonlinearities in the solution and its first derivative. For this class of PDEs, we prove that neural networks trained with gradient descent to minimize the PDE residual objective function will converge to the PDE solution as the network width and training time $\rightarrow \infty$.

cs.LG↗