arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 955 records · Page 53Linked to original sources

Congruence Classes of Supporting the Erdös-Straus Conjecture II: Wild Solutions

In 1948, Erdös and Straus formulated a conjecture : for any positive integer $n>2$, there exist positive integers $n_1,n_2$ and $n_3$ such that \begin{equation}\frac{4}{n}=\frac{1}{n_1}+\frac{1}{n_2}+\frac{1}{n_3},\nonumber\end{equation} which is still open. It is known that the conjecture holds if one can prove it for any prime $n\equiv 1\;(\mbox{mod}\;24)$. If $n=24m+1$ and $n_1\leq n_2,n_3$, then $n_1=6m+k$ with $1\leq k\leq 12m$. A solution $(n_1,n_2,n_3)$ of the above equation is called a {\it tame solution} if $n_2$ and $n_3$ are factors of $(6m+k)(24m+1)$. We call $n=24m+1$ {\it wild} if it does not have any tame solution. Based on the information in our earlier work on tame solutions posed in arXiv, Howerton found that there are only fourteen wild primes of the form $n=24m+1\leq 2.4\times 10^{11}$. In this paper, we derive thirty-four families of wild solutions of the above equation, which contain the solvability of the fourteen wild primes. Together with our earlier tame polynomial solutions, numeric test shows that they cover all the primes of the form $24m+1$.

math.NT↗

Policy as Code: A Coroutine-Bridge Harness for Fast-Reasoning Reliability on CAR-bench

CAR-bench evaluates whether tool-using agents stay reliable under real-world uncertainty, executing every tool inside the evaluator so that each tool-result exchange is a separate agent round-trip. A conventional next-action agent can batch parallel tool calls, but a chain of dependent calls costs it one model call per round of results. We present a coroutine-bridge harness in which the model's only action is to emit a Python program that blocks and resumes in place across evaluator tool exchanges. This decouples model invocation from tool round-trips: on the public test split the agent uses a median of two model calls against seven agent turns per task, resolving a full multi-turn task in a median of 1.8 s of model latency on Cerebras gpt-oss-120b. Because the action surface is executable code, deterministic CAR-bench policies are encoded directly as logic in the tool layer rather than as prompt rules, enforcing compliance at zero reasoning cost. On the official hidden evaluation the harness won Track 2 with 60.0% Pass^3, 4.5x the organizer baseline, at the lowest estimated cost and the fastest median task latency (3.14 s) of any entry scoring above that baseline; the same unchanged harness reproduced an identical 60.0% Pass^3 on GPT-5.5 in the Open track, matching frontier-model agents. A single static prompt, appended with per-task state at the tail, stays byte-identical across calls and across tasks: the frozen submission prompt served 78% of input tokens from cache (86.6% across its warm tail), against 73% over a three-week development corpus in which prompt edits repeatedly reset the cache. This compounds the few-call design into a small fraction of nominal input compute.

cs.AI↗

IronViT: Toward Efficient Generalist Visual Representation Learning

A generalist vision encoder must capture semantic, spatial, language-aligned, and action-relevant cues within a unified representation, yet softmax attention underlying today's most capable visual backbones becomes prohibitively expensive at high resolution. A natural attempt to address both challenges is to distill multiple specialist teachers directly into an efficient architecture. We find that directly coupling these objectives degrades representation quality, as the student must simultaneously reconcile heterogeneous capabilities and adapt them to a different token-mixing architecture. We introduce IronViT, built on a simple principle: consolidate capabilities before constraining computation. IronViT first distills complementary specialists into a softmax attention capability bridge, then progressively transfers the consolidated representation to a hybrid softmax-linear attention encoder. A purpose-built data pipeline further curates the distillation corpus for higher information density and broader domain coverage. Across recognition, retrieval, dense prediction, multimodal understanding, and robotic learning, IronViT is competitive with leading specialist and generalist vision encoders. The softmax bridge achieves the strongest aggregate performance in multimodal understanding and robotic learning among the evaluated backbones, while the hybrid encoder retains broad transfer performance with an efficiency advantage that grows with input resolution. Together, these results show that consolidating capabilities before architectural conversion can yield a generalist visual encoder without inheriting the prohibitive high-resolution cost of conventional softmax attention.

cs.CV↗

Scalable photoacoustic tomography implementations accounting for the spatial impulse response of transducers

Iterative model-based reconstruction in photoacoustic tomography repeatedly applies the forward operator mapping the initial pressure to the transducer signals, and its adjoint. At the scale of current three-dimensional systems, this operator cannot be stored and must be evaluated matrix-free, while accounting for the finite, focused surface of the transducers, whose spatial impulse response degrades the resolution when ignored. Representing the initial pressure by compactly supported radial functions, we show that the measured signal is exactly a temporal convolution between a system kernel gathering the radial function and the electrical impulse response, and a purely geometric quantity accounting for the portion of the transducer surface reached by the wave emitted from a voxel during one time step. Two implementations are proposed, differing only in how this quantity is evaluated: a quadrature over points of the surface, as in existing works, or a closed-form area, which never discretizes the surface. We derive closed forms for planar and cylindrically focused transducers and provide, in the latter case, two accelerations of the resulting elliptic integrals, a lookup table and a trapezoidal approximation, together with the piecewise planar approximation customary in the literature. These implementations reduce the per-voxel geometric computations and are released as an open-source Python package for graphics processing units. The performance of these operators is first demonstrated on a synthetic phantom, where the lookup-table-based operator reaches the accuracy of the exact evaluation ten times faster and outperforms the point discretization on both accuracy and runtime. A second experiment shows that they enable the processing of a realistic vascular phantom at full scale, with a higher peak signalto-noise ratio and a better resolution than the back-projection counterpart. The released implementations are an important step towards the adoption of three-dimensional modelbased photoacoustic reconstructions with finite and focused transducers.

eess.IV↗

Observation of symmetry breaking in time-varying scattering systems

From the Higgs mechanism to magnetic ordering, spontaneous symmetry breaking organizes physics across scales. Whether an analogous transition governs waves scattered by periodically driven structures has remained an open experimental challenge. Here we demonstrate a universal symmetry-breaking transition in a periodically driven finite scattering system. By measuring the multispectral Floquet scattering matrix of a strongly modulated microwave resonator, we observe a transition from an unbroken symmetry regime with scattering eigenvalues of equal modulus to a broken symmetry regime where their moduli split apart, through an exceptional point. At parametric resonance, this transition culminates in coherent perfect absorption, where a tailored multispectral input state is completely absorbed without outgoing radiation, the singular limit of the underlying transition. Because it arises from the general algebraic structure of the Floquet scattering matrix rather than from implementation-specific details, this transition should occur broadly in such systems. Unlike instability phenomena in bulk driven media, it emerges in the eigenstructure of a finite scattering operator that couples waves across multiple frequency channels. Our results establish Floquet scattering as an experimentally accessible platform for the observation of symmetry breaking and for dynamic wave control in driven photonic systems.

physics.optics↗

Rational Design of Low-Dimensional Hybrid Organic/Inorganic Interfaces for Enhanced Second-Harmonic Generation

Hybrid interfaces formed by push-pull organic molecules physisorbed on two-dimensional (2D) semiconductors provide a structurally tunable platform for engineering second-harmonic generation (SHG). However, predicting and optimizing their macroscopic response remains a formidable challenge due to the phase-sensitive interference between the nonlinear responses of the adlayer and substrate. Here, we develop a physics-informed computational screening framework that decomposes the effective second-order susceptibility into its constituent substrate, molecular, and electric-field-induced SHG channels. Validated by fully atomistic first-principles calculations across representative interface structures, this model maps the complete orientation-resolved tensor landscape of chemically modulated carbon-conjugated polar molecules on 2D substrates with distinct symmetry. Crucially, our findings dismantle the conventional reliance on static polar descriptors, demonstrating that the ground-state permanent dipole magnitude is a fundamentally unreliable proxy for macroscopic nonlinear response. Instead, we show that the ultimate criterion for SHG maximization is encoded in the phase-resolved projection of the full anisotropic molecular hyperpolarizability tensor. On non-centrosymmetric substrates, the coherent interference between the physisorbed adlayer and the underlying 2D matrix induces a characteristic Fano-like asymmetry and ranking reversals that are entirely absent on centrosymmetric platforms. By establishing that functionalization topology, dynamic spatial orientation, and substrate point-group symmetry constitute a single, non-separable parameter space, this work provides a systematic blueprint for the rational assembly, predictive discovery, and non-invasive characterization of next-generation low-dimensional hybrid materials for nonlinear optoelectronics.

cond-mat.mtrl-sci↗

Deep learning of longitudinal visual fields predicts glaucoma progression rate and identifies fast progressors

Glaucoma is the leading cause of irreversible blindness, and timely identification of fast progressors is essential to prevent disability. Current practice estimates progression by ordinary least-squares regression of mean deviation (MD) on time, requiring 6--10 visual field (VF) tests over several years to obtain a reliable slope. We present GLAM (Glaucoma Longitudinal Analysis Model), a deep learning framework that ingests longitudinal Humphrey 24-2 total deviation sequences with five clinical features and predicts MD and visual field index progression rates using attention-based fusion and aleatoric uncertainty. On the open-access University of Washington Humphrey Visual Field dataset (4,276 patient-eyes), GLAM achieved an MD-rate mean absolute error of 0.139 dB yr$^{-1}$ ($R^2 = 0.927$; 73.5% reduction over a ridge baseline) and an AUC of 0.990 for fast-progressor detection. VF-only deep learning can match multimodal pipelines for progression prognostication using routinely collected perimetry alone.

cs.CV↗

The quadratic density response function for non-interacting fermions at arbitrary temperature

We develop and implement the quadratic density response function of non-interacting fermions at arbitrary temperature, frequencies, and wave vectors. Starting from a Green's function formulation, we derive the quadratic response and demonstrate its equivalence to the result obtained from the Wigner equation. We further derive the classical limit through a perturbative expansion of the Vlasov equation and demonstrate that the quantum and classical formulations agree in the high-temperature limit. We analyse the limiting behaviour with respect to wavenumber and derive the zeroth harmonic response. Two independent implementations are provided and extensively benchmarked against density-functional theory, canonical path integral Monte Carlo (PIMC), and grand canonical PIMC simulations. As the density response of the interacting electron gas is commonly modelled through the ideal response functions and approximate models for the local field correction, the presented formulation will also allow for more complete explorations of interacting systems. Especially, our efficient implementation, which evaluates the ideal static and dynamic quadratic response functions in less than 0.5 ms on a 1.3 GHz processor, will enable evaluation of quadratic corrections to integrated quantities such as interaction potentials and stopping powers in warm dense matter.

physics.plasm-ph↗

Lower bounds for heat kernels of elliptic operators with unbounded diffusion, drift, and potential terms

We establish a pointwise lower bound for the heat kernel of the elliptic operator $ Λ=(1+|x|^α)Δ+b|x|^{α-2}x\cdot\nabla-|x|^β$, where $d\geq3$, $α>2$, $β>α-2$, and $b\in\mathbb R$. For every $τ>0$, we prove that $$ \begin{aligned} p(t,x,y) &\geq C_τe^{λ_0t} \left(\frac{1+|y|^α}{1+|x|^α}\right)^{\frac{b}{2α}} \frac{(|x||y|)^{-\frac{d-1}{2}-\frac{β-α}{4}}}{1+|y|^α}\\ &\quad\times\exp \left[ -\int_1^{|x|}\sqrt{\frac{s^β}{1+s^α}}\,\mathrm ds -\int_1^{|y|}\sqrt{\frac{s^β}{1+s^α}}\,\mathrm ds \right] \end{aligned} $$ for all $t\geqτ$ and $|x|,|y|\geq1$, where $λ_0<0$ is the largest eigenvalue of $Λ$ and $C_τ>0$ is independent of $t,x,y$. The proof applies the classical Davies - Simon argument in a weighted symmetric setting.

math.AP↗

Parameterized Complexity of Spanner Problems with Independent Weights and Lengths

In this paper, the parameterized complexity of the multiplicative $α$-spanner problem with independent weights and lengths on undirected graphs is considered for the first time. All prior FPT results (except one on DAGs) assume basic instances (i.e., with unit weights and lengths) and are parameterized in the stretch factor $α$ and the (in practice typically non-constant) number of removed edges. We show that several parameterizations do not allow FPT algorithms. However, our exclusion approach generalizes an existing algorithm for basic instances to arbitrary weights and lengths. It is parameterized by the total removed weight and a new tightness parameter. The latter is more precise than $α$ and allows us to also improve the best known result for basic instances. Our second algorithm, called inclusion approach, uses the natural parameterization in the spanner's total weight. We prove that this sole parameter leaves a W[2]-hard problem, but also show FPT algorithms exist when augmented with secondary parameters.

cs.DS↗

Information geometry of emergent symmetry quotients and their weak unfoldings

A regular observed statistical model may converge to a limit in which a previously identifiable signed parameter becomes identifiable only modulo a reflection. We study the local information geometry of this transition. For a twice differentiable Hellinger embedding with an exact limiting reflection, the observed displacement is forced into the two-jet form \(\varepsilonλJ_-+λ^2J_+/2\), up to higher-order terms. The mixed jet restores the sign away from the symmetric face, whereas the even jet is the first tangent inherited by the quotient. After nuisance elimination, a positive Gram determinant yields a nondegenerate cross-cap two-jet. The associated local asymptotic theory has three regimes governed by \(τ_n=\sqrt n\,\varepsilon_n^2\): regular signed LAN, a critical curved Gaussian subexperiment, and a quotient regime with the \(n^{-1/4}\) signed scale. We prove that the same parabolic critical experiment persists for predictive likelihoods along a single stationary dependent trajectory. The limiting quotient has a regular Fisher metric in the invariant coordinate, while its pullback degenerates in the signed coordinate. For a solvable CIR--OU benchmark motivated by coherent sea-clutter observations, we derive the quotient Fisher metric and curvature explicitly and show that the curvature is strictly negative. We also determine the restricted holonomy of the full Amari family: \(\operatorname{Hol}_0(\nabla^{(a)})=SO(2)\) for \(a=0\), whereas \(\operatorname{Hol}_0(\nabla^{(a)})=GL^+(2,\mathbb R)\) for \(a\neq0\). The results separate the intrinsic geometry of the limiting quotient from the transverse geometry of its weak unfolding.

math.ST↗

Dense-Joint-Based Obstacle-Aided Locomotion with a Joint-Repositionable Snake Robot

Obstacle-aided locomotion is a fundamental capability for snake robots to traverse complex environments. However, conventional rigid-link snake robots often suffer from stagnation or jamming caused by their low joint density (i.e., the number of joints per unit length). This results in discontinuous contact with obstacles, unlike the continuous adaptation of biological snakes. To investigate the effect of joint density on obstacle-aided locomotion performance, we utilized a joint-repositionable snake robot mechanism that decouples actuators from joints, enabling a high-density architecture. We developed two experimental models with identical total lengths but different joint densities (high-density and low-density) and conducted comparative propulsion experiments in obstacle environments with varying obstacle diameters. The experimental results demonstrate that the high-density model substantially suppresses the abrupt shifts in reaction forces that cause stagnation in the low-density model. By maintaining smooth contact points, the high-density configuration reduces power consumption and achieves stable, continuous propulsion. These results highlight high joint density as a key factor in improving the environmental adaptability of snake robots in complex terrains.

cs.RO↗

The Marcus-Minc Transform Inequality

We prove the Marcus-Minc conjecture: let n be any integer at least 2, and let A be a nonnegative n by n matrix whose row and column sums are all one. Form a new matrix by subtracting A from the n by n matrix of ones and dividing by n minus one. Then the permanent of A is at least the permanent of the new matrix. We also determine all equality cases.

math.CO↗

Nodal Orbital-Anti-Phase Superconducting State in Bilayer Nickelates

The recent discovery of high-$T_c$ superconductivity in the bilayer nickelate La$_3$Ni$_2$O$_7$ (La-327) under applied pressure and compressive strain opened a new avenue to elucidate the interplay between multiorbital intralayer and interlayer electronically driven Cooper-pairing in bilayer systems. Depending on the details of the electronic structure in the normal state, the superconducting gap in bilayer nickelates is predicted to have either bonding-antibonding $s_{\pm}$-wave symmetry, driven by dominant interlayer Cooper-pairing, or $d$-wave symmetry with substantial intralayer Cooper-pairing. Despite this general picture, the orbital structure of the superconducting gap in these multiorbital systems has been less explored. Here, we analyze the consequences of an orbital-anti-phase structure of the superconducting gap and discuss its possible experimental signatures. We demonstrate that additional pairs of nodes may appear on the $α$ and/or $β$ Fermi surface sheets due to the sign change of the superconducting gap between the involved orbitals. Apart from this additional nodal structure, which is not enforced by the symmetries of the gap function and can be probed in ARPES experiments, the orbital-anti-phase gap modifies the temperature dependence of the superfluid stiffness at low temperatures, providing a concrete experimental prediction to test its realization in bilayer nickelates and related multiorbital systems.

cond-mat.supr-con↗

TP-CRIV: A Framework for Third-Party Challenge-Response Identity Verification of AI Models

Artificial intelligence (AI) models are increasingly deployed through remote services, making model misappropriation a growing concern. Existing approaches, including watermarking, fingerprinting, and model similarity analysis, primarily rely on predefined evidence or direct behavioral comparison and do not explicitly evaluate whether the claimant currently possesses and can utilize model-dependent information relevant to the claimed model identity. In this paper, we propose Third-Party Challenge-Response Identity Verification (TP-CRIV) for AI models. TP-CRIV targets a third-party verification setting in which the verifier has neither white-box nor API access to the claimant's model, can interact with the suspicious deployed service only through its ordinary black-box inference interface, and does not require protocol-specific cooperation from the service provider. Under these constraints, the framework enables the verifier to obtain empirical evidence as to whether the claimant locally possesses a model satisfying a predeclared identity relative to the deployed model. Verification is conducted under fresh, previously undisclosed requirements and network isolation, so that the demonstrated capability cannot rely on online external assistance after challenge disclosure. The resulting evidence is interpreted relative to independently specified and calibrated matching and non-matching operating situations and is statistical rather than cryptographic. We instantiate TP-CRIV for CNN image classifiers using probability-control-based witness generation. Experiments on ten ImageNet-pretrained TorchVision models demonstrate clear same/cross-model separation and finite-challenge verification using independently calibrated thresholds.

cs.CR↗

A Stochastic Optimization Approach to Control-Affine Optimal Control Problems

We consider control-affine optimal control problems on the torus, where the dynamics and cost functions are only accessed through samples. Starting from a weak formulation of such problems, we derive a dual, a primal, and a primal-dual formulation, compatible with stochastic optimization. We show convergence of stochastic first-order methods to the optimal value under generic conditions. In addition, we introduce a computable metric that upper-bounds the performance of suboptimal controllers produced during optimization, under additional regularity assumptions. Preliminary results show that the method can efficiently solve a simple control problem. Finally, we discuss conditions for the boundedness of the optimal occupation measure, a key assumption for the primal and primal-dual approaches.

math.OC↗

Baszta: Data-Centric Fine-Tuning of a Polish Multi-Label Safety Classifier

We develop a multi-label Polish content-safety classifier by fine-tuning allegro/herbert-base-cased (124M) across five categories (hate, vulgarity, sexual content, crime, self-harm) using a Focal + R-Drop objective, and evaluate the resulting model against Bielik Guard (Sójka) on the shared out-of-distribution Gadzi Język benchmark. Both systems are given per-category threshold tuning on the same calibration split. Under that matched protocol our model holds a small but statistically significant lead in micro F1, while an apparent macro-F1 lead does not survive: it was an artifact of comparing a tuned model against an untuned one. We also report what that micro figure is worth. Because Gadzi Język is 97% crime-positive, a classifier that flags crime on every input and nothing else already scores 0.910 micro F1 on the same test split, so micro separates neither system from a degenerate strategy and macro is the column that does. Per-category and per-protocol figures are reported in Section 4. The residual out-of-distribution gap is one of calibration rather than discrimination. Ranking quality stays high while positive probabilities collapse, and per-category temperature scaling recovers the loss where Platt scaling and isotonic regression do not. That recovery turns out to be conditional on the calibration set containing safe text. Gadzi Język contains almost none, so thresholds fitted on it flag crime on every safe input, and a balanced refit buys a deployable operating point at the cost of adversarial recall. We report both operating points rather than only the flattering one. Two changes that are standard practice, per-class cost-sensitive weighting and mean pooling, each raise in-distribution macro F1 while lowering the out-of-distribution figure, which indicates that robustness has to be selected for directly rather than inherited from in-distribution accuracy.

cs.AI↗

MagiCFirm: A Runtime for Magic-State Cultivation with Algorithm-Hardware Co-Design

Magic-state cultivation offers a promising alternative for lowering the cost of non-Clifford operations in fault-tolerant quantum computing (FTQC). However, realizing cultivation in practice exposes two challenges: (i) the lack of an open-source classical runtime layer between logical software and physical control, and (ii) the latency constraints on protocol-specific decisions that determine magic-state readiness. To address these challenges, we adopt an algorithm--hardware co-design approach. At the algorithm level, we develop a two-stage early-escape scheme that identifies an informative subset of detectors offline, constructs a compact decoding problem, and performs partial decoding in parallel with complete decoding at runtime, allowing high-confidence attempts to advance before full decoding completes. At the hardware level, we present MagiCFirm, a configurable runtime that combines offline-compiled microprograms with dedicated datapaths for detector construction, event processing, and protocol control, enabling end-to-end execution of magic-state cultivation. Across evaluated configurations, MagiCFirm reduces wall-clock magic-state preparation time by up to 39.3% at matched logical error rate. For a representative magic-state-bound workload, this translates to an estimated 11% reduction in overall application runtime.

quant-ph↗