arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,441 records · Page 80Linked to original sources

Reinforcement Learning to Accelerate Primal-Dual Hybrid Gradient for Linear Programming

Primal-dual hybrid gradient (PDHG) methods solve large-scale linear programs (LPs) using GPU-friendly matrix-vector products and projections, but their practical performance depends on coordinating algorithm parameters, acceleration, and restarts. We introduce GALLOP, which uses reinforcement learning to jointly learn continuous algorithm parameters and discrete restart decisions without differentiating through the solver. Its generalized accelerated PDHG update combines separate primal and dual extrapolation, history corrections, and restart anchoring with independently adjustable coefficients. We train a dimension-agnostic feedback policy using a groupwise proximal policy optimization objective that clips likelihood ratios separately for different control groups and excludes inactive acceleration controls on restart transitions. We evaluate GALLOP on six LP families and a public item-placement benchmark. On the main evaluation settings across the six families, GALLOP reduces iteration counts by factors of $1.9$-$5.6$ and achieves up to a $16.0\times$ speedup in algorithm wall-clock time over MPAX. With one policy trained per family, the learned policies generalize without retraining to within-family LPs $3\times$-$400\times$ larger than the largest training instances, including Transport LPs with $10.24$ million variables.

math.OC↗

Tense Logic via Truth Degrees: An Algebraic Completeness Result for Kashima's Calculus

We study the minimal tense logic $K_t$ from an algebraic and proof-theoretic perspective. We introduce the degree-of-truth-preserving logic associated with the class of tense Boolean algebras. We then introduce a sequent calculus for this logic and establish its soundness and completeness with respect to tense Boolean algebras by means of an adaptation of the Lindenbaum--Tarski construction. Consequently, this calculus provides an additional syntactic presentation of the minimal tense logic $K_t$. We also provide a second, purely syntactic proof of completeness. Furthermore, we establish an algebraic soundness and completeness theorem for Kashima's Gentzen-style calculus. To this end, we develop an adaptation of the Lindenbaum--Tarski construction to Kashima's nested sequent framework, which allows us to construct the algebraic semantics directly from the proof-theoretic system. This yields a direct algebraic completeness proof for Kashima's calculus and connects its nested-sequent formulation with the algebraic semantics of tense Boolean algebras.

math.LO↗

Range-GRPO: Policy Optimization via Pairwise Relations among Reward Intervals

As the use of large language models (LLMs) expands, post-training has become increasingly important for adapting them to downstream tasks. However, obtaining reliable supervision remains costly, especially in domains without reference answers or executable verifiers. LLM-as-a-Judge provides scalable pseudo-rewards for unlabeled responses, but a single point score does not explicitly represent reward uncertainty. This motivates representing pseudo-rewards as conformally calibrated reward ranges. We propose Range-GRPO, a semi-supervised post-training framework that combines limited labeled data with unlabeled prompts. In Group Relative Policy Optimization (GRPO), learning signals depend on relative reward comparisons within each rollout group. The proposed objective compares reward ranges pairwise rather than reducing them to point rewards, allowing interval uncertainty to affect both the magnitude and direction of these signals. Our theoretical analysis characterizes this distinction and shows that the proposed objective recovers the Dr.GRPO advantage when all reward ranges collapse to points. Empirically, Range-GRPO achieves the highest in-distribution and out-of-distribution average performance among the evaluated semi-supervised methods while requiring fewer training resources.

cs.LG↗

Molecular Dynamics with Nuclear Effects on Quantum Computers

Nuclear quantum effects are critical for describing proton transfer and hydrogen bonding, but their incorporation into quantum chemistry calculations is often computationally prohibitive on classical hardware. A promising alternative are quantum computers due to their linear scaling in the space requirements with system size. We introduce a novel hybrid quantum-classical algorithm for ab-initio molecular dynamics that incorporates nuclear quantum effects via the nuclear-electronic orbital method. The proposed algorithm evaluates ground state energies and forces on the quantum computer using a variational quantum eigensolver, while the molecular geometries are updated classically. We validate our approach through simulations of $\text{H}_2$, $\text{H}_2\text{O}$ and the Zundel ion $\text{H}_5\text{O}_2^+$, comparing the simulated vibrational spectra with experimental data. Upon inclusion of nuclear quantum effects, the proton shuttling movement in the Zundel ion becomes effectively barrierless, and errors in the simulated frequencies improve significantly. Employing compact hardware-efficient ansatz circuits we achieve results comparable to the more accurate UCCSD ansatzes, which hints towards the feasibility of executing our algorithm on near-term quantum devices.

quant-ph↗

The sphere is the only closed surface satisfying the fixed-width Archimedean property

Among Archimedes' many celebrated discoveries is a striking fact: the region of the unit sphere between two parallel planes, both meeting the sphere and separated by a distance $h$, has area $2πh$. We prove that, among connected smooth closed embedded surfaces in $\mathbb R^3$, this property characterizes the unit sphere for a single fixed width.

math.DG↗

Room-temperature Magnetoelastic Coupling in UIr$_4$Al$_{15}$

The interplay between electronic and structural degrees of freedom underpins many emergent phenomena in quantum materials, yet magnetoelastic coupling in metallic systems is typically weak and confined to low temperatures or microscopic length scales. Here, we report giant magnetoelastic coupling slightly above room temperature in the uranium-based intermetallic compound UIr$_4$Al$_{15}$. Using temperature-dependent single-crystal X-ray diffraction together with anisotropic magnetic susceptibility measurements, we directly resolve subtle but reproducible structural distortions coupled to magnetic alignment. Despite the absence of crystallographic symmetry breaking, pronounced anomalies emerge in lattice parameters and selected chemical bond distances near the magnetic transition region, revealing an unusual sensitivity of the crystal structure to magnetic orientation. The coupling enables direct probing of how atomic distances and local chemical bonding govern the electronic and magnetic states in a bulk intermetallic material. Our results establish UIr$_4$Al$_{15}$ as a rare platform in which magnetism and lattice distortions are strongly intertwined near room temperature, opening new opportunities for magnetically responsive quantum materials and functional magnetic sensing applications.

cond-mat.str-el↗

Radiative corrections to the subprocess of annihilation of quark-antiquark pair of prompt photon production in proton-proton collisions at NICA energies $\sqrt{\mathbf{s}}$= 10 GeV

In this work, a comprehensive theoretical and numerical study of the quantum-chromodynamic and quantum-electrodynamic radiative corrections: 1. $q\bar{q} \rightarrow q\bar{q}γ$, 2. $q\bar{q} \rightarrow ggγ$, and 3. $q\bar{q} \rightarrow gγγ$ to the annihilation of a quark-antiquark pair $q\bar{q} \rightarrow gγ$ of prompt photon production in proton-proton collisions at NICA energies $\sqrt{s}$ = 10 GeV was performed. The differential cross-sections and double spin asymmetries $A_{LL}$ of the radiative corrections were investigated analytically using the FeynCalc and modulated in PYTHIA 8.316. Radiative corrections are significant at low collision energies of protons $\sqrt{s}$ and small transverse momentum of photons $p_{T}$. It was shown that radiative corrections to the annihilation of a quark-antiquark pair process satisfy the following relations: $$ 2. σ_{\sqrt{s}}(q\bar{q} \rightarrow ggγ) > 1. σ_{\sqrt{s}}(q\bar{q} \rightarrow q\bar{q}γ) > 3. σ_{\sqrt{s}}(q\bar{q} \rightarrow gγγ) $$ $$ 1. \frac{dσ(q\bar{q} \rightarrow q\bar{q}γ)}{dp_{T}^{2}} > 3. \frac{dσ(q\bar{q} \rightarrow gγγ)}{dp_{T}^{2}} > 2. \frac{dσ(q\bar{q} \rightarrow ggγ)}{dp_{T}^{2}} $$ The polarization of colliding protons has almost the same effect on processes 1. $q\bar{q} \rightarrow q\bar{q}γ$, 2. $q\bar{q} \rightarrow ggγ$, and 3. $q\bar{q} \rightarrow gγγ$ and in this case above indicated relation is also satisfied. A direct comparison between strict leading order analytical calculations in FeynCalc and PYTHIA 8.316 modulations revealed excellent agreement in the central kinematic regions.

hep-ph↗

From Rules to Neural Graphs: Scalable Structured Prediction for Patent Prior Art Search

Patent search requires processing documents routinely exceeding tens of thousands of tokens. Most neural retrieval approaches operate on truncated inputs, limiting their effectiveness. Graph-based retrieval addresses this by representing each patent as a structured invention graph, but constructing these graphs relies on brittle rule-based parsers. We present the neural parser, which adapts biaffine attention from dependency parsing to predict invention graphs directly from patent text. Our local biaffine attention restricts pairwise scoring to a sliding window, reducing complexity from $O(n^2)$ to $O(n \cdot w)$. Since local and global scoring share the same weights, the model trains on short sequences and deploys on documents exceeding 40,000 tokens without retraining. Distilled from 1 million rule-parsed documents, it surpasses its teacher at 3$\times$ lower inference cost: neural graphs improve citation recall by 0.5% on short queries and 1.1% on full documents in a downstream Graph Transformer retrieval system.

cs.IR↗

QK-Wanda: Coupling Queries and Keys for Unstructured Pruning

Wanda (Sun et al., 2024) prunes large language models by scoring weights independently within each linear projection, although queries and keys interact through dot products. We introduce QK-Wanda, which scores query and key weights by their individual deletion costs under an unmasked pre-RoPE reconstruction objective. It augments Wanda scores with information from the opposite projection (keys for query weights, and queries for key weights), allowing both projections to share a pruning budget. Its closed-form scores require no gradients or weight updates; full pruning takes 1.3% longer than Wanda on A100 and 3.1% longer on H200 with the calibration used in our main experiments. We evaluate QK-only pruning across 15 models from TinyLlama, Llama 2, Llama 3, and Qwen2.5, spanning 0.5B-72B parameters. Relative to Wanda, QK-Wanda reduces QK reconstruction error by an average of 60% at 50% sparsity and 45% at 80%. Downstream gains depend on the model. At 80% sparsity on Llama 2 70B, WikiText-2 and C4 perplexity decrease by 20.3% and 13.5%, while mean zero-shot accuracy rises by 5.94 percentage points. Qwen2.5-72B also improves, but Llama-3.1-70B has substantially higher perplexity despite lower reconstruction error. These results show both the promise of coupled pruning criteria and the limits of local reconstruction as a predictor of model quality.

cs.LG↗

Efficient certification of time-reversal symmetry requires entanglement

Time-reversal symmetry is a fundamental principle of physics describing the invariance of physical laws under reversal of the direction of time. We formulate a Bell-inequality-like test of this antiunitary symmetry using only forward access and trusted quantum operations: entanglement converts temporal input--output relations into measurable spatial exchange symmetry. For $n$-qubit unitary dynamics, we prove that reliably distinguishing the time-reversal-symmetric circular ensembles from Haar-random dynamics requires $Ω(\min\{2^{n/2},2^{n-e}\})$ queries for any classically adaptive protocol. Here, $e=\min\{e_{\mathrm s},e_{\mathrm m}\}$ with $e_{\mathrm s}$ and $e_{\mathrm m}$ representing the probe and measurement logarithmic entanglement negativities, respectively. Maximally entangled probes and SWAP measurements reduce this cost to a constant number of queries. Furthermore, we develop a time-reversal symmetry test for arbitrary fixed, compatible probes and measurements, relate its query complexity to their logarithmic negativities, and match the lower-bound scaling in the high-entanglement regime by optimizing the probe and measurement. Our results establish a quantitative connection between entanglement and time-reversal symmetry, bridging two central concepts in quantum information science and fundamental physics.

quant-ph↗

Local and 2 local $\frac{1}{2}$-derivation of $n$-dimensional totally graded filiform Lie algebras

This article provides a complete algebraic description of $\frac{1}{2}$-derivations, local $\frac{1}{2}$-derivations, and 2-local $\frac{1}{2}$-derivations on $n$-dimensional totally graded complex filiform Lie algebras of maximum length. Based on the foundational classification framework established by Janez Bernik (2020), we systematically determine the vector spaces of $\frac{1}{2}$-derivations for the six infinite structural sequences ($m_0(n)$, $m_2(n)$, $W^+(n)$, $m_{0,1}(n)$, $m_{0,2}(n)$, $m_{0,3}(n)$) and the five exceptional one-parameter families ($g_{7,α}$ through $g_{11,α}$). By analyzing the pointwise local evaluation equations via parametric matrix systems, we establish the structural linearity and rigidity of local $\frac{1}{2}$-derivations. In contrast, we demonstrate that the independent parameters residing in the boundary rows of the $\frac{1}{2}$-derivation matrices provide sufficient degrees of freedom to bypass linearity constraints. Exploiting these boundary configurations, we explicitly construct pure non-linear and non-additive 2-local $\frac{1}{2}$-derivations leveraging the homogeneous function of degree one, $f(z_1, z_2) = z_1^3 / (z_1^2 + z_2^2)$, thereby defining the exact boundary where local rigidity fails.

math.RA↗

Conifold factorization in topological recursion and Gromov--Witten theory

We prove a factorization theorem for ordinary topological recursion near a regular nodal degeneration of a spectral curve, allowing logarithmic spectral coordinates. The partition function for closed genera $g\ge2$ factors into the partition function of the normalization, a universal Gaussian vacuum, and the exponential of a connected graph sum. This graph sum has strictly positive order in the vanishing period, and thus the factorization recovers the conifold gap, identifies the term of degree zero in the vanishing period with the free energy of the normalization, and gives finite graph formulas for the coefficients of positive powers of the period. At fixed genus, the regular series converges jointly in the period and the parameters along the nodal locus. The Gaussian neck tensors can be expressed in terms of relative Gromov--Witten invariants of a parametrized $\mathbb P^1$. As examples, we use remodeling to obtain Gromov--Witten factorizations for local $\mathbb P^2$ and local $\mathbb F_0$ in the conifold frame, with backgrounds $\mathbb C^3$ and the resolved conifold, respectively. The analogous local $\mathbb F_1$ factorization has background $\mathcal O(1)\oplus\mathcal O(-3)$ over $\mathbb P^1$.

math.AG↗

Controllable Stochastic Quantization Encoding for Adversarially Robust Spiking Neural Networks

Spiking Neural Networks (SNNs) have attracted increasing attention due to their impressive temporal dynamics, energy efficiency, and brain-inspired mechanisms. Although SNNs have demonstrated promising performance in image classification tasks, recent studies have shown that they remain vulnerable to adversarial attacks, where imperceptible perturbations are added to input images to mislead model predictions. Existing defense methods mainly focus on training strategies, while the role of input encoding remains less explored. An observation is that the robustness advantage of Poisson encoding over direct encoding may benefit from its inherent randomness. Motivated by this, we propose a stochastic quantization encoding method that encodes the input image with controllable randomness adjusted by the quantization scale, thereby improving the adversarial robustness of SNNs. We further show that this method constitutes a general framework that reduces to both Poisson encoding and direct encoding under different choices of the quantization scale. Since it enhances robustness at the input encoding stage, it can be combined with existing training-based defenses for further gains. Experimental results on CIFAR-10 and CIFAR-100 demonstrate the effectiveness of the proposed stochastic quantization encoding method. To sum up, this work highlights the importance of input encoding for the adversarial robustness of SNNs, providing a new perspective for understanding and improving it.

cs.NE↗

Completion Aware Guidance for World Action Models

World Action Models (WAMs) predict visual futures and robot actions, yet they remain susceptible to task-incomplete imagination, where plausible, action-consistent predictions omit the transition needed for task completion. In this paper, we show that this failure is not inherent to the world model backbone, but emerges when adapted for short-chunk control, which can repeatedly favor plausible local continuations over task-completing transitions. To address this, we introduce Completion Aware Guidance (CAG), a training-free sampling method that guides generation toward task completion. Across representative WAMs, CAG improves success from 64% to 70% on a RoboTwin 2.0 subset and from 69% to 75% in zero-shot simulation, while reducing task-incomplete imagination from 79% to 40%.

cs.RO↗

AURAL: Adaptive Latent Reasoning with Joint Chunk for Speech Language Models

Model intelligence and fast response jointly shape the quality of interaction with speech language models, yet remain difficult to achieve together. Explicit chain-of-thought (CoT) improves reasoning and audio understanding, but generating intermediate reasoning tokens delays responses. Describing fine-grained acoustic cues further lengthens CoT and increases latency. Latent reasoning can reduce this overhead, yet existing methods often trail CoT and remain limited by single-path supervision and reasoning budgets that do not adapt to problem difficulty. We introduce AURAL, which models a distribution over multiple plausible reasoning continuations in latent space and jointly predicts chunks of future states to reduce sequential forward passes and reasoning latency. To provide initial supervision for latent reasoning, we construct AuralReason-683K: 683K bilingual speech utterances (about 1,000 hours) with concise CoT for emotion recognition, empathetic dialogue, and general reasoning. AURAL-RL then explores beyond these traces, rewarding concise reasoning that yields high-quality answers and adapting reasoning effort to each problem. Across two backbones, AURAL-RL achieves performance comparable to CoT-RL, with larger gains over the respective supervised checkpoints on most metrics. Analysis further shows that harder questions elicit more latent reasoning steps. On Qwen2.5-Omni, it reduces time to the first answer token by 11.8x, from 1.22 to 0.10 s, versus 0.05 s for direct answering.

cs.CL↗

Sharp higher-order uncertainty principles

We establish a class of sharp higher-order inequalities related to the uncertainty principle. First, by means of the expanding the squares method, we give a concise proof of sharp higher-order Heisenberg uncertainty principles. We provide explicit expressions for the optimal constants and establish the existence of extremal functions. Furthermore, we obtain sharp higher-order Heisenberg uncertainty principles for curl-free vector fields. In particular, when $N=2$, we give an answer to the higher-order version of Open Problem 9 raised by Maz'ya in [Integr. Equ. Oper. Theory (2018) 90:25]. Second, we establish sharp higher-order Heisenberg uncertainty principles that do not involve higher-order tensors. We also prove sharp higher-order Heisenberg uncertainty principles and higher-order hydrogen uncertainty principles for radially symmetric functions.

math.AP↗

Social welfare and price discovery in double auction markets

The tendency of the double auction mechanism to drive prices to competitive equilibrium has been well documented in laboratory experiments, but the phenomenon has lacked a theoretical explanation. This paper studies dynamic double auctions in a pure exchange economy where agents bid their indifference prices implied by their current holdings and preferences. We show that Walras equilibria coincide with the fixed points of the double auction and that repeated double auctions generate bounded sequences of allocations and prices whose cluster points are Walras equilibria with transfers.

q-fin.TR↗

Optimal Universal Coding of Integers

Universal coding of integers (UCI) provides binary codewords for positive integers such that, for every nonincreasing source distribution $P$, the average codeword length stays within $K$ times $\max\{1,H(P)\}$. The smallest constant $K$ is called the minimum expansion factor of UCI $\mathcal{C}$, denoted $C_{\mathcal{C}}^{*}$. The optimal minimum expansion factor $C^*=\inf\{C_{\mathcal{C}}^{*}\}$ is the minimum expansion factor corresponding to the optimal UCI. The optimal minimum expansion factor is currently known to lie in the range $2\le C^*\le 2.0386$. In this paper, we construct a family of one-point plus uniform-tail distributions and prove that, for every universal code, the worst-case ratio is attained by a distribution in this family, so that the family is least favorable for the UCI problem. We further establish an inequality, called the \emph{UCI inequality}, which plays the same role for UCI as the Kraft inequality does for prefix codes: for any real number $B$, it decides whether $B$ lies below or above $C^*$. Through the UCI inequality, we obtain an equivalent definition of $C^*$. By numerical computation, we determine $C^*=2.000124757036101\cdots$, the first fifteen decimal digits being certified. Once $C^*$ is known, we can theoretically construct the optimal UCI.

cs.IT↗