arXiv Science⌕ Search

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,477 records · Page 82Linked to original sources

A Nonlocal Stochastic Optimal Control with Elliptic-Type Smoothing

We study a nonlocal stochastic optimal control problem in which the control acts through the elliptic smoothing operator $\mathcal{S}=(I-Δ)^{-1}$ on $\mathbb{R}^d$. The state is described by the law of a controlled diffusion, equivalently by a controlled Fokker-Planck equation with coefficients depending on $\mathcal{S}u$. We prove the existence of an optimal control by the direct method in the calculus of variations. We then derive a Pontryagin-type maximum principle by spike variations, relying on sharp properties of the operator $\mathcal{S}$. The result yields a pointwise minimisation rule for the optimal control. We discuss the meaning of the optimal control problem in the case of population dynamics.

math.OC↗

Fourth-order perturbation theory for the Frohlich polaron: Analytic structure of the weak-coupling series and the crossover to strong coupling

We calculate the Rayleigh-Schrodinger perturbation series for the ground-state energy and the effective mass of the three-dimensional Frohlich polaron through fourth order in the coupling constant alpha. The third- and fourth-order coefficients of the effective mass are new. Because the Frohlich interaction is a form-bounded perturbation of an isolated nondegenerate ground state at fixed total momentum, the series has a nonzero radius of convergence, and we analyze it with the ratio and Pade methods appropriate to convergent series. The coefficients are strikingly regular and are described by a simple power-law singularity at alpha near 8.5 with an exponent near 1.4. The couplings at which successive truncations of the inverse mass vanish, the first of which is the textbook breakdown value alpha = 6, are slowly converging estimates of this radius. Pade approximants built from the weak-coupling coefficients alone agree with diagrammatic Monte Carlo results for the mass to better than one percent for alpha up to 5. Exact results on the analyticity of the ground state and on the strong-coupling asymptotics imply that the coefficients cannot keep a constant sign, although all known coefficients do, so that the apparent singularity must be a complex-conjugate pair close to the positive real axis. It marks the crossover to strong coupling: the weak-coupling series locates this crossover accurately but cannot be continued through it.

cond-mat.str-el↗

Spacing statistics for a point scatterer on the cubic three-torus

We study the new eigenvalues of a fixed point interaction on the cubic three-torus. Their mean-one consecutive spacings converge to a probability law independent of the interaction parameter. We characterize this law by the consecutive zeros of a random meromorphic function built from three-squares congruence densities. Its small-gap distribution has the form $s^5Ψ(\log_2(s^{-2}))+o(s^5)$, where $Ψ$ is continuous, positive, and one-periodic. The proof combines periodic mean-square approximation, a weighted Hilbert-transform estimate, and Fourier estimates that retain the interactions between primes.

math.NT↗

Spectral Efficiency Analysis of Massive MIMO-OFDM Systems with Double Quantization

Massive multiple-input multiple-output (MIMO) achieves a high spectral efficiency (SE) by employing a large number of receive antennas in the uplink. Employing many receive antennas requires a large number of analog-to-digital converters (ADCs) and a high fronthaul data rate. Low-resolution ADCs can reduce power consumption and cost, while fronthaul quantization can reduce the required fronthaul data rate, at the expense of quantization distortion at both stages. In wideband orthogonal frequency-division multiplexing systems, ADC quantization is performed in the time domain, while fronthaul quantization may be applied after the discrete Fourier transform to transmit only active subcarriers. In this study, we investigate the achievable uplink SE over fronthaul links under double quantization and imperfect channel estimation. We derive a linear minimum mean-squared error channel estimator from the double-quantized pilot observations and obtain an achievable SE using the use-and-then-forget bound. Numerical results demonstrate that the ADC and fronthaul quantization resolutions have comparable impacts on the achievable SE. When the two resolutions differ, the lower resolution becomes the dominant factor limiting the SE. Furthermore, our numerical results show that six-bit ADC and fronthaul quantization achieve more than 90% of the SE attained in the ideal case.

cs.IT↗

Gaussian Equivalence for Multi-Head Self-Attention

A theoretical understanding of multi-head self-attention is fundamental to the study of modern neural networks. Using random matrix theory, we establish Gaussian equivalence for multi-head self-attention: replacing softmax attention with rescaled scores plus Gaussian noise preserves the limiting spectral law of the centered output. This equivalence also covers value and output projections that depend on the keys. The resulting laws separate the effects of head allocation and projection widths, and distinguish spectrum-preserving across-head sharing from within-head key--value dependence.

stat.ML↗

A second-moment proof of quenched equals annealed for the Potts model on random regular graphs

Ferromagnetic spin models on random regular graphs, such as the Ising and Potts models, have been studied extensively. Several equivalent variational and recursive representations of their limiting pressure are known, and these representations can be used to establish agreement between the quenched and annealed pressures. We give a concise overview of these representations and highlight the connections among the different formulations. We then derive analogous representations for the exponential growth rate of the second moment of the partition function, and prove that it has the same order as the square of the annealed pressure, up to a subexponential error. Combined with concentration of the quenched pressure, this identity yields a transparent second-moment proof of quenched--annealed agreement on random regular graphs. By bringing these representations and arguments together in a unified presentation, we aim to provide an accessible entry point to the literature on ferromagnetic spin models on random regular graphs.

math.PR↗

The tableaux monoid: a rewriting-theoretic approach

We study the tableaux monoid from the perspective of rewriting theory. We first construct a finite quadratic convergent presentation based on column generators and prove that it yields a biautomatic structure. We then extend this presentation to a finite coherent presentation in the sense of polygraphic rewriting. These results provide canonical normal forms, yield homological and algorithmic consequences, and produce a higher-dimensional presentation that can be used to describe actions of the tableaux monoid on categories.

math.CO↗

Breaking the $\sqrt{3}$ Barrier for Maximum Weighted $3$-Set Packing

We give a deterministic polynomial-time $1.6908$-approximation for Maximum Weighted $3$-Set Packing, breaking the $\sqrt3$ locality-gap barrier of squared-weight local search. The approximation ratio for this problem progressed from Berman's $2$ [Ber00] to Neuwohner's $2-\frac{1}{63{,}700{,}992}+ε$ [Neu21]. Thiery and Ward then obtained $1.786$ [TW23], while Thiery subsequently improved the bound to $1.761+ε$ and finally to $\sqrt3 \approx 1.732051$ through a layered exchange analysis [Thi23]. Thiery also proved that $\sqrt3$ is a locality-gap lower bound for the squared-weight objective even with exchanges of arbitrary size. Our algorithm performs in two phases and combines two objectives. Phase~I computes a bounded-exchange local optimum for the squared-weight potential and analyzes it through Thiery's layered framework, while strengthening the terminal analysis by preserving internal tree-edge slack for nonsingleton components and exploiting the incidence structure of $3$-sets for final singletons. This yields an augmented structural inequality with residual positive claw gain under the original objective. Phase~II switches to the original objective and recovers sufficient residual gain through an auxiliary weighted $9$-Set Packing instance. A covering argument transfers the structural bound through the high-girth lift used only in the analysis.

cs.DS↗

Matching of signal, noise and hardware timescales for filtering and forecasting of correlated noise signals

Physical reservoir computing exploits the nonlinear dynamics of physical systems to process time-dependent data with greater energy efficiency than conventional machine learning approaches. However, physical reservoirs have fixed intrinsic response timescales, whereas real-world signals combine deterministic and stochastic components across multiple timescales. Here we show, using a nanoporous niobium oxide reservoir, synthetic noisy signals and cryptocurrency-price volatility, that the relationship among noise correlation time, reservoir memory and forecast horizon determines whether correlated noise is filtered or predicted. Noise varying faster than the relevant reservoir memory and forecast horizon is averaged by the reservoir, whereas the temporal structure of slower-varying noise is sufficient for algorithmic forecasting. We introduce the reservoir memory horizon and forecasting regime index to distinguish these operating regimes. These contributions demonstrate that timescale matching can guide the encoding of input time series and development of physical reservoir architectures that filter, analyse and predict stochastic signal components across distinct temporal scales.

cs.LG↗

Evolve on the Host, Predict on the Edge: Deploying Online Neuroevolutionary Architecture Search for Cross-sectional Stock Return Prediction

Accurate forecasting models are usually large, expensive to update online, and fixed in architecture once trained. We apply ONE-NAS, an online neuroevolutionary architecture search that evolves a population of small recurrent networks as each window of data arrives, to daily cross-sectional stock return prediction, and pilot it on a host and endpoint pipeline: the host runs the search and ships each generation's champion genomes over TCP/IP to a Raspberry Pi 4B, which predicts online. On the Pi a single champion predicts a 50-stock window in 24.6~ms and the ensemble of 40 island champions in 556~ms, far inside the daily decision cycle. On four panels of US mid-cap equities over 2022--2024, reading the population as a rank-mean ensemble of island champions returns $+27.5\%$ net of realised transaction costs, against $+11.3$ to $+14.8\%$ for online LSTM, online GRU and monthly-retrained LSTM baselines and $+4.5\%$ for the single best genome used in prior ONE-NAS work.

cs.CE↗

Video-to-Model: Automatic Modeling of Deformable Linear Objects

This paper presents a video-to-model framework for automatically modeling the motion of a deformable suture thread from an input video. We utilize a recently developed CBF--CLF--QP numerical model that simplifies the characterization of deformable string motion through the selection of a small number of parameters. A perception module first localizes and tracks the thread in video, producing an ordered sequence of thread nodes. The observed thread motion is then processed by a spatio-temporal CNN network that estimates the effective parameters of a structured CBF--CLF--QP model. These parameters are used to simulate the thread under a user-defined needle velocity input. Experiments using unseen thread configurations and motion demonstrate that the framework can reliably reconstruct the thread behavior from video, automatically configure the structured model, and reproduce the expected thread motion with low tracking error. The proposed approach reduces the need for manual parameter tuning and provides a step toward automatic video-based modeling of deformable linear objects.

cs.RO↗

Can we make sense out of negative kinetic energy?

Many Quantum Field Theory models, either designed to improve the ultraviolet behavior via higher derivative terms or aimed at providing unified descriptions of Gravity, inevitably introduce negative-energy 'ghosts' states. These states raise fundamental questions regarding the stability of the theory. In this study, we investigate the time evolution of simplified quantum mechanical models consisting of two coupled normal and ghost subsystems. Although the full spectrum of these models is unbounded from below and, therefore, lacks a true ground state, we focus the time evolution of a 'would-be ground state': a zero-energy eigenstate of the decoupled system. By analyzing its metastability under interactions, we gain insights into the dynamics of ghost-like states. Contrary to the standard paradigm, our results demonstrate that, across a broad range of couplings, the would-be ground state does not decay and remains dynamically stable over long times. This provides evidence towards resolving the stability problems of Hamiltonians that are unbounded from below. We also discuss the connections with the corresponding classical and semiclassical dynamics.

quant-ph↗

A mathematical perspective on nonplanar on-shell forms

On-shell forms are differential forms on the Grassmannian which arise in particle physics. They are defined using bipartite graphs with $n$ distinguished ``boundary'' vertices. Mathematical investigation of on-shell forms has largely focused on the case where the graph is planar, in which case one can utilize combinatorial tools pioneered by Postnikov in the study of the totally nonnegative Grassmannian. In this article, we investigate on-shell forms for arbitrary graphs. We first discuss how to extend various tools from the planar case to arbitrary graphs. We then prove a determinantal formula for a class of on-shell forms which first appeared in physics literature. We explain the relation between this class of forms and the hypertree divisors of $M_{0,n}$, introduced by Castravet--Tevelev.

math.CO↗

Learning to Accumulate Knowledge with Mutual Information

Large language model (LLM) agents can improve their performance by reusing knowledge distilled from past interactions. However, curating new experiences into a knowledge bank that becomes more useful as it grows remains challenging. Effective knowledge accumulation should limit redundant overlap among entries and ensure that new knowledge contributes beyond what the bank already provides. Yet training a curator with Group Relative Policy Optimization (GRPO) on standalone task success can reinforce general guidance even when it duplicates existing knowledge. Therefore, we propose Knowledge Weaver, a reinforcement learning framework that trains a language model to curate reusable knowledge from agent trajectories. We couple feedback inspired by token-wise mutual information (MI) with marginal success rewards to guide knowledge accumulation. Together, these signals encourage the curator to preserve distinct information from experience and produce entries that improve task success when added to existing knowledge. Standalone success rewards also favor entries that are useful on their own. On ALFWorld and WebShop, Knowledge Weaver achieves mean success rates of 54.0\% and 42.0\% with k=10 retrieved entries, exceeding GRPO by 16.9 and 18.7 percentage points, respectively. Its knowledge banks also outperform the evaluated prompt-based and established banks, including human-written banks, in overall ALFWorld success rate and WebShop score with the executor frozen. Our codebase is available at https://github.com/LaoKuiZe/Knowledge-Weaver.

cs.AI↗

Building Macroeconomically Relevant Climate Indices via the Assemblage VAR

What should a macroeconomically relevant climate index contain? The composition is inherently ambiguous, and aggregation choices affect structural inference. We introduce the Assemblage VAR, which jointly estimates nonnegative aggregation weights and VAR parameters by maximizing the system likelihood-gain criterion, effectively outsourcing aggregation to observed macroeconomic dynamics. Two variants operate in component-space and rank-space, reweighting named subcomponents or emphasizing regions of the cross-sectional distribution. Applied to disaggregated U.S. climate data from the Actuaries Climate Index and NOAA, VARs using the assembled climate measures yield contractionary impulse responses that are substantially larger than those estimated using fixed-weight benchmarks. Weights emphasize high-wind variables and distributional tails over slow-moving components such as sea level.

econ.EM↗

CANDO: Cooperative Agentic Network for Layout Design Optimization

Layout generation for real-world facilities is a challenging problem, requiring reasoning over irregular site boundaries, heterogeneous orientations, access-aware placements, and motion-planning feasibility. Yet, most existing layout benchmarks in the generative AI space target simpler placements over rectangular domains and rely on distributional metrics such as FID and IoU that reward conformity to dataset priors, thus discounting design innovation. Motivated by these gaps, we introduce ALPS-Bench, a benchmark of $1,000$ professionally annotated real-world facility layouts paired with an instance-specific scoring protocol grounded in a structured design manual. As a strong baseline for ALPS-Bench, we propose CANDO, a training-free multi-agent framework in which specialized agents iteratively refine layouts through a verification-grounded loop, concentrating reasoning on strategic spatial decisions. We demonstrate that CANDO surpasses state-of-the-art trained and LLM-based baselines on the widely adopted PubLayNet, RICO, and PKU-PosterLayout benchmarks, establishing cooperative agentic design as a broadly effective recipe for constraint-aware layout synthesis.

cs.MA↗

The Komar charges of type~II supergravity and supersymmetric localization

We compute the on-shell closed, generalized Komar charges associated to an arbitrary Killing vector of the two 10-dimensional $\mathcal{N}=2$ supergravity theories which are the low-energy effective field theories of the type~IIA and IIB superstrings. We study different rewritings of this charge in terms of Page charges and forms with special properties that are used in the derivation of thermodynamical relations in black $p$-brane spacetimes. The, we show that, when the Killing vector is the supersymmetric Killing vector built as a bilinear of a Killing spinor, the momentum maps of the different fields of the theory are

hep-th↗

Optimal Investment to Reach a Financial Goal: A Stochastic Control Framework

We develop a framework for an investor who trades until she either reaches a financial goal or an exogenous deadline arrives. Analogous to utility functions over wealth, we measure satisfaction with the timing of reaching a goal by a discount function. For a continuous-time market where a stochastic factor drives the dynamics of stock prices and the financial goals, the investor maximizes the expected discount at the goal reaching time and the expected utility of the funding ratio if the goal remains unreached by the deadline. This setup leads to a new class of stochastic control problems. We establish Bellman's principle of optimality and characterize the value function as a viscosity solution to the associated Hamilton-Jacobi-Bellman equation. When the deadline is infinite and the goal is constant, the HJB equation reduces to a form related to the backward heat equation, for which we provide a complete characterization of smooth solutions. When analytical solutions are unavailable, we develop an approach based on Howard's algorithm to obtain numerical solutions. Our analysis shows that optimal investment policies for goal reaching problems can be decreasing in the drift of risky assets and need not converge to full risk-free investment even as volatility diverges to infinity.

q-fin.MF↗