arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,657 records · Page 92Linked to original sources

Intersecting a curve in an abelian variety with multiples of another curve

Levin asked what can be said about the locus of points lying on a given curve in $\mathbb{G}_m^n$ that have a non-zero integer multiple on another given curve in $\mathbb{G}_m^n$ for $n \geq 3$. We give a definite answer to the abelian analogue of Levin's question, proving what is predicted by the Zilber--Pink conjecture in this case. An important ingredient in our proof is a strengthening of a height inequality by Vojta and Rémond.

math.NT↗

Edge Accuracy Is Not Enough: Why Dynamics-Learned Structure Fails to Transfer to Inverse Problems

A natural strategy for inverse problems with scarce labelled data is to transfer relational structure learned from abundant forward-simulation data. We show this strategy fails systematically, even when it satisfies the standard theoretical justification for why structure should help. We prove that approximate structure provides estimation-error benefits whenever the edge error satisfies $Δ< n^2 - kn$, reducing sample complexity from $O(n^2)$ to $O(kn+Δ)$. Structure learned via Neural Relational Inference (NRI) from dynamics prediction satisfies this condition, yet on a source-localisation task across 180 CFD-simulated hydrogen-leak scenarios and 180 acoustic scenarios, it degrades performance by 116% and 201% relative to a flexible, task-optimised attention baseline, while a physics-based prior (Green's function) degrades by only 69-72%. Four independent lines of evidence show this is not a tuning failure: NRI improves only 0.5% when given 18x more training data (versus 16.6% for the task-optimised baseline, $p<0.001$); performance is insensitive to the NRI edge threshold across a wide range; the dynamics-learned graph overlaps the task-optimal graph on only 6% of edges; and two further dynamics-derived structure estimators (correlation- and mutual-information-based) show no measurable benefit over a structure-free baseline, with the correlation-based estimator performing markedly worse. We formalise this gap as a statement about approximation error that the edge-accuracy condition cannot control, and we provide a lightweight transferability test (Jaccard similarity against a partially-observed target-task graph) that separates successful from failed transfer in all four domain/structure pairs we evaluate, using under an hour of computation and 15-20% of target-domain data; we present this as a heuristic calibrated on few cases, not a validated general threshold.

cs.LG↗

RoBART: Bayesian Additive Regression Trees with Tree-Specific Rotations

Bayesian additive regression trees (BART) can require many splits to approximate boundaries misaligned with the predictor axes. RoBART assigns each tree a rotation shared by all internal nodes, retaining axis-aligned splits in rotated coordinates and constant leaves. We jointly propose a Givens rotation sequence and cutpoints on the resulting grid by Metropolis-Hastings and establish reversibility with respect to the conditional posterior with leaf means integrated out. For additive functions with component-specific rotations and anisotropic Hölder smoothness, we prove posterior contraction in empirical $L_2$ distance and for the noise standard deviation. Under the stated prior, design, and grid conditions, with fixed numbers of predictors, trees, and components and no more components than trees, the rate is a sum of componentwise rates determined by smoothness and the number of rotated coordinates used. We also establish a posterior contraction lower bound showing that there exist functions for which RoBART adapts to the intrinsic dimension but axis-aligned BART does not.

stat.ML↗

Angular Baryon Acoustic Oscillations with DESI DR1

We present the angular BAO analysis using DESI DR1 data as a complementary approach to DESI's primary 3D BAO analysis. We perform the analysis using the two-point angular correlation function $w(θ)$ in tomographic redshift shells of $δz = 0.02, 0.05$, spanning the redshift range $0.4 < z < 2.1$ using the DESI tracers: LRG1, LRG2, LRG3+ELG1, ELG2 and QSO. The main benefit is that it requires no fiducial cosmology to convert redshifts into distances and yields a direct measurement of the transverse dilation parameter $α_{\perp}(z)$. We fit each tracer with a single shared $α_{\perp}(z)$ using templates constructed in a fiducial cosmology and a mock-based covariance from 1000 EZmock realisations. We measure $α_\perp$ to $2.5$ -- $4.2\%$ precision in each tracer bin. Compared to the $D_M/r_{\rm d}$ measurements obtained from the DESI DR1 3D analysis for the same tracers, our results are consistent within the statistical precision. We find no evidence for a shift in the transverse BAO scale between the 2D and 3D analyses. Flat $Λ$CDM fits to $α_\perp(z)$ alone give $Ω_{\rm m} = 0.198^{+0.045}_{-0.068}$, $1.6σ$ below the DESI 3D value, and the pull is dominated by the $z =1.49$ QSO bin, whose $α_\perp(z=1.49)=1.058\pm0.028$ sits $\sim 2σ$ above the fiducial. Restricting to LRGs and ELGs only, our $Ω_{\rm m}$ and $H_0 r_{\rm d}$ agree with the 3D transverse-only fit over the same bins to better than $0.1σ$. With an additional BBN prior, the full 2D sample gives $H_0 = 63.9^{+1.9}_{-3.3}\, {\rm km\,s^{-1}\,Mpc^{-1}}$. This pipeline is implementable on forthcoming DESI data releases, where the increase in volume and sampling density will determine if the QSO offset persists and will also tighten 2D BAO constraints.

astro-ph.CO↗

Extremal Sasaki manifolds and weighted K-stability

We remove the log-concavity assumption on the weight in the analytic weighted Yau--Tian--Donaldson correspondence established in [arXiv:2406.10939, arXiv:2407.09929, arXiv:2503.22183] and discuss applications to Sasaki and conformally Kähler Einstein--Maxwell geometries.

math.DG↗

Engineering of Chirped Apodized Sources for Quantum Spectroscopy

Quantum sources of light are of paramount importance for the development of photonic quantum technologies. The growing demand for bright, efficient, and spectrally engineered photonic sources for applications in quantum sensing, spectroscopy, and imaging has motivated the development of tailored nonlinear crystals. Here, we report on the development and characterization of chirped, apodized KTP crystals specifically engineered for quantum spectroscopy. Our source exhibits an unprecedented bandwidth in a highly non-degenerate configuration, demonstrating its potential to support quantum spectroscopy protocols over broad spectral ranges while enabling significantly shorter measurement times.

quant-ph↗

Finite-Sample Approximation of Hessian-Guided Perturbed Wasserstein Gradient Flows

Wasserstein gradient flow extends gradient descent to probability measures. Its Hessian-guided perturbed variant (PWGF) adds Gaussian perturbations to escape saddle points in nonconvex problems. We investigate when its approximation by finitely many interacting particles remains accurate over growing time horizons. Our analysis retains the curvature accumulated along the population-driven reference path: negative curvature can amplify approximation errors, while subsequent positive curvature can damp their influence. This captures favorable scenarios in which temporary instability is compatible with accurate tracking over growing horizons. Under regularity assumptions and a prescribed common perturbation schedule, we prove particle and objective-value tracking bounds on a high-probability event for reference paths satisfying explicit conditions on accumulated curvature. To handle state-dependent Gaussian jumps, we construct a population-first coupling that preserves the reference particles' conditional independence and reduces jump errors to covariance comparison. We verify the conditions in a variance-plus-cosine model, where curvature recovery yields a growing-horizon tracking guarantee. We also establish local attraction, transverse descent, and positive second variation in two regions of a regularized matrix-factorization model, motivating a positive-negative-positive curvature pattern.

cs.LG↗

Is the Light Neutralino Dark Matter Still Viable in the MSSM After the LZ-2024 Results?

In the Minimal Supersymmetric Standard Model with heavy sfermions, a sub-hundred-GeV Bino-like neutralino can thermally achieve the observed dark matter abundance only through $h$- or $Z$-resonant annihilation. Since the couplings governing these resonances simultaneously determine the spin-independent and spin-dependent scattering off nucleons, reproducing the relic density imposes irreducible lower envelopes on the scattering cross sections. We evaluate these bounds including the complete one-loop corrections to these vertices, which enhance the direct-detection rates by up to $50\%$ and $30\%$, respectively, adopt the FLAG-2024 nucleon matrix elements with their most adverse $3σ$ shifts, and confront the predictions with a likelihood-based recast of the LZ data. We find that the LZ-2024 results exclude both resonant scenarios at conservative one-sided significances exceeding $5.9σ$ and $4.1σ$, respectively, irrespective of whether the neutralino constitutes all or only part of the dark matter, and independently of LHC electroweakino searches.

hep-ph↗

Dynamics of Work Extraction in Multipartite Atomic Systems: Role of Correlations and Relative Entropy

We investigate the dynamics of quantum ergotropy in a multipartite system of two-level atoms interacting with a quantized field. The role of multipartite quantum mutual information and quantum relative entropy is analyzed to understand how correlations and the distinguishability between the passive state and the product of local states affect the extractable work. Numerical results for systems containing two to five atoms show that the maximum ergotropy does not increase linearly with the system size, while the intervals of zero ergotropy are gradually modified and suppressed for larger systems. We also study the effect of the average thermal photon number and different initial atomic states. Increasing the thermal photon number mainly reduces the relative-entropy contribution, whereas the initial state strongly influences both the magnitude and persistence of ergotropy. Among the considered states, the partially entangled state provides the largest and more persistent extractable work, while the mixed GHZ state gives the smallest ergotropy. These results highlight the importance of multipartite correlations, thermal effects, and state preparation in controlling quantum work extraction and may be useful for future studies of quantum batteries and quantum power.

quant-ph↗

The Domination Reciprocal Mean Square Index

We introduce the Domination Reciprocal Mean Square index \(DRMS\), a new domination-degree-based topological index obtained by replacing the ordinary degree in the reciprocal mean square index with the domination degree. We establish elementary bounds in terms of graph size and extremal domination degrees, derive a one-parameter family of sharp lower bounds via the power mean, and obtain sharp bounds in terms of several well-known domination-degree-based indices, with a full characterization of the equality cases. We also compute closed-form expressions for \(DRMS\) on several standard graph families and for the corona product of two graphs.

math.CO↗

Evaluating Sequence Assembly Strategies for Differentially Private Synthetic Time-Series Forecasting

Differentially private time-series generators commonly produce fixed-length synthetic windows, whereas downstream forecasting models often require long continuous training sequences. How these windows are assembled after generation can therefore alter the effective synthetic data presented to a forecaster, even when the trained generator remains unchanged. We study this post-generation sequence assembly process by systematically varying overlap rates and window-weighting schemes and evaluating the resulting sequences in terms of boundary continuity, statistical and temporal fidelity, and Train-on-Synthetic-Test-on-Real (TSTR) forecasting utility. Across four types of public datasets (ETTh1, ETTm1, Weather, and Appliances) and five forecasting models, the results reveal a clear forecaster-dependent assembly principle: downstream TSTR utility is jointly shaped by the forecaster, overlap rate, and window-weighting scheme, leading to distinct assembly preferences across forecasting models. Increased overlap generally improves boundary continuity, but improvements in continuity or individual fidelity diagnostics do not consistently reduce forecasting error, indicating that these diagnostics alone are insufficient for selecting assembly configurations. Complete five-forecaster assembly grids, together with matched Train-on-Real-Test-on-Real (TRTR) references, further characterize these regularities and quantify assembly-dependent utility relative to real-data training. We then validate the identified principles through additional analyses of robustness and generator variability.

cs.LG↗

Neutral Is Not Free: Evaluating Downside Risk in Neutral Launches

Evaluating "neutral launches" (e.g., infrastructure upgrades) using traditional confidence interval overlap is flawed: it is dangerously permissive with scarce data and excessively restrictive with abundant data. To resolve this, this paper introduces Expected Bayesian Loss (EBL), a continuous metric that quantifies both the probability and expected severity of metric degradation. Computable directly from standard frequentist estimates, EBL explicitly penalizes empirical noise and high-variance experiments. Validated against expert decisions, EBL provides experimentation platforms with a rigorous, tunable guardrail that aligns statistical safety with institutional risk appetite.

stat.AP↗

Generative AI in Publishing: An Editors' Panel on Ethics and Policies

Generative artificial intelligence (GenAI) is reshaping scholarly research faster than journals have developed stable norms for its use. This article presents an edited thematic account of a 2026 Joint Statistical Meetings panel that brought together editorial perspectives from mathematical statistics, data science, biomedical statistics, and general statistical scholarship. The discussion examines journal policies, disclosure, authorship and research integrity, peer-review confidentiality, researcher training, editorial workload, access, and possible future models of scholarly publishing. Panelists shared commitments to human accountability, the protection of confidential submissions, and disclosure of consequential assistance, while offering different recommendations on assistance with research ideas and proofs, disclosure requirements, automated review, and policy enforcement. By distinguishing shared principles from unresolved implementation questions, the article clarifies the choices facing statistical publishing and outlines an agenda for evaluating policies and practices as GenAI evolves. The account seeks to foster continued discussion of GenAI in scientific communication and encourage statistical organizations to develop more robust operational standards.

stat.OT↗

Masked Feature Encoding for Large-Scale Whole Slide Image Representation

Whole slide image (WSI) analysis in computational pathology follows a multiple instance learning (MIL) pipeline where patch embeddings are extracted independently and aggregated for slide-level prediction, but within-slide variance from staining, scanner, and local texture can overwhelm the discriminative signal. We propose Masked Feature Encoding for Multiple Instance Learning (MFE-MIL), a feature-space masking framework that trains a lightweight MLP adapter jointly with a window-based masked reconstruction branch and a MIL classification head. The two objectives are complementary. Classification guides the adapter to suppress within-slide patch variance, while window-based masked reconstruction provides an auxiliary regularizer for the adapted features without using patch coordinates, coordinate graphs, or segmentation preprocessing. The raster patch-extraction order is used only as a weak implicit prior. At inference, the decoder is removed, leaving only the adapter and MIL head. Across CAMELYON16/17, PANDA, and TCGA-BRCA with four diverse encoders, MFE-MIL improves ACC/F1 for nearly all tested aggregator-encoder settings and AUC in most, outperforms coordinate-based spatial methods (CAMIL), and achieves higher AUC than 2DMamba on three of four datasets (UNI). On five TCGA survival cohorts it improves the average concordance index for every aggregator tested, its most consistent gain. Code is available at https://github.com/AtlasAnalyticsLab/MFE-MIL.

cs.CV↗

Why Software Engineering Is Indispensable in the Age of Coding Agents

Can AI make Software Engineering (SE) -- the discipline -- obsolete? And can it make software engineers -- the professionals -- redundant? This paper argues that the rise of capable AI coding agents makes SE and software engineers essential, not obsolete: the missing foundation without which AI-assisted development produces misleadingly plausible, unverifiable, and ultimately untrustworthy software. Three structural properties of large language models (probabilistic generation, agnosticism, and semantic statelessness) create a structural vacuum that no amount of training can eliminate. Filling it requires four knowledge levers: methodological knowledge, domain knowledge, design choices, and process choices. All four must be reified as persistent artifacts, and each requires the software engineer as methodologist, mediator, and custodian.

cs.SE↗

From Prompts to Trees: Effective LLM-Guided Tree Generation for Few-Shot Tabular Classification

While Large Language Models (LLMs) possess rich world knowledge and impressive generalization capabilities, their direct application to tabular data classification is hindered by high inference costs and limited interpretability. In contrast, decision trees are fast and transparent but often underperform in low-data regimes. In this work, we propose a novel framework that bridges these paradigms by distilling LLM knowledge into interpretable decision trees under a few-shot learning setting. Instead of directly prompting the LLM to generate full trees, which is often unstable and inefficient, we develop a three-stage paradigm that prompts the LLM to generate rules and organize the rules into a tree. Experiments on multiple real-world tabular datasets demonstrate that our method achieves superior accuracy and interpretability with significantly lower prompting overhead compared to existing baselines.

cs.LG↗

Expected mixed volumes of convex hulls of random walks and Lévy processes

Let $C_1,\ldots,C_k$ be the convex hulls of independent partial-sum processes in $\mathbb{R}^d$ whose increments are exchangeable within each process. We express the expected mixed volume $\mathbb{E} V_d(C_1[m_1],\ldots,C_k[m_k])$, $m_1+\cdots+m_k=d$, through mean absolute determinants of disjoint block sums of the increments; no general-position assumption is needed. For random walks with i.i.d. integrable increments the block sums are independent, and the formula extends the expected-volume formula of Barndorff-Nielsen and Baxter and of Vysotsky and Zaporozhets to mixed volumes. We then prove a continuous-time counterpart: for independent Lévy processes with finite first moments, the expected mixed volume of the closed convex hulls of their paths is an explicit integral of mean absolute determinants over a product of simplices. For symmetric stable processes the integral is evaluated in terms of the associated zonoids. As a geometric application, we compute the mean mixed volume of random projections of mutually orthogonal canonical orthoschemes.

math.PR↗

On determination of convex bodies by their section functions

For a convex body $K$ its parallel section function in the direction $θ\in S^{n-1}$ is defined by $$A_{K,θ}(t) = \mathrm{vol}_{n-1} \left( K \cap \{ x \in \mathbb{R}^n ~ | ~ \langle x, θ\rangle = t \} \right) .$$ We study to what extent the body $K$ is determined by partial information about these functions. As applications, we characterize convex bodies with locally separable section functions, establish partial results on the homothety conjecture for bodies of flotation, and answer a question of Barker and Larman concerning the determination of convex bodies from section functions at infinitely many distances from the origin.

math.MG↗