arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,603 records · Page 89Linked to original sources

Sheeting property and rigidity for stable solutions to the Allen--Cahn equation

We prove a sheeting theorem for stable solutions to the Allen--Cahn equation in ambient dimensions $2$ through $10$. We show that weak convergence of the measures $\varepsilon|Du_\varepsilon|^2\,dX$ to a hyperplane with integer multiplicity implies that the zero sets are disjoint smooth graphs. As a consequence, stable entire solutions in $\mathbb R^N$, $N\le7$, with $O(R^{N-1})$ energy growth are one-dimensional. When $W''(-1)=W''(1)$, every blowdown of an entire solution with finite Morse index and this energy growth has multiplicity one for $4\le N\le10$. For even potentials, we also obtain multiplicity one for limits with bounded energy and Morse index on closed manifolds of dimensions $3$ through $7$, under bumpiness or positive Ricci curvature.

math.AP↗

Character expansions and affine Jacobi-Trudi identities

We provide character expansions of certain specialisations of multiparameter Hall--Littlewood polynomials of types $\mathrm{B}_n$, $\mathrm{C}_n$, and $\mathrm{BC}_n$ for rectangular shapes. We use these expansions to establish the six Jacobi-Trudi-type identities conjectured by Ole Warnaar in 2025.

math.CO↗

Language-model ratings of depression reflect the rater more than the patient

Depression has no diagnostic blood test. Language models promise tireless, consistent assessment, but can accurate raters disagree about individuals? We pre-registered 880 language-model raters, crossing 11 open models with prompting and scoring choices, and applied them to 189 interviews against the eight-item Patient Health Questionnaire. Model choice explained 30.0% of summed-symptom score variance, stable participant differences 10.5%. Two randomly drawn raters with area under the receiver operating characteristic curve (AUC) >= 0.70 disagreed on screening decisions for 40% of participants, on average. Average over-rating governed how many were flagged, yet equal-capacity raters chose differently for about one participant in five. A locked analysis of 86 new interviews reproduced the main pre-registered findings. Exploratory recalibration with 40 labelled participants raised accuracy from about 60% to 75% and halved disagreement, leaving one participant in five decided differently. Calibration repaired much of the rater dependence without securing agreement about individuals.

cs.CL↗

X-OPM: Explainable Automatic Digital On-Chip Power Modeling for Enhanced Robustness

Proactive power management systems reduce processor dynamic power through runtime power prediction and power-aware scheduling. Accurate, stable and low-overhead digital on-chip power meters (OPMs) are crucial for improving the prediction quality. Recent studies have explored various modeling methods, including using linear models, decision trees, and multi-layer perceptrons (MLPs) to construct OPMs. However, most current approaches train models end-to-end without analyzing the physical interpretability of features, affecting their ability to generalize to unseen workloads. Grounded in the design principles of synchronous digital VLSI circuits, X-OPM introduces a robust feature engineering framework that uses tree-based models to capture feature interactions and linear models for prediction. It also incorporates a human-in-the-loop workflow to balance model accuracy against modeling effort. Evaluated on a commercial C906 vector processor, X-OPM consistently achieves $R^2 > 0.93$ across all workloads with sampling window size set below $8$ cycles. In contrast, state-of-the-art methods including APOLLO, COBIT, and standard MLPs fail to generalize across all test cases. Layout with commercial EDA tools shows that X-OPM incurs an area overhead below $0.1\%$, which is on par with lightweight tree-based and linear models, and significantly smaller than MLP-based models.

cs.AR↗

Replica Fragmentation and Glassy Dynamics in Parity Learning

We study how independently trained Transformer neural networks reconstruct a binary string from its local domain walls. Runs sharing the data and training protocol can realize different functions. We treat them as replicas and measure truth alignment $m$, prediction confidence $q_{\mathrm{self}}$, and cross-replica agreement $q_{\mathrm{cross}}$. Confident disagreement defines the finite-size replica fragmentation that we call glass-like. With small training sets, replicas predict all training examples correctly but remain confident in incorrect predictions for unseen inputs, a regime we call memorization. With larger sets, runs can generalize and then retreat. Retreat occurs when outputs start to deviate from truth while confidence remains high. The frontier between learned and unlearned outputs recedes toward shorter strings. Many later recover as the frontier advances again. The self--cross gap $q_{\mathrm{self}}-q_{\mathrm{cross}}$ clearly distinguishes the three learning regimes of memorization, retreat, and recovery. With overall and position-dependent truth alignment subtracted, their residual correlations also differ: memorizing replicas have weakly and uniformly correlated residuals, while retreat has the largest fraction of replica pairs whose residuals are anti-correlated. That is, on the inputs where one replica of such a pair does better than its average, the other tends to do worse. These finite-size observations distinguish persistent memorization from ongoing retreat--recovery dynamics. The thermodynamic and long-time limits, and extensions to other learning tasks, remain open.

cond-mat.stat-mech↗

Explicit tensors beyond the linear flattening barrier

We construct explicit n x n x n tensors of border rank at least 3n-o(n), improving the previous record of (2 + $\varepsilon$)n due to (Landsberg and Michalek 2025). We also prove that the linear flattening method cannot be used to establish lower bounds on border rank beyond 2n-1. When n is odd, we prove that this bound is achieved by Koszul flattenings. When n is even, a result of (Landsberg 2015) shows that Koszul flattenings can achieve 2n-2, leaving an open gap of size one. Our 2n-1 bound improves the best known 6n-4 linear flattening barrier due to (Garg et al. 2019) and (Buczyński 2026). Combined, these results show that our 3n-o(n) construction, as well as the construction of Landsberg and Michalek, provide explicit examples of tensors with higher border rank than any linear flattening can achieve.

cs.DS↗

Ultra-small-angle graphene twistronics with monolayer spacers: From interlayer hybridisation and moiré effects to high-temperature magnetotransport oscillations

Based on the analysis of moiré superlattice characteristics and interlayer hybridisation in small-angle twisted graphene multilayers with a monolayer hexagonal boron nitride (hBN) spacer, we predict spectral features and quantum transport effects in these systems. Here, we consider both a twisted bilayer and double bilayers with a spacer, discuss double bilayers in two stacking orders - with and without inversion symmetry - and also compare the results obtained with self-consistent Hartree calculations with and without in-plane periodic charge distribution, offering computationally efficient models for these graphene/hBN/graphene moiré systems. We note parametric intervals where the appearance of minibands can be enhanced by tuning with an out-of-plane displacement field, and identify the miniband edges and van Hove singularities across a broad range of densities and displacement fields. These system host Lifshitz transitions at which the Fermi lines form a network across moiré Brillouin zones, and to which Brown-Zak oscillations converge at low magnetic field and elevated temperatures. We analyse interlayer hybridisation across the same parameter space, discuss how hybridisation affects in-plane magneto-transport in ultra-high-mobility structures, and predict that semiclassical magneto-oscillations of interlayer hybridisation can lead to 1/B-periodic quantum oscillations of resistivity persistent to high temperature.

cond-mat.mes-hall↗

Closing the realism gap in physics-based gait simulations with a learned state prior

Predictive simulation of human movement is a promising tool for studying ``what-if'' scenarios in human movement and its underlying motor control, yet its realism is often limited. To address this gap, we incorporate a learned state prior that is trained on a large-scale dataset of human gait kinematics and external forces into predictive simulations. Resulting gait simulations yield kinematics and kinetics across diverse walking and running speeds that better match experimental data than current physics-based simulations, achieving accuracy comparable to data-driven models that reproduce learned data. Furthermore, our method enables robust hypothesis testing by demonstrating how varying optimality assumptions, muscle weakness, and footwear choices influence predicted gait. We also show that this prior generalizes well beyond its training data, successfully reconstructing full-body kinematics for curved running and cutting maneuvers from sparse marker sets. Ultimately, these results suggest that state priors should be broadly integrated into predictive simulations.

cs.CE↗

Budget-Constrained Multi-Consensus Decentralized Gradient Descent

We investigate decentralized gradient descent (DGD) with emphasis on efficient communication and computation resource utilization under budget constraints. As a first step toward the broader communication-computation allocation problem, we consider and analyze a \textit{multi-consensus decentralized gradient descent} (mcDGD) scheme, where the number of consensus rounds and the stepsize are allowed to vary across iterations. Building on a unified analytical framework for DGD, we derive finite-time convergence bounds that explicitly characterize the interaction between consensus quality and optimization dynamics. Our analysis requires only convexity of the local objective functions while assuming smoothness and strong convexity of the global objective. The resulting bounds enable a principled consensus-allocation strategy under resource constraints, for which we show that equal allocation of consensus rounds across iterations is optimal under our stepsize rule, up to integer rounding. Numerical experiments corroborate the theoretical findings and demonstrate favorable communication-computation tradeoffs compared with existing multi-consensus decentralized optimization baselines.

math.OC↗

Exact Catalysis Cannot Overcome the Gaussian-Steering Barrier for Remote Wigner Negativity

Gaussian steering sets a sharp threshold for creating Wigner negativity at a distance. A measurement by Bob can make Alice's part of a shared Gaussian state Wigner negative only if Alice can steer him with Gaussian measurements. We ask whether multiple copies, ancillas, or catalysts can lift this requirement when Alice is restricted to Gaussian operations. We introduce the steered negativity, a multiplicative monotone whose logarithm equals the Gaussian steerability for Gaussian states. It bounds the negativity that Bob can herald and rules out activation by copies, by ancillas of Bob, and by catalysts without remote negativity of their own. For a strictly unsteerable Gaussian state and a single round of feed-forward to a Gaussian unitary of Alice, we further prove that even a Wigner-negative or steerable catalyst cannot help if it is returned exactly. Gaussian steering thus remains a barrier to remote Wigner negativity beyond the single-copy setting.

quant-ph↗

Cylindrical Geodesic Flow Matching for Quasiperiodic Physiological Signal Transformation

Paired translation between quasiperiodic physiological waveforms (i.e., recovering a target oscillatory signal from the source) is central to the interpretation of cardiovascular signals derived from wearables placed at different body locations. This source-to-target mapping in these problems carries inherent geometric structure: the phase wraps around the cycle and must be treated as a circular variable, the amplitude remains strictly positive, and the beat-to-beat alignment can drift unpredictably across cycles and subjects. While deep neural networks have been used for phase estimation and complex-valued signal modeling, prior work does not explicitly learn phase transport between paired signals. Consequently, neither endpoint-supervised regression nor the standard affine path used in flow matching accounts for this phase--amplitude structure. We introduce \emph{cylindrical geodesic flow matching} for paired cardiovascular waveform translation. We show that the standard affine path used in flow matching distorts intermediate amplitude and instantaneous frequency when interpolating between quasiperiodic signals; replacing it with a closed-form geodesic on the phase--amplitude cylinder eliminates these artifacts and converts each training pair into dense, geometry-consistent velocity supervision. On zero-shot photoplethysmography and limited-support seismocardiography adaptation benchmarks, our method consistently outperforms interpolation baselines and matches or exceeds direct supervised prediction, reducing Hilbert Transform, $L_2$, and Dynamic Time Warping distance by up to ${\sim}15\%$ over the strongest competing baseline. These results suggest that bridge geometry is a critical inductive bias for flow matching on oscillatory signal translation.

cs.AI↗

Groupoid strict comparison for Ample Groupoids: Fiberwise Supramenability and Topological Amenability

In this paper, we first introduce fiberwise supramenability for locally compact Hausdorff étale groupoids with compact unit space. Then, for such a minimal ample groupoid $\mathcal{G}$, we show that if $\mathcal{G}$ is fiberwise supramenable or topologically amenable, then the clopen type semigroup $S(\mathcal{G})$ is almost unperforated. Consequently, if $\mathcal{G}$ is $σ$-compact, then it has groupoid strict comparison. We then present several applications, including Matui's AH conjecture and pure infiniteness for groupoids and groupoid $C^*$-algebras.

math.DS↗

Dynamical vertex approximation for retarded interactions: Application to the Hubbard-Holstein model

We extend the dynamical vertex approximation (D$Γ$A), in its most commonly used ladder-version, to systems with retarded electronic interactions. Such an extended scheme, which we name "dynamical-$U$ D$Γ$A", allows us to study strongly interacting electrons coupled to bosons including non-local spatial correlations on all length scales on top of the purely local ones captured by dynamical mean-field theory. We demonstrate the applicability of the method in the specific case of the coupling to a single dispersionless phononic mode, by studying the Hubbard-Holstein model on the square and cubic lattices, focusing on regimes of sizable Hubbard interactions where strong electronic correlations compete with phonon-mediated interactions. As illustrative examples, we consider the antiferromagnetic phase in the three-dimensional model and the d-wave superconductor in two dimensions, analyzing the effect of the electron-phonon coupling, including its frequency dependence, on both instabilities.

cond-mat.str-el↗

Wiki-Talkie: Multilingual Benchmarking of Persona-Based Agents on Real-World Discussions

LLMs are increasingly deployed as autonomous agents in social environments, making it critical to study their ability to faithfully simulate human interactions. Central to this is grounding agents in realistic user personas, yet existing datasets rely on fictional personas and are limited to a handful of languages, lacking the empirical grounding necessary to evaluate behavioral fidelity across diverse populations. We introduce Wiki-Talkie, a multilingual dataset of real-world conversations from Wikipedia Talk pages across five languages spanning two language families: Germanic (German, English) and Romance (Spanish, French, Italian), paired with personas derived from real user communities and encompassing sociodemographic attributes, self-descriptions, and behaviorally grounded interaction traits. Using Wiki-Talkie, we evaluate agent interactional behavior on a next-turn generation task across various persona conditioning strategies. Our evaluation assesses whether agents collectively reproduce the distributional behavioral patterns observed in human discussions. Results show that user's comment history exemplifying interaction behavior consistently outperforms explicit persona information. In addition, models systematically underproduce negative or extreme sentiments, while over producing references and suggestions, revealing biases toward agreeableness and positivity. Crucially, these patterns hold robustly across languages, with small cross-lingual differences.

cs.CL↗

How Much Evidence Should a Coding Agent's Self-Correction Carry? Adaptive Dirichlet Evidence for Self-Distillation

Execution feedback lets coding agents revise programs and learn from their own corrections. A correction's learning weight should reflect both the transitions supported by its executions and the amount of evidence behind that support. We introduce Effective-Evidence Self-Distillation (EESD), which represents these quantities separately. Normalized execution relevance determines relative transition support and an effective pseudo-count mass; a Dirichlet posterior then produces an uncertainty-penalized weight for KL-anchored correction learning. Under a symmetric prior, changing mass preserves category ordering, and effective mass yields a supervised coefficient bounded by its matched fixed-mass counterpart. Across four model-domain history sweeps, increasing visible observations from one to eight reduces future-outcome NLL by 55.0-59.3%. At eight observations, effective mass achieves lower NLL than fixed mass in all four comparisons. In the primary matched DeepSeek/RunBugRun study, argmax predictions agree on all 3,000 examples, with the largest NLL gain under concentrated relevance. After one correction-learning round, DeepSeek/CodeARC all-tests Pass@1 increases from 15.0% to 20.4%, with a paired 95% source-bootstrap interval of [+2.8, +8.0] percentage points. The twelve-setting downstream evaluation establishes the model-domain scope of this update. These results show how separating evidence support from evidence mass changes probability estimation and correction learning in coding agents.

cs.AI↗

Gauge-covariant Hadamard descent in Yang-Mills theory

We study the dimensional reduction of non-Abelian Yang-Mills theory from five to four dimensions. In this framework, Hadamard descent is formulated as a restriction to gauge-equivalence classes that are invariant, up to gauge transformations, under translations along the reduced direction. The central result is that, prior to gauge fixing, the adjoint scalar of the reduced theory is not the connection component $A_u$ itself, but the gauge-covariant combination $\widetildeΦ=ε-ι_YA$, which follows naturally from the covariant Cartan identity. The usual identification with $A_u$ is recovered, up to the sign convention adopted here, in the $u$-independent gauge, while global holonomy data remain additional information to be specified separately. We show that the reduced theory takes the form of a Yang--Mills--Higgs theory with the Higgs field in the adjoint representation. As an example, we focus on the $U(2)$ gauge theory and analyze the Higgs mechanism, together with the associated symmetry breaking pattern, when the reduced theory is expanded around the ground state of the parent theory. We then include fermions and show that, in the $u$-independent gauge, the extra-dimensional component of the gauge connection generates a pseudoscalar Yukawa coupling which, on the vacuum configuration, can be converted into a fermion mass term by a chiral field redefinition, while the gauge interaction remains vectorial.

hep-th↗

Optimal GHZ extraction from MABK violations

We determine the least Greenberger-Horne-Zeilinger (GHZ) extractability compatible with a Mermin-Ardehali-Belinskii-Klyshko (MABK) Bell score for every number of parties greater than two. Extractability is the largest squared overlap with a GHZ state obtainable by separate local quantum channels. Between the biseparable bound and the quantum maximum, the exact minimum is the affine interpolation from one half to one. The bound holds for normal states on tensor products of arbitrary local dimension and arbitrary binary measurements. We settle the previously unresolved range of six or more parties with an analytic proof that is uniform from four parties onward. We also determine the exact minimum when every local system is a qubit, and show that it lies strictly above the unrestricted bound at every interior score. One qutrit and qubits at all remaining parties attain every point of the bound with measurements fixed as the score varies. For at least four parties and strict interior scores, we classify all states attaining the bound with the minimum product of local support dimensions. For at least four parties on one qutrit and qubits elsewhere, we also prove that, when the extractability excess above the affine minimum is small, the trace-norm distance from the attaining mixture, up to local unitaries, is bounded by a constant times the square root of that excess. The exponent $1/2$ is optimal, and the constants are independent of the number of parties on every fixed interior interval of normalized scores.

quant-ph↗