arXiv Science⌕ Search

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,549 records · Page 86Linked to original sources

Postcritically finite endomorphisms I: Kobayashi hyperbolicity and tautness

Let $f:\mathbb{P}^N\to\mathbb{P}^N$ be a postcritically finite endomorphism of degree at least two. Combining Yamanoi's hyperbolicity results with the dynamical property of $f$, we prove that the complement of its postcritical divisor is taut unless $f$ is a monomial power map or admits a nontrivial equivariant rational fibration with a polarized divisorially postcritically finite base map. We establish analogous alternatives for polarized PCF endomorphisms of rationally connected projective manifolds and, for Kobayashi hyperbolicity, on normal rationally connected projective varieties. In dimension two, the non-taut case can be described, up to iteration and rational semiconjugacy, by monomial power maps and skew products with monomial fiber maps.

math.DS↗

Weighted Gagliardo--Nirenberg Inequalities of Grushin Type: Symmetries, Compactness, and Optimizers

In this paper, we study a family of weighted Gagliardo--Nirenberg interpolation inequalities associated with Grushin--type operators on $\mathbb{R}^n \times \mathbb{R}^m$. Specifically, we examine how to control the weighted norm $\||x|^βu \|_{L^B}$ by balancing fractional derivatives $D_x^s u$ with weighted derivatives $|x|^b D_y^r u$. Using two--parameter scaling, spatial translations, and Littlewood--Paley decompositions, we map out exact parameter conditions and demonstrate that the scaling exponents are both necessary and optimal. For multi-parameter weighted extensions, we leverage Khintchine-type random sums to show why $Q \ge 2$ is strictly required. When $x$ and $y$ are radially symmetric, two key advantages emerge: the allowable parameter range for the weights expands, and compact embeddings of the energy space into $L^p$, $2 < p < B$, are completely restored. We establish this compactness through two complementary tools, geometric covering arguments and a generalized three--norm version of Strauss's lemma, that yields explicit pointwise weighted decay estimates. Finally, in the scale--invariant case $β=0$, a concentration--compactness analysis proves that the inequality's sharp constant is achieved by an optimal function, with minimizing sequences remaining precompact up to the natural three-parameter group of symmetries.

math.AP↗

Finite volume discretization and normalized gradient flow for computing minimizers of the rotating Gross-Pitaevskii energy

We study the approximation by a finite volume scheme of the stationary Gross-Pitaevskii equation with angular momentum rotation. We prove that, under standard mesh regularity assumptions, the discrete global minimizers converge to the continuous one as the mesh diameter tends to zero. Moreover, under a spectral gap assumption on the Hessian of the discrete energy, we prove that solutions to a normalized continuous gradient flow converges exponentially fast to such discrete minimizers. Additionally, we show that this previous discrete spectral gap assumption can be derived from similar hypotheses in the continuous setting. We support our theoretical results with numerical simulations, which broaden the scope of the analysis.

math.NA↗

Federated Zeroth-Order Optimization with Direction Aggregation and Variance Reduction

We study constrained nonsmooth nonconvex stochastic optimization in federated settings, where clients access only stochastic function evaluations. Existing federated zeroth-order methods primarily combine local zeroth-order updates with model averaging. However, they struggle with client drift induced by local projected updates and sampling variance in stochastic zeroth-order estimates. In this paper, we propose a federated zeroth-order framework based on direction aggregation, equipped with two local sampling schemes. Specifically, FedZOO constructs a local direction from a minibatch shared across multiple spherical queries, thereby controlling the stochastic errors arising from data sampling and gradient approximation. In contrast, FedVRZO forms its local estimator from independent sample--direction pairs, so that its sampling error is controlled directly by the number of sample--direction pairs. In both algorithms, clients compute the mean of the directions evaluated along their projected local trajectories, and the server performs a single projected update using the weighted aggregate of the client directions. Furthermore, we establish convergence guarantees for FedZOO and FedVRZO, respectively. Experiments on black-box adversarial attacks demonstrate the effectiveness of the proposed methods.

cs.DC↗

The rank of $3\times 3$ matrix multiplication over $\mathbb{F}_2$ is 23

The rank of the tensor of $3\times 3$ matrix multiplication over the field with two elements is at most $23$ by Laderman's algorithm, and Rudich and Rousseau recently proved that it is at least $22$. We prove that it equals $23$. Hence Laderman's algorithm uses the fewest multiplications among all bilinear algorithms over $\mathbb{F}_2$ and among all bilinear algorithms with integer coefficients. The proof uses the substitution method in the form developed in recent work of D'Ambrosio, Wang and Yang et al.: a subspace $S$ of the space of first factors contains at most $r-R(S)$ first factors of a decomposition of length $r$, where $R(S)$ is the rank of the tensor modulo $S$. We raise the known lower bounds on $R(S)$ for $111$ of Wang's $496$ symmetry classes of subspaces. One of these bounds, $R(S)\ge 21$ for a point spanned by a matrix of rank one, forces the $22$ first factors of a decomposition of length $22$ to be distinct. A $27\times 27$ flattening of the tensor gives further constraints on the ranks of the first factors, and a separate enumeration shows that, when at least $14$ first factors have rank one, no line in a certain orbit of lines contains two first factors. A computer search then lists, up to symmetry, all sets of $22$ matrices that satisfy these constraints, and an exact completion search shows that none of them is the set of first factors of a decomposition. The computation emits certificates, which are checked in the Lean 4 proof assistant by checkers whose soundness is proved in Lean. The largest checks are evaluated as compiled code, so the proof relies on the Lean compiler in addition to its kernel.

math.RA↗

SmartBike: A Low-Power Edge IoT System for Cycling Safety and Real-Time Awareness

Urban cycling is a sustainable mode of transportation, but cyclists remain highly vulnerable in dense traffic because of limited situational awareness and the lack of active safety assistance on conventional bicycles. This paper presents SmartBike, a low power Internet-of-Things (IoT) platform for cyclist safety, combining blind spot monitoring, rider usage detection, localization, and wireless edge-to-gateway communication. The proposed system consists of a bike-mounted IoT node and a gateway assisted monitoring interface. The node integrates three time-of-flight sensors for blind spot detection, an IMU module for motion sensing and on-board speed estimation, a GNSS module for localization, and a saddle mounted piezo-electric transducer for event-driven wake up. To satisfy the energy constraints of the IoT node, the design employs piezo-triggered activation, application specific sensing rates, and compact connectionless Bluetooth Low Energy (BLE) advertisements for low overhead transmission. The firmware is implemented in Zephyr RTOS using a lightweight state machine architecture, while the gateway performs passive BLE scanning, payload decoding, and real-time dashboard visualization. The BLE payload is reduced to 11 bytes, and power analysis shows that sensing accounts for more than 90% of the total energy budget, while wireless communication represents less overall consumption. With a 12600 mWh battery, the estimated lifetime reaches approximately 21 months under typical commuting conditions (Average 40 min per day).

eess.SP↗

Self-Supervised Speech Representations for Cross-Speaker Dysarthria Detection During Awake Craniotomy

Detecting intra-operative speech impairment during awake craniotomy is essential for preserving language function. However, automated detection remains challenging because operating-room recordings contain substantial acoustic interference, clinically relevant speech events are rare, and available cohorts are small and heterogeneous across speakers. This study presents a systematic component-wise evaluation of a pipeline for distinguishing dysarthric from no-trouble speech in the DATABRASE corpus of awake-craniotomy recordings. The pipeline incorporates speaker diarization to isolate patient speech, a multi-view representation combining handcrafted acoustic descriptors with multilayer wav2vec 2.0 embeddings, speaker-conditional normalization and transferability-based feature selection to improve cross-speaker robustness, and a cascaded classifier comprising a gradient-boosted first stage and a neural second stage. Evaluation was conducted under strict speaker-independent conditions using leave-one-speaker-out cross-validation. The results show that cross-speaker performance is influenced more strongly by the speech representation than by classifier choice. The AUCs of three classifiers differed by no more than 4.7%, whereas replacing conventional acoustic descriptors with the multilayer self-supervised representation produced AUC improvements of 18.2%-26.1%. Diarization-conditioned feature extraction and the proposed classifier cascade provided additional consistent gains. These findings indicate that reliable patient-specific speech isolation and strong pretrained representations are more important than increased classifier complexity in low-resource intra-operative settings. They also quantify the potential performance gains that may be achieved through patient-specific preoperative calibration.

cs.LG↗

From Suppression to Repair: Mitigating Object Hallucination in Large Vision-Language Models via Localized Distribution Alignment

Object hallucination remains a major obstacle for large vision-language models (LVLMs) to generate reliable content. An intuitive mitigation strategy is to suppress hallucination-related components in hidden representations. However, these components may also contain useful information, and suppressing them can weaken the model's multimodal capabilities. In this paper, we propose ResOT, a training-free method that repairs representations at inference time through localized distribution alignment. Specifically, ResOT projects dominant hallucinated directions away from the faithful subspace, forming a low-dimensional residual subspace for intervention. Within this subspace, ResOT uses Gaussian optimal transport (OT) to align the hallucinated distribution with the faithful one. The resulting map defines repair targets with minimal changes to the original representations. At inference, ResOT adaptively controls how far each token state moves toward its OT target. Experiments on three representative LVLMs show that ResOT substantially reduces object hallucination while improving image caption quality and multimodal performance across multiple benchmarks. Code will be released.

cs.CV↗

Strichartz estimates for Klein-Gordon equations on asymptotically hyperbolic manifolds with applications

In this paper, we establish global-in-time Strichartz estimates for Klein-Gordon equations on non-trapping asymptotically hyperbolic manifolds. Our result is a positive-mass counterpart of the global Strichartz estimates of Sire, Sogge, Wang, and Zhang for shifted wave equations, and recovers the same range of admissible pairs. The main new feature is the low-frequency analysis, where the positive mass changes the spectral phase and leads to a different dispersive mechanism. Combining the resulting low-frequency estimates with microlocalized high-frequency estimates yields the full global Strichartz estimates. As applications, we obtain small-data global existence and scattering for power-type and certain derivative semilinear equations, as well as small-data global existence for a class of smooth semilinear equations without imposing a null condition.

math.AP↗

Manifold Regression

Conventional statistical modeling typically assumes a one-to-one mapping between input and output variables, where one-to-one is used in the predictive sense that a specified input results in a unique output. Many scientific and engineering systems, however, exhibit non-one-to-one(noto) input-output relations. In a noto relation, the same input may correspond to multiple admissible outputs, so the conventional statistical models may not be appropriate. This paper develops manifold regression, a parametric modeling framework for such noto input-output relations, which represents the underlying input-output relation through a latent manifold. The manifold is specified upto an unknown finite dimensional parameter and we estimate the unknown parameters from ordinary input-output observations by regularized profile optimization with a robust solver. Once the manifold has been learned, prediction is obtained by slicing the estimated manifold along a specified coordinate value. Because a slice may contain multiple latent roots, the resulting manifold prediction is naturally set-valued. This formulation extends predictive learning beyond conventional one-to-one statistical models while containing classical regression and inverse prediction as special cases. Theoretical results are established to justify the manifold regression model and its sliced prediction sets; and case studies are used to demonstrate stable performance across representative noto cases.

stat.ME↗

A rank bound for bases and circuits in binary matroids

Let \(b(M)\), \(d(M)\), and \(r(M)\) denote the number of bases, the number of circuits, and the rank of a matroid \(M\), respectively. We prove that every nonempty simple binary matroid with no coloops satisfies \[ 2b(M)\ge(r(M)+1)d(M), \] with equality if and only if \(M\) is isomorphic to the Fano matroid. This confirms a conjecture recorded by Oxley in 1983. For the same class, we prove that deleting any element leaves at least as many bases as there are circuits in the original matroid: \(b(M\backslash e)\ge d(M)\) for every \(e\in E(M)\). We determine all equality cases and deduce a sharp linear lower bound for basis growth under successive series extensions. The main counting step is a joint estimate for the three largest possible circuit sizes, obtained from contraction-normalized fundamental-circuit counts and an exact folded-cube edge correspondence.

math.CO↗

Solution-Based Synthesis of Fe-Co-Ni Prussian Blue Analogue Powders: A Comparative Structural, Spectroscopic and Thermal Study

Prussian Blue (Fe-PBA) and its Co and Ni analogues (Co-PBA and Ni-PBA) were synthesized by an additive-free aqueous co-precipitation route using K$_4$[Fe(CN)$_6$] and FeCl$_3$, CoCl$_2$, or NiCl$_2$, respectively. The three powders were systematically compared by X-ray diffraction, ATR-FTIR and Raman spectroscopy, FEG-SEM/EDS, and TG/SDTA. All compositions exhibit the characteristic cubic cyanide-bridged PBA framework, with apparent lattice parameters of 10.10 +/- 0.02, 10.01 +/- 0.02, and 10.10 +/- 0.03 Angstrom for Fe-PBA, Co-PBA, and Ni-PBA, respectively. The dominant C$\equiv$N stretching band in ATR-FTIR shifts from 2062 to 2071 and 2087 cm$^{-1}$ across the Fe-Co-Ni series, indicating composition-dependent changes in the local cyanide environment. Fe-PBA contains less potassium than the Co- and Ni-containing powders and exhibits a substantially larger low-temperature mass loss. Sharp reflections assigned to crystalline KCl are observed in all three diffraction patterns, while additional unassigned reflections in Ni-PBA indicate the presence of at least one further crystalline phase. These results show that differences among Fe-, Co-, and Ni-based PBAs cannot be attributed solely to transition-metal identity, as precursor oxidation state and washing efficiency also influence the composition and thermal response of the resulting powders.

cond-mat.mtrl-sci↗

Controlling transitions between nonequilibrium states through active bath engineering

We develop a family of control protocols for finite-time transitions between active nonequilibrium steady states using the noise-color (correlation rate) as the sole control parameter. By reverse-engineering the second moment dynamics of an active Ornstein-Uhlenbeck process, we determine the time-dependent correlation rate required to realize a prescribed evolution of the system's state. We experimentally implement these protocols with a micrometer-sized optically trapped particle coupled to an engineered active bath, demonstrating transitions substantially faster than the natural relaxation while keeping the confining potential, noise amplitude, and temperature fixed. Our approach extends engineered swift-equilibration methods to far-from-equilibrium steady states by directly controlling the temporal correlations of an active environment. We show how physical constraints impose a speed limit on finite-time transitions, while the freedom in choosing the prescribed system evolution can be exploited to eliminate control discontinuities, minimize the admissible transition time, or optimize a thermodynamic cost.

cond-mat.stat-mech↗

Investigating the nature of magnetic turbulence in Tycho's SNR using X-ray observation

Supernova remnants (SNRs) are widely regarded as the primary sources of Galactic cosmic ray acceleration and the particles are energized at the shock front of SNRs through the diffusive shock acceleration (DSA) mechanism, gaining energy by repeatedly crossing the shock. Magnetic turbulence plays a crucial role in scattering these particles back and forth, making the investigation of turbulence essential to understanding the acceleration process. In this study, we apply the two-point correlation method to derive the magnetic energy spectrum from the non-thermal X-ray flux image of Tycho's SNR. This image was obtained using Poissonian-based Generalized Morphological Component Analysis (pGMCA). The turbulence length scale observed in the resulting spectrum is shorter than the turbulence scale derived from previous studies of radio observations of Tycho's SNR. Additionally, we use the X-ray rim thickness measured across Tycho which is an indicator of magnetic field strength to estimate the magnetic energy spectrum. The turbulence scale inferred from this method is consistent with that obtained from the non-thermal X-ray flux analysis.

astro-ph.HE↗

Probability-Signature Dynamics: Unpacking Modular Addition Learning Within Two-Layer Networks

Neural networks trained on modular addition tasks often develop Fourier-structured representations that support exact generalization. While prior work has identified these Fourier circuits, the mechanism by which gradient-based training selects them from the data distribution remains unclear. We address this question using probability signatures, which express leading gradient interactions through conditional statistics of the training distribution. For modular addition, these signatures are cyclic shift operators and are diagonalized by the discrete Fourier transform, yielding approximately decoupled Fourier-mode dynamics. This explains the emergence of Fourier sparsity, frequency matching, and phase alignment. The same framework resolves a puzzle under label noise: corrupted examples can show faster early loss decrease than clean examples, despite lacking a coherent generalization rule. We show that noise increases conditional label collisions, strengthening early shared-coordinate reinforcement. Finally, this method can be applied to other operators. Taking XOR as an example, we observed the predicted frequency in experiments.

cs.AI↗

Recovery Guarantees for Posterior Sampling of One-Bit Compressed Sensing

We study the sample complexity of noisy one-bit compressed sensing for signals drawn from a prior distribution. By characterizing the effective distributional complexity of the prior via its approximate covering number, we prove that posterior sampling achieves accurate recovery with high probability when the number of measurements scales with the logarithm of the approximate covering number, up to a one-bit separation gap factor. This upper bound is robust to learned prior mismatch. Specifically, we show that posterior sampling with an approximate prior remains reliable, provided that the learned prior distribution is sufficiently close to the true signal distribution in Wasserstein distance. In addition, we establish a sample complexity lower bound for any reliable method of noisy one-bit compressed sensing, showing that our upper bound is nearly matched in its main prior dependent term. To approximate the ideal posterior sampling process for real world scenarios, we instantiate posterior sampling through a plug-and-play algorithm with diffusion priors. Experiments on the FFHQ and ImageNet datasets demonstrate the effectiveness of our proposed approach.

cs.LG↗

On the Risks of using LLM-Generated Tests for Regression Testing

Software is under constant evolution: developers continuously add features, fix bugs, and refactor code, and any of these changes may break existing functionality. Regression testing guards against such effects by capturing expected behavior in test cases. LLM-based test generation aims to automate this process by generating regression tests directly from the code under test. This is beneficial when the implementation is correct, but problematic when the code contains faults: the generated tests may then encode and preserve incorrect behavior. To investigate this risk, we apply LLM-based regression test generation to pull requests merged into the main branch of software projects and study the impact of the generated tests on subsequent project evolution. We distinguish between fault-revealing tests, which assert correctly implemented behavior, and fault-enforcing tests, which assert faulty behavior. Across 145 pull requests from SciPy, Qiskit, and pandas, 8%-17% of the generated tests are fault-enforcing, while only 2.4%-4.8% reveal faults. Fault-enforcing tests persist over time: after several subsequent commits, 83%-91% of them are still relevant and pass. They also accumulate: when the faults of all pull requests are combined in one codebase, 83%-92% remain enforced at the end of the commit history, and the developer-written test suite detects only 14%-30% of them. Our results reveal a fundamental risk of LLM-generated regression tests: without manual validation, they may encode faulty behavior as expected behavior, allowing bugs to persist across software revisions and largely evade developer-maintained test suites. LLM-based regression testing can thus give rise to a new form of technical debt.

cs.SE↗