arXiv Science⌕ Search

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,567 records · Page 87Linked to original sources

Boundary integral equation based solvers for elastic transmission problems in complex media

We present a boundary integral equation (BIE) framework for two-dimensional time-harmonic elastic transmission problems. Using the Helmholtz decomposition, the elastodynamic displacement fields are expressed in terms of scalar pressure and shear potentials, so that only the four classical Helmholtz boundary integral operators are needed. The key step is a rewriting of the Navier traction operator in the local normal--tangential frame of the interface, in which second-order normal derivatives are eliminated in favour of tangential derivatives and curvature terms. This leads to a one-parameter family of indirect combined field integral equations, depending on a coupling constant $α\ne 0$, which we prove to be uniquely solvable. The principal part of the resulting operator is, however, defective, and the discretized systems are severely ill-conditioned. Using the pseudodifferential calculus of the Helmholtz BIOs, we construct explicit left and right analytical preconditioners and prove that the preconditioned operator is a compact perturbation of the identity. Numerical experiments based on high-order Nyström discretizations confirm that the preconditioner clusters the spectrum: the number of GMRES iterations becomes independent of the discretization and is reduced by one order of magnitude or more, also in the high-frequency regime and for non-convex scatterers.

math.NA↗

Elastic polymer networks of exceptional strength by deconcentrating tension

The strength of a polymer network is orders of magnitude below that of a polymer chain, because the network concentrates high tension in a small fraction of polymer strands. Here we show that the strength of a network can be greatly amplified by recruiting a larger fraction of strands to bear high tension. We develop approaches to fabricate hydrogels of exceptional strength while maintaining low hysteresis. The strength of the hydrogel increases from ~0.05 MPa for a regular network, to ~1 MPa for a highly entangled network, and further to ~10 MPa for a prestretched interpenetrating network. Similar amplifications of strength are achieved for elastomers. Furthermore, experimental data suggest a scaling relation between strength and strand length. This work provides design principles for creating elastic and strong polymer networks.

cond-mat.soft↗

Magnon band splitting without altermagnetism in CuF2

Conclusive identification of an altermagnetic state requires going beyond mere symmetry arguments. We illustrate this in a combined computational and experimental study of the rutile-like material CuF$_2$, which is on the list of predicted altermagnets. Using ab initio and linear spin-wave calculations supplied by magnetization measurements, we show that CuF$_2$ in its experimental monoclinic structure can be described by a spin-$\frac12$ model of weakly coupled square-lattice layers with the in-plane coupling $J_1\simeq 115$ K and two synergistic antiferromagnetic interplane couplings amounting to 4% and 8% of $J_1$, respectively. Driven by long-range superexchange, these interlayer couplings are oblique to the square planes, resulting in the unit-cell doubling in the magnetically ordered state, thus effectively suppressing any altermagnetic band splitting. Concurrently, we identify unusually strong Dzyaloshinskii-Moriya interactions, $|\mathbf D|/J_1\simeq 0.3$, that produce spin canting and, together with order-by-disorder effect, pin the Néel vector to the crystallographic $b$-axis. Additionally, DM anisotropy promotes magnon band splitting, but these bands remain non-chiral. Our results highlight the importance of relativistic effects even in $3d$ magnets with altermagnetic symmetries.

cond-mat.str-el↗

Searching for Charged Lepton Flavor Violation

Searches for Charged Lepton Flavor Violation (CLFV) are powerful tools to probe physics beyond the Standard Model (BSM). This chapter provides a comprehensive survey of CLFV theory, focusing primarily on low-energy muon and tau transitions while outlining complementary searches from high-energy colliders. Grounded in the Effective Field Theory (EFT) framework, we systematically review key observables including radiative and purely leptonic decays of muon and tau as well as $μ\to e$ conversion in nuclei and $τ$ hadronic decay channels. CLFV involving light new physics particles is also discussed. This review captures the modern synergy between EFTs, novel light-mediator scenarios, and hadronic and nuclear physics in the ongoing hunt for BSM signals.

hep-ph↗

Partition regular linear equations over Sidon sets

In this article, we show that the Sidon subsets of $[N]^d$ are Fourier uniform, and we prove a dense model lemma for Sidon sets. Using these results and a higher-dimensional version of Rado's theorem, we prove that given any partition regular linear equation in $s \geq 5$ variables with nonzero coefficients, and any finite partition of a sufficiently dense Sidon subset $S$ of $[N]^d$, the number of monochromatic solutions to this equation is $\gg |S|^s N^{-d}$ for all large $N$. As a corollary of the Fourier uniformity result, we show that dense Sidon subsets of $[N]^d$ are equidistributed in certain arithmetic and Bohr structures. Our proofs are motivated by the arguments of Ortega and Prendiville.

math.CO↗

postshock: An R Package for Donor-Adjusted Forecasting After Structural Shocks

We present the postshock R package, which implements and extends a donor-based framework for forecasting when a structural shock is known and the target response of interest is not yet observed. The package estimates shock effects from historical donor episodes, balances donors using specified matching features, and transfers the resulting adjustment to a target-series forecast. It provides integrated workflows for conditional mean forecasting through ARIMA and ARIMAX models and conditional variance forecasting through GARCH-X models. Additional functionality includes structured donor pools, control-shock regressors, automated GARCH-X order selection, processed data objects, and reproducible empirical workflows.

stat.CO↗

ConventionPlay: Capability-Limited Training for Robust Ad-Hoc Collaboration

Ad-hoc collaboration often requires agents to identify and adhere to some shared convention within a cooperative task. Existing work on reinforcement learning (RL) for ad-hoc collaboration focuses on training agents that adapt to the conventions established by their partners. These methods fail to consider the possibility that while some partners might follow only a single fixed convention, others may themselves be capable of adapting to multiple conventions. Here we present ConventionPlay, an RL-based approach that teaches agents to discover their partner's optimal convention by training against a learned population of partners that exhibit different degrees of adaptability across conventions. Some of these partners follow a single, fixed convention, while others are able to adapt to a subset of the possible conventions for the task in question. The existence of partners that support a limited subset of conventions forces agents trained against this population to actively probe their partner's capabilities, and steer their partner towards the most effective joint strategy that they are capable of following. Our experimental results demonstrate that agents trained via ConventionPlay achieve superior performance to existing ad-hoc collaboration methods against test populations of partners that are compatible with multiple conventions.

cs.MA↗

Anytime-valid detection of LLM weight exfiltration

A compromised LLM inference server can leak model weights by encoding payload bits in otherwise plausible token choices. A replay of the same prompt in a trusted server can expose such deviations, but benign numerical nondeterminism also causes token mismatches. Patient attackers can therefore hide within normal variation unless evidence is combined across responses. We introduce a prompt-level e-process that calibrates whole-response mismatch events on trusted benign traffic and accumulates evidence sequentially while, under a calibration-transfer assumption, controlling the probability of any false alarm over an unbounded monitoring horizon. We evaluate it on four models against a seed-blind attack and a stronger seed-aware attack that hides payload bits only in near-ties to remain stealthy, analyzing the channel capacity vs detectability trade-off. Compared with a hard per-token alarm, the e-process combines weak evidence across responses while providing explicit anytime false-alarm control.

cs.CR↗

Einstein-Cartan and Palatini-Cartan Gravity from Graded Poisson Geometry

Using notions from graded Poisson geometry, we develop a framework that allows for concise formulations of field theories such as Einstein-Cartan and Palatini-Cartan gravity. We show that Einstein-Cartan gravity is equivalent to a variant of Palatini-Cartan gravity in which the vielbein fields are constrained to satisfy appropriate symmetry relations, which we refer to as constrained vielbein gravity. To that end, we relate these theories via a diffeomorphism of graded manifolds. We then show how to obtain the full Palatini-Cartan theory by removing the symmetry constraint. Finally, we show how our graded geometric framework allows for a direct incorporation of the minimal BV extension of the graded constrained vielbein and Palatini-Cartan theories.

hep-th↗

Open-Vocabulary Audio-Visual Event Localization via Complex-Valued Fusion

Open-Vocabulary Audio-Visual Event Localization (OV-AVEL) labels each video segment with an event class, including classes that were never seen during training. The dominant pipeline uses a frozen multimodal foundation model (e.g. ImageBind) to embed the visual frame, the audio mel-spectrogram, and each candidate class name into a shared space, then computes two cosine similarities for each segment against each class: visual-text and audio-text. Existing methods then collapse this pair into a single scalar score with a fixed rule (geometric mean, weighted average) before taking the argmax. Instead, we compute complex-valued similarities and learn their fusion using a complex-valued neural network (CVNN). Each modality's standard representation becomes the real part of our pipeline, and a paired companion stream supplies the imaginary part. We use imaginary part of iHSV for visual modality and CycleGAN-translated phase spectrogram for audio modality as these companion streams. This results in two complex similarities, which are then fused. While the vision and audio encoders remain frozen, only the temporal-attention blocks and the fusion CVNN are trained. The four-stream complex architecture sets a new state of the art on both OV-AVEL benchmarks. On the open (unseen-class) split of OV-AVEBench we reach 66.5/59.1/54.1% Acc/Seg-F1/Event-F1 (+1.6/+4.1/+6.6 over the previously reported fine-tuned baseline), with consistent gains for seen classes as well. We also modify AVE dataset for this task and observe that our architecture reaches 60.7/51.9/50.4% Acc/Seg-F1/Event-F1, achieving state-of-the-art OV-AVEL results on it as well. We also propose a two-stream alternative, which also sees great improvements over the baseline.

cs.CV↗

Arithmetic Progressions in Midpoint Colourings

We introduce an asymmetric variant of a classic problem by Roth about colourings of integers and midpoints. The new problem cannot be solved with the standard approach by Erdös-Sárközy-Sós which uses symmetry. We describe a Fourier Analytic approach which gives asymptotically tight bounds. In the $\mathbb{F}_3^n$ setting, the density increment we show is efficient and together with the Freiman-Ruzsa Theorem provides an affine subspace of near optimal codimension within the distinguished set. This approach adopted to the setting of the integers from $1$ to $N$ via Bohr sets and Bogolyubov-Ruzsa Lemma gives a progression of length being a power of $N$ which only depends on the number of colours. We then use Chang's Lemma and Balog-Szemerédi-Gowers Theorem to further improve the dependence on the number of colours.

math.CO↗

SemField: A Simple, Linear, Continuous, yet Robust Semantic Watermark

The rapid proliferation of Large Language Models (LLMs) necessitates reliable watermarking techniques to identify AI-generated text and ensure appropriate attribution. While token-based watermarks are vulnerable to paraphrasing, a central challenge for semantic watermarking is to turn sentence meanings into a stable, well-calibrated document-level signal. To this end, we introduce SemField, a simple, training-free semantic watermark that embeds a continuous, linear signal directly into the sentence embedding space. Using a shared secret key to define a specific Gaussian direction, SemField iteratively evaluates candidate sentences and selects those that maximize the alignment of the cumulative document aggregate with this targeted direction. We also propose SemField-PL, a polarity-locked variant that first determines the optimal orientation from the initial sentence and continuously reinforces it throughout the generation process. The document-level aggregated formulation guarantees exact invariance to sentence reordering and provides theoretical bounds against structural tampering, such as sentence insertion and deletion. Lastly, with extensive evaluations across three models, we demonstrate that both variants outperform 12 recent baselines. Across clean detection and four paraphrasing attacks, both variants achieve a mean True Positive Rate (TPR) of 88.2% to 90.4% at a 1% False Positive Rate (FPR), all while maintaining comparable perplexity and naturalness as human-generated content. We open-source our implementation at https://github.com/declare-lab/SemField

cs.CR↗

VEDJE: Video-Efficient Discriminative Joint Encoder for Scalable Video-Text Retrieval

Finding the right video often requires distinguishing similar scenes in which different events occur. Joint matching improves retrieval, but processing rich video representations for each query is costly. VEDJE compresses features within sampled frames while keeping their representations separate in a reusable cache. Feature-change prediction supplies an auxiliary training signal that improves retrieval from the compressed cache without adding work at query time. On MSR-VTT, MSVD, DiDeMo, and ActivityNet, VEDJE improves R@1 over matched first-stage retrievers in both retrieval directions. On MSR-VTT, it reaches 59.8 text-to-video R@1 with a fine-tuned VideoCLIP-XL first stage. In the VideoPrism configuration, shrinking the per-video cache fourfold to 12 KiB preserves text-to-video recall within 0.2 points. These results show that accurate video search can operate on compact evidence, encoded once and reused as new queries arrive.

cs.CV↗

Dictatorial profiles: a proof of Arrow's theorem

We study social welfare rules with strict individual and social preferences. A preference profile is dictatorial for a rule if the social preference coincides with some agent's preference and no agent with a different preference can change the social preference unilaterally. We establish that (i) a rule is dictatorial if and only if all of its profiles are dictatorial; and (ii) if a rule satisfies unanimity and independence (of irrelevant alternatives), then all its profiles must be dictatorial. These results constitute a proof of Arrow's Impossibility Theorem.

econ.TH↗

Instrumentation and Stabilization of Electric Arcs for Plasma Smelting Reduction

The Hydrogen Plasma Smelting Reduction (HPSR) is a promising technology for low-CO$_2$ steel production. However, unstable arc conditions limit energy efficiency, hydrogen utilization, and electrode wear. In this paper, we discuss the influence of the electrical supply characteristics, which can be a source of additional instability. For active stabilization, a fast current controller using a buck converter with an adaptive state controller is used. For analyzing arc stability, a stereoscopic measurement system for reconstructing the arc is presented. The results show a clear correlation between the arc lengths and the measured arc voltage at constant current. Therefore, arc stability can be determined by the measured voltage. These findings build the basis for dynamic arc control by fast electrical interventions.

eess.SY↗

DADP: Dynamic Activity-Dependent Pruning, A Reverse Hebbian-Inspired Structural Pruning Method

Modern neural networks are heavily over-parameterized. This redundancy incurs substantial compute and memory overhead during training and inference. Existing pruning methods rely on post-hoc magnitude thresholds or static initialization heuristics. Consequently, they often require manual per-layer sparsity targets or expensive retraining cycles. We propose Dynamic Activity-Dependent Pruning (DADP), a biologically inspired structural plasticity mechanism. During training, DADP measures connection importance via the accumulated product of pre-synaptic activations and post-synaptic error gradients. Using a single global threshold instead of fixed layer budgets, DADP dynamically allocates sparsity across network depth while naturally inducing neuron- and channel-level pruning. Across MLP, VGG-16, ResNet-18, BiLSTM-CRF, and MiniBERT architectures, DADP matches or outperforms Magnitude, SNIP and RigL, retaining 73.67% accuracy (dense baseline: 76.06%) at 99% sparsity on ResNet-18. Finally, matrix-based Shannon entropy and effective rank measurements confirm that DADP preserves latent feature diversity at extreme sparsities without representation collapse.

cs.LG↗

GRPODropout: Less is More for Online Reinforcement Learning Rollouts

Reinforcement learning (RL) methods such as GRPO substantially improve large language model reasoning but often suffer from policy entropy collapse: the loss of sampling diversity weakens exploration and limits further improvement. Existing methods address this issue either through algorithm-level interventions, such as reward modification and entropy/KL regularization, or through token-level reweighting. We investigate a complementary perspective: entropy collapse can also be mitigated by changing which generated rollouts contribute to policy updates. Under the same sampling budget, not all rollouts contribute positively to an update, and selectively excluding some can improve learning. To address this, we propose GRPODropout: before the standard update, we use a simple strategy that selectively removes a small number of high-probability positive-advantage rollouts and recenters the retained advantages. To motivate this design, we develop a rollout-level theoretical analysis that guides method design and threshold selection. The method changes only rollout usage, and adds negligible computational overhead. Experiments show higher accuracy than original GRPO and higher actor entropy while using fewer rollout samples for updates, illustrating "less is more." This work provides insight into RL rollout usage: removing some rollouts can improve performance. Code is available at https://github.com/hexuandeng/GRPODropout/.

cs.LG↗