arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,459 records · Page 81Linked to original sources

Chaining Skills to Hijack LLM Agents

LLM agents use skills to improve performance on specialized tasks. To complete a user request, an agent may invoke several skills in sequence, allowing information produced under one skill to guide the next. Because skills may come from open-source repositories, this handoff can also carry attacker-controlled claims into later decisions. In this paper, we introduce APEX, which constructs and refines adversarial skill chains tailored to a user task and an attacker-selected action. The key insight is that an agent-written record of genuine task progress can carry a false claim of user approval across skills: an upstream skill induces the agent to create the record, and a downstream skill uses it to direct the attacker-selected action. Across four targeted-action families and six models on SkillsBench, the chains induce the selected action in 512 of 690 attempts (74.2%). On GPT-5.4, the full chain succeeds in 84.3% of attempts, compared with 17.4% when the workflow is merged into one skill. We further evaluate a prompting defense that asks the agent to check skill-produced files against the original request. On GPT-5.4, it lowers targeted-action success from 84.3% to 59.1%, while the verifier test-pass rate across 72 benign native-skill tasks falls from 86.7% to 56.3%. These results highlight the need for defenses that prevent attacker-directed actions while preserving legitimate task performance.

cs.CR↗

A conservative micro-continuum-cellular automaton method for multispecies biofilm dynamics in complex flows

We develop a conservative micro-continuum-cellular automaton method for simulating multispecies biofilm dynamics in complex flows. The proposed method couples the Darcy-Brinkman-Stokes equations, reactive transport, suspended bacteria, and biofilm dynamics with a two-stage cellular automaton algorithm for biofilm redistribution and interface evolution. While treating biofilms as evolving porous media, we ensure conservative redistribution of multispecies biomass across partially occupied cells. Donor and recipient cell volumes are explicitly accounted for to conserve biomass and preserve species composition on non-uniform meshes. The proposed method is assessed against diffusion-dominated benchmark cases, including single-species fingering and multispecies stratification, and is further evaluated through a mesh-convergence study for flow and growth over a rectangular bump. The framework is then applied to counter-diffusional biofilms in a membrane-aerated biofilm reactor as a canonical example. The results demonstrate that the framework captures the expected biofilm morphology and stratification in systems involving coupled flow, substrate transport, and biofilm dynamics. The proposed method provides a flexible computational approach for simulating multispecies biofilm dynamics in complex flows and geometries.

physics.flu-dyn↗

Towards Optimal Policy Improvement

Practical Reinforcement Learning (RL) algorithms learn to solve Markov Decision Processes (MDPs) through iterative policy improvement in the presence of approximate evaluation. We study policy improvement from first principles, defining optimal policy improvement as producing the best policy attainable in a single update under specified constraints. We show that optimal improvement restricted to a set of states is equivalent to solving an induced MDP, characterizing planning with an explicit or implicit model as a path towards optimal policy improvement. Because practical methods commonly solve such induced problems through iterative improvement in the form of greedification, we take steps towards optimal greedification under the central practical constraint of approximate evaluation. We formulate greedification under this constraint as probabilistic decision-making under uncertainty and derive a novel operator that is optimal with respect to the resulting objective. Empirically, the operator and its practical gradient-based approximations improve aggregate performance across GumbelAlphaZero, SAC, ReBRAC and Generalized Policy Iteration, in experiments spanning discrete and continuous actions, model-based and model-free, online and offline RL.

cs.LG↗

Strip Decoupling Inequalities for the Paraboloid in $\mathbb{R}^3$

Large-cap or large-ensemble decoupling inequalities were initially studied by Demeter. These are intimately connected to the reverse square-function conjecture and restriction conjecture in Fourier restriction theory. By strips, we mean a partition of the $R^{-1}$-neighborhood of the paraboloid into $1\times R^{-1/2}\times R^{-1}$ curved large-caps. In this article, we prove the sharp $\ell^2(L^{10/3})$ decoupling estimate for strips in $\mathbb{R}^3$. We reduce the strip decoupling inequalities to a planar incidence problem, which is dealt with via two-ends Furstenberg incidence estimates. Our method is inspired by Wang--Wu on the restriction conjecture.

math.CA↗

Single-Pulse Optical Switching combined with Current-Induced Motion of Skyrmionic Spin Textures

Magnetic skyrmions are promising nanoscale information carriers because their position can be manipulated by electrical currents. The additional ability of deterministic control of the spin structure would allow the use of magnetic skyrmions as multi-state bits, expanding the capabilities of skyrmion devices further. Such control could be realized by coupling all-optical helicity independent switching of magnetization (AO-HIS) to magnetic skyrmions. This is realized in a Pt/Ir/CoB/Gd/Pt multilayer that is engineered to stabilize magnetic skyrmions and, at the same time, enable AO-HIS by single ultrafast laser pulse. Laser excitation is used to write skyrmionic textures, whose polarity is selected by a small applied out-of-plane field. A single 30 fs laser pulse is then used to toggle the magnetization of an illuminated region containing skyrmionic textures, reversing their polarities without requiring any magnetic field or current pulses. This establishes deterministic optical switching of the skyrmion spin structure as a new manipulation channel, distinct from previously reported optical nucleation and annihilation. The material is then patterned into a wire, to combine the optical manipulation with current-induced skyrmion motion. These results establish a route toward opto-spintronic skyrmion devices in which light programs the internal state of a skyrmion and electrical currents control its position.

cond-mat.mes-hall↗

Managing Context and Communication in Distributed Agentic UAV Swarms

Unmanned aerial vehicle (UAV) swarms increasingly rely on language-model agents to provide adaptive mission-level reasoning in uncertain environments. Fully distributed control, in which each UAV hosts an independent Small Language Model (SLM), removes reliance on a centralized coordinator but introduces an information-management problem: long-running interaction histories can degrade the reasoning context, while indiscriminate information dissemination increases communication and inference overhead. We address these challenges with a distributed UAV-agent architecture that enables continuous local SLM control through an event-driven reason-act-observe lifecycle. Runtime knowledge is represented as structured atomic notes and organized into core, local, and peer-specific memory. A deterministic interest-aware gossip engine selectively disseminates these notes according to recipient-specific semantic novelty and recency. We evaluate the architecture using ten UAVs in a simulated search-and-rescue mission. Our approach completes all experimental runs, whereas unrestricted flooding messages completes only 70-85\%, and delegating forwarding decisions to the SLM prevents mission completion in every run. Compared with unrestricted flooding, our approach approximately halves inference-token consumption, reduces transmitted data, and achieves lower survivor-count error.

cs.MA↗

Flow separation and turbulence production over trapezoidal protrusions

The influence of rear-face inclination on turbulence around a surface-mounted protrusion depends on the separated flow approaching its trailing edge. This dependence is examined experimentally using stereoscopic particle image velocimetry over twelve trapezoidal protrusions with upper-surface length-to-height ratios $L/h=1,2,3,4$ and rear-face angles of $30^\circ$, $45^\circ$, and $90^\circ$. Measurements were conducted at a height-based Reynolds number of approximately $5\times10^4$, with an incoming boundary-layer thickness comparable to the protrusion height. For $L/h=1$ and 2, a connected reverse-flow region extends over the upper surface into the wake. For $L/h=3$ and 4, upper-surface reattachment separates the two regions. Length governs this change in topology and the variation in turbulence levels at common spatial locations, whereas rear-face inclination has its strongest influence close to the surface. From $L/h=2$ onward, shallower faces have higher near-face turbulent kinetic energy, while vertical faces have lower energy but a larger spanwise fraction. A decomposition of the in-plane deviatoric production separates the effects of Reynolds-stress and mean-strain magnitudes from those of their relative principal-axis orientation. At the inclined faces, this contribution changes from negative to positive between $L/h=1$ and 2, before upper-surface reattachment occurs. Its subsequent increase with length is driven mainly by increasing tensor magnitudes, with the largest contribution at $L/h=4$ and $30^\circ$, where strong production remains close to the rear face. These results show that local turbulence production depends jointly on the flow state approaching the trailing edge and the position of the energetic shear layer relative to the rear surface.

physics.flu-dyn↗

On-shell renormalization of sine-Gordon by the quantum inverse scattering method

In the quantum inverse scattering method, the sine-Gordon model is solved on a lattice, and its continuum and infinite-volume limits become uniform only after the length of the box is rescaled. We show that this rescaling, supplemented by an on-shell condition, defines a renormalization scheme. The whole dependence on the cut-off is absorbed into a factor $Z_L$ multiplying the length of the box, with an exponent fixed by the scaling dimension of the vertex operator; neither the field nor the coupling $β$ is renormalized, no subtraction scale is introduced, and the mass parameter receives a finite renormalization, fixed by requiring that the lightest breather has the mass of the boson in the linearized theory. The soliton and breather masses follow from the eigenvalues of the monodromy operator as functions of parameters that do not run, and reduce to the classical and semiclassical results as $β\rightarrow0$. In particular, the mass ratios are produced directly from the regularized construction, while the relation between the lattice parameters and the physical scale is absorbed into $Z_L$ and never needs to be computed. The scheme is a reparameterization of the theory which makes explicit the interplay between integrability and renormalization. We compare it with Coleman's normal ordering and with conformal perturbation theory, and we argue that it extends to ultralocal integrable theories with a massive spectrum, possibly including asymptotically free models that admit an ultralocal lattice regularization.

hep-th↗

Optimal Momentum Methods for Stochastic Multilevel Compositional Optimization

This paper investigates stochastic multi-level optimization where the objective is a nested composition of several smooth non-convex functions. We assume that only stochastic estimates of the gradient and function values for each level are accessible. Consequently, obtaining an accurate estimate of the overall gradient is challenging due to the nested structure. To address this, we employ a momentum-based estimator with mini-batches to track the function values of each level, which are subsequently used to construct momentum gradient estimators. We establish an optimal sample complexity of $\mathcal{O}(ε^{-4})$ for finding an $ε$-stationary point, avoiding the stronger average smoothness assumption commonly relied upon in prior literature. Furthermore, by employing a normalization technique, we attain the same rate without requiring problem-dependent constants to set hyperparameters. To achieve the optimal rate without mini-batches, we further develop a batch-free method that incorporates a first-order approximation and a clipping technique for function value estimation. Finally, we validate the effectiveness of our proposed methods through experiments on risk-averse portfolio optimization and hierarchical tilted empirical risk minimization.

math.OC↗

LiDARFlow: Real-Time Panel-Based MAV Guidance in Unknown Environments

This paper presents a guidance algorithm for micro aerial vehicles operating in unknown, cluttered environments using only onboard sensing. The method is based on a panel formulation originally derived from aerodynamic potential-flow theory and generates smooth, collision-free guidance vectors from locally perceived obstacles. The approach is extended to unknown environments by constructing and updating the obstacle representation online from onboard LiDAR measurements. The resulting obstacle-avoidance field is integrated with a nominal guiding vector field to produce the final control input. The system is experimentally validated in indoor flight tests under two scenarios: waypoint navigation and directional guidance. In both cases, the vehicle successfully completes its task while avoiding all obstacles in real time using only onboard perception. The results demonstrate that the method is computationally lightweight and suitable for onboard implementation, with pointcloud processing identified as the main practical limitation. These results support the feasibility of lightweight onboard guidance in unknown environments.

cs.RO↗

Adversarial Robustness in Fake Quantum Simulators

This paper investigates the performance scalability and adversarial robustness of Quantum Machine Learning (QML) models deployed on noise-model-based fake simulators. We conduct a dual-phased study, first benchmarking the computational throughput of Qiskit's Aer simulation engine across varying hardware architectures, and second, evaluating the effectiveness of Projected Gradient Descent (PGD) attacks and adversarial retraining strategies under realistic noise conditions. Our results quantify the runtime tradeoffs and scaling behavior for medium-scale simulations (projected up to 8 qubits) and demonstrate that high adversarial-to-benign retraining ratios (50/50) are essential for achieving practical model robustness for 4-qubit classifiers under realistic noise conditions.

quant-ph↗

Digitally enhanced Multi-wavelength Stabilization using a Passive Fiber Frequency Reference

High-stability frequency stabilization between lasers is essential for applications such as precision metrology, frequency dissemination, and quantum communications. We demonstrate the use of code-based multiplexing of a passive fiber-based interferometer for frequency stabilization and transfer between lasers separated by 85 GHz in frequency using a single shared passive reference. Out-of-loop characterization places an upper bound of sub-kHz/$\sqrt{\text{Hz}}$ on the transferred frequency noise at Fourier frequencies above 40 mHz, with a differential fractional stability of $7\times10^{-13}$ at 1 second. We also demonstrate continuous laser tunability and characterize the residual length noise coupling. Finally, we provide a framework for scaling the architecture to higher laser counts, offering a pathway to scalable, frequency agile stabilization of multiple light sources without the overhead of an optical frequency comb.

physics.optics↗

Characterization and Quantification of Immiscible Polymer Blend Compatibilization by Phyllosilicate Clays

Phyllosilicate clays are widely used in polymer nanocomposites owing to their high anisotropy and tunable surface polarity. Their distribution and interface localization in polymer blends can be used to tune the properties of polymer-clay nanocomposites (PCNCs). A coarse-grained (CG) force field for clays can aid molecular simulations in PCNC development. Here, we developed MARTINI-3 parameters, a CG force field with high chemical specificity, for phyllosilicate clays with diverse surface polarities. Initial interactions for the clay functional groups were determined from hydration free energies, obtained by applying the Lifshitz theory to experimental surface tension data. These were fine-tuned using the structural, thermodynamic, and dynamic properties of thermoplastic starch (TPS)-clay composites from all-atom (AA) molecular dynamics (MD) simulations. The radial distribution function and two-body excess entropy of TPS components, not used in CG parameterization, were accurately estimated, establishing the robustness of the parameters. We investigated the effect of clay surface polarity on polymer segmental dynamics and structure-property relationships. The CG parameters were then used to study the effect of dodecyltrimethylammonium (C12TAB)-modified montmorillonite (MMT), an organically-modified MMT, on TPS-polyethylene (PE) blend morphology using large time- and length-scale MD simulations. We observed compatibilization of the TPS-PE interphase by the amphiphilic clay particle, reducing the TPS-PE interfacial tension from 45 mN/m to 13.06 mN/m. We found good agreement between MARTINI-3 estimates for properties of model MMT-based PCNCs and those from AA simulations and experimental data, establishing grounds for the transferability of the parameters to other systems.

cond-mat.soft↗

The hidden advantage of mask resampling: a theory of masked autoencoders

Why can masked prediction learn useful representations that unmasked reconstruction misses? We study this question in a high-dimensional model of a masked autoencoder (MAE) trained on data with shared latent structure and heterogeneous noise. We prove that masked linear reconstruction can recover the latent feature at linear sample complexity in regimes where unmasked linear reconstruction, equivalent to PCA, fails. The analysis also quantifies the statistical advantage of mask resampling, an established ingredient of masked pretraining. By introducing a fixed collection of $K$ masks per sample, we characterize its effect on feature recovery and downstream performance, identifying regimes where greater mask diversity lowers sample complexity. Guided by this prediction, we find that random cropping and flipping in standard image-training pipelines can obscure the advantage of mask resampling by renewing the prediction task even when the patch mask is fixed. Removing these transformations reveals a downstream advantage for dynamic over static masking in CNN autoencoders and vision transformers. A complementary BERT pilot finds benefits from greater mask diversity on downstream language tasks. Our results separate the benefit of the masked prediction objective from that of mask diversity, and show how a tractable theory can guide experiments that uncover advantages hidden by standard training practices.

stat.ML↗

Beyond Pointwise Error: A Multi-Metric Evaluation of Spatial Climate Downscaling

Climate downscaling aims to reconstruct fine scale spatial fields from coarse resolution inputs. Evaluating the quality of these reconstructions is challenging: low pointwise error can come at the cost of fine scale variability, while realistic spatial variability can be achieved with inaccurate local structures. The evaluation metric can therefore change which method appears to perform best. This work presents a multi metric benchmark comparing five spatial downscaling methods on ERA5 temperature, wind, and precipitation fields. Five criteria assess complementary properties: pointwise error, structural similarity, distribution error, spectral error, and gradient error. The results reveal a systematic trade off between spatial fidelity and fine scale variability. Some methods perform best on pointwise and spatially aligned metrics, but lose high frequency content, while others preserve substantially more spectral variability at the cost of less accurately positioned local structures. Consequently, method rankings change across metrics and variables. These results show that there is no single best downscaling method. Multi metric evaluation is therefore essential for assessing which properties of a climate field are preserved.

cs.LG↗

Protocol Integration of Physical Layer Deception into EAP-TEAP Wi-Fi Authentication

Credential-based Extensible Authentication Protocol (EAP) authentication cannot distinguish a legitimate credential holder from an adversary using compromised credentials. Physical Layer Deception (PLD) complements credential-based authentication by exposing a deceptive primary object over a primary transport while a separate recovery object travels with differentiated reliability over a secondary channel. Existing PLD studies remain, to our knowledge, at the physical/link-model level; using PLD's activation/deactivation mechanism as an authentication gate creates an authentication-specific design requirement, since an all-inactive attempt would exercise no recovery path. We present a batched PLD-based re-verification step for Enterprise Wi-Fi's TEAP/RADIUS/IEEE 802.11 authentication chain, implemented end to end across the server, access point, and device in the open-source hostap 2.12 codebase. Each attempt carries three rounds, at least one active, with no dedicated activation flag. Across four campaigns totaling 1593 attempts, the prototype evaluates batched recovery behavior, rejects the implemented naive credential-bearing attacker in all 30 attempts, measures successful-path latency, and evaluates the security-reliability trade-off for one, two, and three active rounds under two modeled recovery regimes. The evaluation exercises the protocol and software-MAC behavior directly and analyzes informed and retry-seeking attackers under the software recovery model.

cs.CR↗

Evaluating Physical Consistency and Plausibility in Generative Scenario Models for Autonomous Driving

Generative AI models are increasingly used for scenario generation in autonomous driving. While they can generate realistic-looking scenarios, they often provide limited transparency into learned representations and consistency with real-world vehicle dynamics. This lack of formal assurance limits their use in safety-critical validation and certification workflows. To address this aspect, we introduce a layered evaluation protocol that complements existing methods by assessing models across five layers. The first four layers inspect internal representations and network layers through kinematic alignment, statistical baseline comparison, latent controllability, and activation analysis. The fifth layer evaluates model outputs against vehicle dynamics constraints such as lateral jerk thresholds. We demonstrate the protocol on a Variational Autoencoder (VAE)-based scenario generator. Although standard output-level metrics and visualizations suggest that the generated scenarios are realistic, our protocol provides deeper insight into the extent to which the model's latent space aligns with kinematic features and whether visually plausible trajectories satisfy vehicle-dynamics constraints. We further apply the protocol to additional generative models, demonstrating its applicability beyond the VAE architecture.

cs.AI↗

Viability problems for SDEs in a general Gelfand triple setup using a variational approach

We establish viability conditions for a broad general class of stochastic differential equations formulated within the variational approach,% \[ dX_{t}=A\left( X_{t}\right) dt+B\left( X_{t}\right) dW_{t}. \] The underlying functional space setup is given by a general Gelfand triple $V\subset H\subset V^{\ast},$ where $H$ is a separable Hilbert space in which the stochastic evolution equation (SEE) is considered. Our approach to deriving viability conditions is inspired by the techniques developed by Aubin and Da Prato for finite-dimensional forward stochastic differential equations. We establish necessary and sufficient conditions for the viability of closed random constraint sets in terms of adapted variational tangent (and contingent) sets. Furthermore, we illustrate the applicability of the proposed framework by discussing several important classes of stochastic differential equations that fit naturally within our variational setting.

math.PR↗