arXiv Science⌕ Search

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 793 records · Page 44Linked to original sources

Nonequilibrium Dirac condensate in bosonic Kitaev chain

We explore a nonequilibrium Bose-Einstein condensate (BEC) in a bosonic Kitaev chain. Within a mean-field framework, we show that bosonic superfluid order develops once the pairing strength exceeds a critical threshold. The condensate order parameter adopts an alternating phase between even and odd sites, giving rise to a spontaneous sublattice ordering. By mapping the emergent sublattice degrees of freedom onto a pseudospin, the elementary Bogoliubov excitation is described by a non-Hermitian Dirac Hamiltonian, displaying two prominent features: a re-entrant bosonic Andreev bound state localized around boundaries, and a macroscopic degeneracy of exceptional Majorana bosons at the critical point.

cond-mat.mes-hall↗

A Rapid Pipeline for Training and Deploying ML Models on WeBe Band

Developing optimized machine-learning algorithms for edge devices with limited computational and memory resources is challenging, time-consuming, and highly dependent on device-specific constraints. In this work, we streamline an edge ML workflow to enable rapid development, optimization, and deployment of machine-learning (ML) models directly on the WeBe Band, a wrist-worn wearable device designed for multimodal physiological data monitoring. The proposed system automatically generates hardware-efficient ML models that can be easily integrated into the WeBe core firmware, supporting AutoML, hardware-aware quantization, and performance profiling to build models that meet desired latency targets while remaining compatible with device memory and power limitations. The proposed framework tightly integrates the open-source Piccolo AI ecosystem with an automated pipeline that generates deployable firmware artifacts, performs hardware-aware model compilation, and supports over-the-air (OTA) deployment. The system supports multiple lightweight model classes, including classical machine-learning algorithms and neural networks, and provides built-in on-device profiling tools to evaluate inference latency and memory footprint under realistic execution conditions. Experimental results demonstrate clear trade-offs between model complexity and deployability on a microcontroller, showing that classical models offer strong real-time performance while lightweight neural networks require careful resource management. Rather than proposing new learning architectures, the current work mainly focuses on system-level automation, deployability, and enabling researchers and developers to rapidly iterate on models and evaluate them directly on target hardware. Although demonstrated on the WeBe Band platform, the workflow is designed to be extensible to other ML-powered edge devices.

cs.AI↗

Orthonormal Strichartz estimates for the Schrödinger equations with Hamiltonian on Wiener amalgam spaces

The main objective of this paper is to investigate orthonormal Strichartz estimates for Schrödinger equation on the Wiener amalgam space $\mathcal{W}(\mathcal{F} L^p, L^q)$. More precisely, we first examine improvements in the time-integrability exponent of the existing Strichartz estimates established by Cordero and Nicola in \cite{NFC} for Schrödinger equations associated with $\mathcal{H}^{+}=-\frac{1}{4π}Δ+|x|^2$. We then extend these improved Strichartz estimates from a single initial datum to systems of orthonormal families, providing, to the best of our knowledge, the first results in this direction in the setting of Wiener amalgam spaces. We further extend the classical Strichartz estimates in Wiener amalgam spaces from a single initial datum to systems of orthonormal families of initial data for the Schrödinger equation associated with the operator $\mathcal{H}^{-}=-\frac{1}{4π}Δ-|x|^2$ and Hamiltonian operator of the form $\mathcal{H}_{\mathcal{A}} = -\frac{1}{4π} B \nabla \cdot \nabla$, where $\mathcal{A} = \begin{pmatrix} 0 & B \\ 0 & 0 \end{pmatrix} \in \mathrm{Sp}(d,\mathbb{R})$ with $B = B^*$ and $\det B \neq0$. A key ingredient of our approach is Stein's complex interpolation theory for Wiener amalgam spaces, combined with a duality argument inspired by the work of Frank and Sabin. As an application of these orthonormal estimates associated with $\mathcal{H}^{+}, \mathcal{H}^{-}$, and $H_{\mathcal{A}},$ we establish local and small-data global well-posedness for the Hartree equation with infinitely many particles, for non-trace-class initial data.

math.AP↗

Unrolling Iterative Lanczos Algorithm for Ideal Low-pass Graph Filter Approximation

Low-pass (LP) filtering is a fundamental operation in graph signal processing (GSP). Among finite-order nodal-domain methods, Lanczos-based filtering provides more accurate approximations of ideal LP filters than Chebyshev polynomial methods. We show that the approximation of Lanczos filtering can be further improved through algorithm unrolling and data-driven parameter learning. The key insight is that, because ideal LP filtering is a projection operation into the low-frequency eigen-subspace $\cS_K$, instead of approximating individual eigen-pairs of a graph Laplacian $Ł$ as done in classical Lanczos, an unrolled Lanczos network can directly approximate $\cS_K$. Specifically, we first establish a theorem identifying properties of the Lanczos tridiagonal matrix $\T_m$ that promote accurate approximation of the low-frequency eigen-subspace $\cS_K$. Guided by this theory, we relax the orthogonality constraint on Lanczos vectors, resulting in Ritz vectors that better span $\cS_K$. To ensure numerical stability, we constrain $\T_m$ to be similar to a symmetric matrix, thereby guaranteeing real-valued eigenvalues. Experimental results show that our unrolled Lanczos network achieves superior ideal LP filter approximation compared to classical Lanczos and Chebyshev methods.

eess.SP↗

Winding Number Statistics of a Parametric Chiral Symplectic Random Matrix Ensemble

The winding number is a simple topological invariant. In the case of chiral symmetry it characterises gapped phases of Fermions. We study statistical properties of this topological index or invariant in a chiral symplectic setting using Random Matrix Theory. We consider ensembles of Hamilton matrices in the symmetry class CII (quaternionic matrices) according to the classification in the tenfold way. We set up a parametric random matrix model and derive expressions for parametric correlations of the winding number density as well as for the discrete winding number distribution. We found a super-universality in the limit of large matrix dimensions for the one- and two-point correlators for the bulk of parameters, meaning that the results agree up to rescaling with those of the class AIII (complex matrices). In this context we employ a new method of unfolding and discover the Gaussian behaviour of the winding number distribution.

math-ph↗

Light-induced rectified orbital magnetization in electron-hole bilayers

Circularly polarized light can rectify orbital motion into a static magnetization through the inverse Faraday effect (IFE), but in electron-hole bilayers the electron and hole contributions cancel exactly when their properties are equivalent. We show that electron-hole bilayers in transition-metal dichalcogenide platforms avoid this cancellation through effective-mass asymmetry alone, and that the surviving orbital IFE is sensitive to interlayer coupling. In the weak-coupling regime, a two-component Drude description yields an induced magnetization of order one Bohr magneton per carrier for representative terahertz driving. In the strong-coupling regime, where the carriers bind into interlayer excitons, we treat the relative motion using a hydrogenic model with a Rytova-Keldysh interaction and obtain a reduced but finite response. We give closed-form expressions that allow estimates across a broad parameter range and identify where the response is largest.

cond-mat.mes-hall↗

Can Classical Semantic-Extractive Summarization Be Evaluated in Hindi? A Replication Study

We replicate the distributional-semantics extractive summarisation method of Mohd, Jan and Shah (2020) and adapt it to Hindi, substituting a Devanagari-appropriate component at every language-specific step. The system is evaluated on two independent corpora --- the Hindi portion of XL-Sum and FIRE ILSUM 2.0 Hindi --- under a Devanagari-aware ROUGE implementation validated against the XL-Sum authors' own multilingual scorer, with all comparisons drawn as 1000-resample paired bootstraps. In its published equal-weight configuration the replicated system is significantly worse than a three-sentence lead baseline on both corpora, trailing Lead-3 by 0.042 ROUGE-1 Fon XL-Sum and by 0.265 on ILSUM. A feature ablation shows that sentenceposition is the only feature that contributes: position alone reproduces the lead baseline exactly, removing position gives the weakest configuration,and a validation-tuned weighting can at best equal Lead-3 and never exceed it. TextRank fails identically, making this a class-level rather than an implementation-level result. A selection analysis shows the remaining features steer extraction towards long, entity-dense body sentences while the references reuse the article lead.Current Hindi benchmarks therefore cannot reward non-lead content selection, motivating purpose-built evaluation resources.

cs.CL↗

From Passive Execution to Active Exploration: Agentic Embodied Manipulation in Realistic Environments

Recent advances in agentic systems have substantially enhanced the long-horizon capability of embodied manipulation. However, many existing frameworks still follow a passive execution paradigm, which limits their applicability to real-world scenarios involving textual semantic cues, distractors, and initially invisible targets. To bridge this gap, we propose an agent-based active exploration framework that enables robots to dynamically interact with the environment rather than merely execute predefined instructions. Specifically, our framework consists of three collaborative modules: a planning module for high-level task reasoning, a perception module for visual scene understanding, and an execution module for low-level manipulation. This design allows the robot to actively acquire task-relevant information, adapt its behavior based on environmental feedback, and complete manipulation tasks under partial observability. Furthermore, we introduce a fine-grained perception-execution interleaving strategy, which tightly couples visual feedback with skill execution to improve exploration robustness. We evaluate our method on a realistic Find-and-Place task, demonstrating its effectiveness in challenging environments where target objects must be actively discovered before manipulation.

cs.RO↗

DAWN: Noise-Robust Quadruped Parkour via Depth-Denoising World Models

Vision-based legged locomotion methods assume clean depth at training time and rely on hand-tuned post-processing filters at deployment. However, filter parameters are rarely disclosed, hindering reproducibility, and performance degrades substantially when depth noise is left unaddressed. Building noise robustness directly into the learning pipeline would eliminate this dependency. While such robustness has been explored for proprioceptive inputs, analogous approaches for depth perception remain largely absent in legged locomotion. We propose DAWN (Denoising and Alignment in World models for Noise-robustness), a noise-robust perception framework for legged locomotion, which builds noise robustness directly into a world model via two modifications: (1) feeding noisy depth to the encoder while keeping clean depth as the reconstruction target, forcing the model to implicitly denoise its input; and (2) applying contrastive learning to align the latent states of noisy and clean depth. Importantly, DAWN is not tied to a specific noise model, requiring no manual tuning to the noise distribution at deployment. Furthermore, it incurs no additional inference cost over existing world model-based methods. Without any manual filter calibration -- relying solely on the learned noise-robust representation -- DAWN achieves zero-shot quadruped parkour on a Unitree Go1: traversing stairs up to 18 cm, clearing gaps up to 70 cm, and mounting steps up to 45 cm from raw depth observations. Ablation studies show that denoising and contrastive alignment contribute at complementary levels -- reconstruction and representation, respectively -- and yield additive gains when combined. Videos and code are available at: https://dawn-parkour.github.io/

cs.RO↗

A Support-Enhanced Granular-Jamming Gripper for RL-based Grasping with Continuum Manipulators

Continuum manipulators provide dexterous motion in confined spaces, but structural compliance, hysteresis, and load-dependent deformation leave residual position and orientation errors that can undermine reliable contact with rigid grippers. To address this limitation, this paper presents a lightweight support-enhanced granular-jamming gripper tailored to a continuum manipulator. The gripper maintains compliance before jamming while establishing a direct load path to the continuum manipulator tip after jamming. To improve its grasping performance, we systematically designed membrane materials, particles, filling ratios, and the internal support structure, and further identify geometry-dependent grasp boundaries with respect to contact offset and object shape. Building on these results, we construct a physical manipulation system integrating the continuum manipulator, granular-jamming gripper, visual feedback, tendon actuation, and pneumatic control. We then train a reinforcement-learning-based reaching controller in a randomized simulation and deploy it on the physical system, demonstrating how positioning control and contact level mechanical adaptation can complement each other in a modular grasp-and-release task. By introducing an adaptive structure that relaxes the need for highly accurate modeling and positioning control, this work explores a design paradigm that integrates physical and embodied intelligence.

cs.RO↗

Where Does Exactly-Once Live? Model, Harness, and Tool-Contract Effects on Duplicate Side Effects in LLM Agents

When a tool-using agent's write times out or returns a server error, the action may already have taken effect. Retrying blindly duplicates it -- a second charge, a second announcement, a second deployment -- while giving up skips required work. We ask where exactly-once behaviour should be enforced: in the model, in the agent harness, or in the tool contract. We introduce LIMBO, a deterministic sandbox of six services with realistic contracts (optional idempotency keys, eventually consistent and missing read paths) and twelve fault modes injected at the service boundary, including late commits, redelivery and partial batches; every episode is graded against a ledger of committed effects. Across 25,930 episodes spanning nine recent models, three production agent harnesses, two contract variants and fifteen recovery conditions, the answer depends on the fault. When an immediate read-back can reveal what happened, the model decides: frontier models instructed to act exactly once almost never duplicate a write whose acknowledgement was lost (0.5%), weaker models often do, and the model explains 53% of the explained variance. When it cannot -- the request is still in flight, or the transport delivered it twice -- the same frontier models duplicate in 56% and 74% of episodes, and the contract explains 81%. We prove that no verification-only policy is exactly-once under late commits without a bound on in-flight time. Waiting works when such a bound is short and known, but with heavy-tailed in-flight delays even an hour of waiting per episode falls short of offering an idempotency key on every write, which lowers the duplicate rate from 28% to 4% because agents use keys when they exist. The harness barely matters, a guard that attaches keys transfers across harnesses unchanged, and agents reported success in 90% of the episodes in which they had duplicated an effect.

cs.LG↗

Downside-Controlled Online Forecast Combination under Delayed and Revised Outcomes

Post-hoc correction adjusts a forecaster that cannot be retrained, such as a foundation model, but a correction fitted where errors are stable can hurt where they shift. We aim for downside control: not much worse than the starting forecast. We combine the frozen forecaster, a static corrector and an online corrector on the simplex, using only losses that mature after the horizon. Across seven benchmarks and four base models, two of them foundation models, the worst deterioration over 28 pairs at the main horizon is 0.15% and gains reach 11.5%. On day-ahead load for seven European bidding zones it lowers mean MSE in all seven zones, while single correctors raise mean MSE by up to 102% where the published forecast is most accurate. Three empirical conditions on expert speed, stream length and outcome alignment, each fixed by a documented failure, delimit its scope. Learning from the provisional outcome improves four zones on the settled one; learning on the settled outcome restores all seven.

cs.LG↗

Novel higher-dimensional regular black holes in general relativity coupled to nonlinear electrodynamics

We present a formalism for constructing higher-dimensional regular black holes in general relativity coupled to nonlinear electrodynamics, and propose a new class of electrically charged solutions that unifies the Bardeen and Hayward models in a single analytic form valid for any $D\geq4$. The Bardeen and Hayward solutions are shown to have a correct Maxwellian weak-field limit in five and six dimensions, respectively. The matter sector is written as a Hamiltonian density $\mathcal{H}(P)$ containing only fixed coupling constants, so that the mass and the charge enter as integration constants. The resulting theory has a two-parameter family of solutions, of which a one-parameter subfamily is regular, with the mass tied to the charge. We locate the radius at which the Lagrangian $Ł(F)$ splits into two branches and show that it is always hidden inside the horizon for $D=4$ but can lie outside it near extremality for $D\geq5$. The null and weak energy conditions hold everywhere, while the strong energy condition is violated near the core. Building on our recent result that the first law holds in its standard form once the theory is fixed, we derive the horizon potential in closed form, the extended first law with the couplings as thermodynamic variables, and a generalized Smarr formula.

gr-qc↗

Evaluation-efficient quantum architecture search with ZX-calculus-based topological reuse

Variational quantum algorithm is a leading approach for quantum chemistry and many-body physics on noisy intermediate-scale quantum devices. The performance depends strongly on the structure of the parameterized quantum circuits. Quantum architecture search (QAS) can automate ansatz design. However, it requires repeated training and evaluation of many candidate circuits, leading to high evaluation cost. In this work, we propose a noise-aware quantum architecture search framework based on ZX-calculus topological reuse (ZX-QAS). The framework encodes the search space with a ternary Gray-code mapping and integrates a noise-aware quantum neural network with a ZX-calculus topological reuse mechanism. The effectiveness of the framework is validated through ground-state energy estimation tasks and one-dimensional transverse-field Ising model tasks under noisy conditions. The results show that the ZX-QAS exhibits stable convergence and remarkably reduces the cost of expensive evaluations under noisy conditions.

quant-ph↗

TraceGuard: Adaptive Multimodal Poison Filtering through Cross-Feature Rank Agreement

Multimodal training relies on image-text corpora collected from external sources, creating opportunities for attackers to poison the data. Stealthy attacks can preserve plausible image-text pairs while concealing the differences used by detectors, so apparently clean data can still redirect the trained model. We therefore ask which properties a poison set must preserve for the attack to remain effective. A small poison set must still exert enough collective influence during training to induce the attacker's target behavior. We analyze this influence in terms of how often an attack pattern occurs and how strongly the examples carrying it jointly affect the model. This analysis motivates six corpus-level features that examine cross-modal neighborhoods, recurring text, and changes after text-span erasure without training the victim model. We introduce TraceGuard, an adaptive rank-based filtering method that uses agreement among complementary feature rankings to identify suspicious examples. It refines the selected set through shared patterns and adapts the removal threshold to each corpus without knowing the attack or poison rate. Across 19 attack configurations spanning image-text learning, generative vision-language model fine-tuning, and encoder-transfer tests, TraceGuard removes an average of 98.4% of poisoned examples and 5.4% of clean examples. After training on the filtered corpora, the residual attack metric is at most 1% in 13 configurations. Matched-removal controls and ablations support the contributions of sample selection and adaptive removal. Stress tests also identify detection failures under adaptive attacks and unnecessary removal on poison-free corpora.

cs.CR↗

Variable-Horizon Model Predictive Control for Switched Systems

This paper investigates model predictive control (MPC) for switched systems subject to control and state constraints. A variable-horizon switched MPC approach is proposed. By steering the system state into a well-designed switching feasible set, the proposed method structurally decouples the dwell-time conditions from the MPC constraints, thereby relaxing the dwell-time requirements to match those of the unconstrained switched systems. Furthermore, algorithms are developed to construct this switching feasible set and characterize the domain of attraction, ensuring both persistent feasibility and closed-loop asymptotic stability. To further decouple the prediction horizon length from strict dwell-time bounds, advanced short-horizon switched MPC schemes are designed, which expand the overall domain of attraction. Simulations illustrate the efficacy of the proposed methods.

eess.SY↗

Language Specificity vs. Domain Diversity: Benchmarking Transformers for Bangla Medical NER

Medical Named Entity Recognition (NER) for low-resource languages remains a challenging task due to high linguistic variability and a scarcity of domain-specific annotated corpora. This work presents a comprehensive empirical benchmark evaluating three fine-tuned transformer encoders-BanglaBERT, multilingual BERT (mBERT), and XLM-RoBERTa-against GPT-4o mini under zero-shot and few-shot prompting configurations for Bangla medical NER. In contrast to prior studies that evaluated large language models on limited subsets of only 50 samples, we conduct a large-scale evaluation across the full test set of 3,179 samples, providing statistically robust and reproducible baselines. Our fine-tuned XLM-RoBERTa model achieves an F1- score of 0.5959, establishing a new state-of-the-art and surpassing the previously reported best result of 0.5848. Crucially, we demonstrate that the language-specific BanglaBERT model consistently underperforms its multilingual counterparts with an F1-score of 0.4937, indicating that pretraining domain diversity can outweigh language specificity in highly specialized clinical settings. Furthermore, we present a detailed per-entity-type analysis for this task, revealing that Medicine and Specialist categories are recognized with high reliability, achieving F1- scores above 0.83, while the Symptom category remains the most challenging with an F1-score of 0.4367 despite being the most frequent training class. Finally, fine-tuned transformer models outperform the optimal prompting configuration by a factor of 3.76, confirming that prompt-only pipelines remain inadequate for structured clinical entity extraction in low-resource language environments.

cs.LG↗

ELF-REG: Scaling Continuous Diffusion Language Models to Reasoning Tasks

Fully continuous diffusion language models (dLMs) denoise continuous representations without intermediate discretization, then decode all response tokens in parallel at the final step. Their performance on challenging reasoning tasks remains less established than that of autoregressive (AR) LLMs and masked dLMs. We scale Embedded Language Flows (ELF) to mathematical reasoning and code generation on GSM8K, MATH-500, HumanEval, and MBPP. We introduce ELF-REG, which improves learning with representation alignment and entanglement (REPA+REG), where a frozen AR teacher supervises intermediate denoiser features and supplies a global representation that is jointly denoised with the response. ELF-REG-L achieves 55.96% pass@1 on GSM8K at 64 network function evaluations (NFE), and 13.39% on MATH-500 and 22.56% on HumanEval at 128 NFE. It outperforms the evaluated comparable-scale dLMs in pass@1 on GSM8K and code, and improves MATH-500 pass@1 from 10.55% for the ELF-L baseline to 13.39% with ELF-REG-L. Without few-step training, the same task-specific checkpoints support strong low-NFE performance through early-stop, which decodes an intermediate clean prediction without completing the denoising trajectory. At 16 NFE, ELF-REG-L reaches 41.21% HumanEval pass@10, outperforming recent continuous dLMs of comparable scale.

cs.CL↗