arXiv ScienceSearch

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 199 records · Page 11Linked to original sources

Learning to Remember: Attentive Reinforcement Learning for Edge Serverless Autoscaling

In edge computing, the stochastic and bursty nature of serverless workloads challenges autonomous resource orchestration. Traditional reactive controllers, such as the Kubernetes Horizontal Pod Autoscaler (HPA), suffer from reaction latency, leading to Service Level Objective (SLO) violations during traffic spikes and resource flapping during ramp-downs. While Deep Reinforcement Learning (DRL) offers a pathway toward proactive management, standard agents suffer from \textit{temporal blindness}, an inability to exploit the recent temporal context in non-Markovian edge environments. To bridge this gap, we propose a stability-aware autoscaling framework unifying short-horizon temporal context and control via an Attention-Enhanced Double-Stacked LSTM architecture integrated within a Proximal Policy Optimization (PPO) agent. Unlike shallow recurrent models, our approach employs a learned attention mechanism that weights recent historical states non-uniformly, suppressing high-frequency jitter while preserving the trend that precedes demand shifts. We validate the framework on two independent Kubernetes clusters using real-world Azure Functions traces. Against the single-layer LSTM ablation and the static HPA baseline, our approach reduces P90 latency by $\approx$67\%, and holds average latency within the 50ms hard SLO for 98.8\% of the run against 49.6\% and 43.5\% respectively. Against Kubernetes Event-Driven Autoscaling (KEDA), it matches latency performance at 75\% fewer replica-steps and 59\% less churn, with P90 hard-SLO violation bursts of at most 5 consecutive intervals against up to 24 for KEDA. These results indicate that mitigating temporal blindness through deep attentive memory improves the reliability and stability of Kubernetes autoscaling under bursty edge workloads.

cs.DC

LLM-Based Agents for Identifying Bug-Introducing Commits

Śliwerski, Zimmermann, and Zeller (SZZ) just won the 2026 ACM SIGSOFT Impact Award for asking: When do changes induce fixes? Their paper from 2005 served as the foundation for a wide array of approaches aimed at identifying bug-introducing changes (or commits) from fix commits in software repositories. But even after two decades of progress, the best-performing approach from 2025 yields a modest increase of 10 percentage points in F1-score on the most popular Linux kernel dataset. In this paper, we uncover how and why LLM-based agents can substantially advance the state-of-the-art in identifying bug-introducing commits from fix commits. We propose a simple agentic workflow based on searching a set of candidate commits and find that it raises the F1-score from 0.64 to 0.81 on the most popular Linux kernel dataset, a bigger jump than between the original 2005 method (0.54) and the previous SOTA (0.64). We also uncover why agents are so successful: They derive short greppable patterns from the fix commit diff and message and use them to effectively search and find bug-introducing commits in large candidate sets. Finally, we also discuss how these insights might enable further progress in root cause understanding and repair.

cs.SE

Dissolution of carbonate stones caused by CO2 pollutant: Numerical modelling of erosion scenarios

In this paper, we introduce a mathematical model of carbonate-stone erosion driven by the penetration of CO2-derived acidity into the water-filled pore network. The model couples transport of the reactive aqueous species, calcite dissolution and porosity evolution, thereby representing the feedback by which dissolution increases porosity and modifies subsequent transport. Such model is formulated as nonlinear reaction-transport system in porous media governed by Darcy flow. We propose a numerical algorithm based on finite difference approximation that relies on level-set method at the boundaries and we show numerical tests that are in accordance with the literature in terms of the advancement of the erosion front.

math.NA

Softmax gradient policy for variance minimization and risk-averse multi armed bandits

Algorithms for the Multi-Armed Bandit (MAB) problem play a central role in sequential decision-making and have been extensively explored both theoretically and numerically. While most classical approaches aim to identify the arm with the highest expected reward, we focus on a risk-aware setting where the goal is to select the arm with the lowest variance, favoring stability over potentially high but uncertain returns. To model the decision process, we consider a softmax parameterization of the policy; we propose a new algorithm to select the minimal variance (or minimal risk) arm and prove its convergence under natural conditions. The algorithm constructs an unbiased estimate of the objective by using two independent draws from the selected arm's distribution. We provide numerical experiments that illustrate the practical behavior of these algorithms and offer guidance on implementation choices. The setting also covers general risk-aware problems where there is a trade-off between maximizing the average reward and minimizing its variance.

cs.LG

Sona: Personalized Soundscape Mediation to Support People with Sound Sensitivity

People with sound sensitivity (PWSS) often manage distressing sounds with earplugs and noise-canceling headphones that broadly suppress their surroundings, limiting access to useful auditory cues. We present Sona, a mobile system for personalized, real-time soundscape mediation, informed by prior sound sensitivity research and an online survey of 68 PWSS. Sona selectively attenuates multiple overlapping user-chosen sounds at adjustable strength, suggests targets from ambient sound recognition, and lets users add custom targets from short recordings without retraining the model. In an in-situ evaluation with ten PWSS, participants reported that Sona made their soundscapes more manageable. The study also surfaced uneven attenuation across sound type and context, tensions between managing filters and attending to ongoing activities, and difficulty interpreting personalization outcomes. These findings highlight the need to design for the quality of the residual soundscape, balance user control with interaction demands, and support guided, interpretable personalization.

cs.SD

A Halo: The Trigger to a New Era of Nuclear Correlations

In this contribution to the Halo-40 Proceedings, we discuss two topics regarding halo phenomena: The first is the pairing anti-halo effect on the neutron radius of halo nuclei and its restoration due to the coupling to the continuum; the second is the soft dipole excitation of deformed halo nuclei. We demonstrate the importance of Hartree-Fock-Bogoliubov and the relativistic Hartree-Bogoliubov theory in continuum for properly taking into account the halo nature of extended wave functions in calculations of neutron radii, as well as the soft dipole excitations of halo nuclei. It was shown that the anti-halo effect is very sensitive to the continuum coupling induced by Bogoliubov-type quasi-particles, which largely cancels the anti-halo effect on the neutron radius. The soft dipole excitations of deformed halo nuclei Ne-31 and Mg-37 are discussed within the deformed Woods-Saxon model. We point out that the sharp peak just above the threshold in the dipole response is created by the halo effect, and its strength can be used to identify the magnitude of deformation and the halo configuration in the Nilsson level scheme.

nucl-th

Safe learning-based control via function-based uncertainty quantification

Uncertainty quantification is essential when deploying learning-based control methods in safety-critical systems. This is commonly realized by constructing uncertainty tubes that enclose the unknown function of interest, e.g., the reward and constraint functions or the underlying dynamics model, with high probability. However, existing approaches for uncertainty quantification typically rely on restrictive assumptions that encode smoothness properties of the unknown function, such as a known norm in a function space. Moreover, these methods usually struggle with discontinuities. In this paper, we model the unknown function as a random function from which independent and identically distributed realizations can be generated. We then construct uncertainty tubes via the scenario approach that hold with high probability. Our uncertainty tubes rely solely on sampled realizations and can therefore accommodate discontinuities represented by the sampling model. We integrate these uncertainty tubes into a safe Bayesian optimization algorithm with which we safely tune control parameters on a real Furuta pendulum.

eess.SY

On the p-part of the conductor of a generalised character

We show that the $p$-part of the conductor of a generalised character of a finite group is equal to the conductor of its generalised decomposition numbers. We use this to show that $p$-parts of conductors of irreducible characters are preserved under isotypies and perfect isometries that arise in the context of stable equivalences of Morita type with endopermutation source. We apply this to blocks with abelian defect and Frobenius inertial quotient.

math.RT

DQC1-completeness of normalized trace estimation for functions of log-local Hamiltonians

We study the computational complexity of estimating the normalized trace $2^{-n}\mathrm{Tr}[f(A)]$ for a log-local Hamiltonian $A$ acting on $n$ qubits. This problem arises naturally in the DQC1 model, yet its complexity is only understood for a limited class of functions $f(x)$. We show that if $f(x)$ is a continuous function with approximate degree $Ω(\mathrm{poly}(n))$, then estimating $2^{-n}\mathrm{Tr}[f(A)]$ up to constant additive error is DQC1-complete, under a technical condition on the polynomial approximation error of $f(x)$. This condition holds for a broad class of functions, including exponentials, trigonometric functions, logarithms, and inverse-type functions. We further prove that when $A$ is sparse, the classical query complexity of this problem is exponential in the approximate degree. Together, these results identify the approximate degree as the key parameter governing the complexity of normalized trace estimation: it characterizes both the quantum complexity (via efficient DQC1 algorithms) and the classical hardness, yielding an exponential quantum-classical separation. Our proof develops a unified framework that cleanly combines circuit-to-Hamiltonian constructions, periodic Jacobi operators, and tools from polynomial approximation theory, including the Chebyshev equioscillation theorem.

quant-ph

Thermal Hall resistivity and transverse entropy production in a phonon gas

Most theories of the phonon thermal Hall effect ignore phonon-phonon interactions. Here, by recalling the Senftleben-Beenakker effect in molecular gases, we argue that a magnetic field, by influencing collisions between neutral non-chiral [quasi-]particles, can induce a Hall response. Our study of two insulators with distinct crystal structures, layered honeycomb WS$_2$ and ferroelectric perovskite LiNbO$_3$, finds that $κ_{xx}$ and $κ_{xy}$ peak at nearly the same temperature in both materials, as reported in other insulators. We show that the amplitude of transverse thermal \emph{resistivity} in clean and simple insulators is of the order of $|W_\perp/B|\simeq \frac{e}{k_B u}$, where $e$ and $k_B$ are fundamental constants and $u$ is the binding energy density of the crystal. In complex and dirty insulators, $|W_\perp/B|$ is much larger and has a significant temperature dependence. Nevertheless, the peak thermal Hall \textit{angle} in all insulators remains roughly the same.

cond-mat.mtrl-sci

On the triviality of inhomogeneous deformations of $\mathfrak{osp}(1|2n)$

We specify a symmetrized mixed-oscillator deformation family of $B(0,n)=\operatorname{osp}(1|2n)$, with even mixed coefficients and one odd square-zero parameter. For every $n\geq1$, we derive its bracket from a faithful oscillator realization and exhibit an odd cochain whose coboundary is the recovered deformation coefficient. The resulting even change of generators is an exact isomorphism over the exterior parameter algebra. For $n=1$, the cochain agrees with the normalization of Bakalov-Sullivan. We give the source relations and the even-central specialization explicitly, together with a Lean 4 formalization.

math.RT

Value Mirror Descent for Reinforcement Learning

Value iteration-type methods have been extensively studied for computing a nearly optimal value function in reinforcement learning (RL). Under a generative sampling model, these methods can achieve sharper sample complexity than policy optimization approaches, particularly in their dependence on the discount factor. In practice, they are often employed for offline training. In this paper, we consider discounted Markov decision processes with state space S, action space A, discount factor $γ\in(0,1)$ and costs in $[0,1]$. We introduce a novel value optimization method, termed value mirror descent (VMD), which integrates mirror descent from convex optimization into the classical value iteration framework. In the deterministic setting with known transition kernels, we show that VMD converges linearly. For the stochastic setting with a generative model, we develop a stochastic variant, SVMD, which incorporates variance reduction commonly used in stochastic value iteration-type methods. For RL problems with general convex regularizers, SVMD attains a near-optimal sample complexity of $\tilde{O}(|S||A|(1-γ)^{-3}ε^{-2})$. Moreover, we establish that the Bregman divergence between the generated and optimal policies remains bounded throughout the iterations, even under the presence of model misspecification. This property is absent in existing stochastic value iteration-type methods but is important for enabling effective online (continual) learning following offline training. Under a strongly convex regularizer, SVMD achieves sample complexity of $\tilde{O}(|S||A|(1-γ)^{-5}ε^{-1})$, improving performance in the high-accuracy regime. Furthermore, we prove convergence of the generated policy to the optimal policy. Overall, the proposed method, its analysis, and the resulting guarantees, constitute new contributions to the RL and optimization literature.

math.OC

Enhanced ShockBurst for Ultra Low-Power On-Demand Sensing

On-demand sensing requires battery-powered Internet-of-Things (IoT) and implantable medical devices to remain in deep sleep and activate wireless communication only when data transmission is required. In such systems, battery lifetime depends strongly on radio active time. This work investigates how communication architecture and physical layer (PHY) configuration influence radio active time by comparing connection-oriented Bluetooth Low Energy (BLE) with connectionless Enhanced ShockBurst (ESB) on identical BLE-compatible hardware. Under identical 2 Mbps PHY configurations, ESB reduces wake-up latency and energy consumption to approximately one-twentieth of BLE by eliminating connection establishment and maintenance overhead. Increasing the ESB PHY rate from 2 to 4 Mbps further shortens packet airtime by approximately 52% and reduces transmission energy by approximately 43%. Finally, a first-in, first-out (FIFO)-triggered implantable loop recorder prototype demonstrates that jointly optimizing communication architecture, PHY configuration, and buffered transmission enables sleep-wake operation and reduces total system power consumption by approximately 60% compared with conventional BLE operation. These results identify minimizing radio active time as a key design principle for ultra-low-power on-demand sensing and provide practical guidance for battery-powered sensing systems.

eess.SY

Linearly Solvable Continuous-Time General-Sum Stochastic Differential Games

This paper introduces a class of continuous-time, finite-player stochastic general-sum differential games that admit solutions through an exact linear PDE system. We formulate a distribution planning game utilizing the cross-log-likelihood ratio to naturally model multi-agent spatial conflicts, such as congestion avoidance. By applying a generalized multivariate Cole-Hopf transformation, we decouple the associated non-linear Hamilton-Jacobi-Bellman (HJB) equations into a system of linear partial differential equations. This reduction enables the efficient, grid-free computation of feedback Nash equilibrium strategies via the Feynman-Kac path integral method, effectively overcoming the curse of dimensionality.

math.OC

Scheduling Coflows in Multi-Core OCS Networks with Performance Guarantee

The coflow abstraction captures application-level communication patterns and enables coordinated scheduling of parallel flows to reduce job completion times in distributed systems. Modern data center networks (DCNs) are employing multiple independent optical circuit switching (OCS) cores operating concurrently to meet the massive bandwidth demands of application jobs. However, existing coflow scheduling research primarily focuses on the single-core setting, while studies of multi-core fabrics have largely considered electrical packet switching (EPS) networks. To address this gap, this paper studies the coflow scheduling problem in multi-core OCS networks under the not-all-stop reconfiguration model, in which the reconfiguration of one circuit does not interrupt other circuits. The challenges stem from two aspects: (i) cross-core coupling induced by traffic assignment across heterogeneous cores; and (ii) per-core OCS scheduling constraints, namely \textit{port exclusivity} and \textit{reconfiguration delay}. We propose an approximation algorithm that jointly integrates cross-core flow assignment and per-core circuit scheduling to minimize the total weighted coflow completion time (CCT) and establish a provable worst-case performance guarantee. Furthermore, our algorithm framework can be applied to the multi-core EPS scenario with a corresponding approximation guarantee for packet-switched fabrics. Trace-driven simulations using real Facebook workloads demonstrate that our algorithm can reduce the total weighted CCT and tail CCT.

cs.DC

TiAb Review Plugin: A Browser-Based Tool for AI-Assisted Study Selection in Systematic Reviews

Server-based screening tools impose subscription costs, while open-source alternatives require coding skills, and full-text screening has remained outside the scope of no-code open-source tools. We developed TiAb Review Plugin, an open-source Chrome browser extension that provides no-code, serverless artificial intelligence (AI)-assisted study selection covering both title and abstract (T&A) screening and full-text screening. It uses Google Sheets as a shared database and Google Drive as a PDF store, and users supply their own large language model (LLM) API key. For T&A screening, it offers manual review, LLM batch screening, and machine learning (ML) active learning. For full-text screening, it retrieves open-access PDFs from PubMed Central, Europe PMC, Unpaywall, OpenAlex, and publisher pages, supports blinded dual review with structured exclusion reasons and adjudication, optionally obtains an LLM judgment with page-anchored evidence, and computes PRISMA 2020 flow counts. We re-implemented the default ASReview algorithm (TF-IDF with Naive Bayes) in TypeScript and compared it with the Python original using 10-fold cross-validation on six datasets. For LLM T&A screening, we compared 16 parameter configurations on a benchmark dataset, validated the best (Gemini 3.0 Flash, low thinking budget, TopP 0.95) on five public datasets (1,038 to 5,628 records; 0.5% to 2.0% prevalence), and benchmarked nine further models from four developers. The TypeScript classifier produced top-100 rankings identical to ASReview on all six datasets. LLM T&A screening achieved recall of 94% to 100% with precision of 2% to 15%, and work saved over sampling at 95% recall (WSS@95) of 46.3% to 89.3%. No additional model exceeded the 96.1% recall of the reference configuration; the most recent models traded recall for precision. The classification accuracy of the full-text stage has not yet been evaluated.

cs.DL

Joint Interference Detection and Identification via Adversarial Multi-task Learning

Precise interference detection and identification are crucial for enhancing the survivability of communication systems in non-cooperative wireless environments. While deep learning (DL) has advanced this field, existing single-task learning (STL) approaches neglect inherent task correlations. Furthermore, emerging multi-task learning (MTL) methods often lack a theoretical foundation for quantifying and modeling task relationships. To bridge this gap, we establish a theoretically grounded MTL framework for joint interference detection, modulation identification, and interference identification. First, we derive an upper bound for the weighted expected loss in MTL frameworks. This bound explicitly connects MTL performance to task similarity, quantified by the Wasserstein distance and learnable task relation coefficients. Guided by this theory, we present the adversarial multi-task interference detection and identification network (AMTIDIN), which integrates adversarial training to minimize distributional discrepancies across tasks and uses adaptive coefficients to model task correlations dynamically. Crucially, we conducted a quantitative analysis of task similarity to reveal intrinsic task relationships, specifically that modulation identification and interference identification share a substantial feature overlap distinct from interference detection. Experiments demonstrate that AMTIDIN outperforms its independently trained single-task counterparts and MTL baselines under the evaluated conditions of limited training data, short signal lengths, and low signal-to-noise ratios (SNRs)

cs.LG

GLIMPSED: Direct evidence for a fast active galactic nucleus-driven outflow from a z=6.64 little red dot host galaxy

We report the discovery of GLIMPSED-329380, a z=6.64 galaxy behind Abell S1063, which shows signs of an extreme ionised outflow driven by an active galactic nucleus (AGN). The deep JWST/NIRSpec medium grating observations show spatially resolved structures of a host galaxy containing the very fast outflow and an AGN, which we analyse separately. The outflow, mainly traced by broad [O III]λ5008 and Hα emissions in the host, reaches a full-width-at-half-maximum velocity of ~5350km/s, velocities only observed in AGN-dominated systems. From the Balmer decrement, we observe that while the narrow emission lines show no dust attenuation, the outflowing gas is dusty. We use emission lines diagnostics to infer gas abundances within the host galaxy. The oxygen abundance is 12+log(O/H)~7.96 (~18% solar) and the host is slightly nitrogen-enriched with log(N/O) ~ -0.68. Unless the extreme outflow turns out to be very dusty, the mass loading factor (>5-10%), and the kinetic energy (>1e43erg/s) of the extreme outflow does not hint at a significant impact on the galaxy. The AGN component shows many similarities with little red dots (LRDs): a characteristic 'V shape', exponential profile in hydrogen lines, numerous detection of forbidden [Fe II] lines, a Balmer break, and a broad absorption feature at ~4550Å. This detection of a fast outflow in an LRD, rare in surveys dominated by low-resolution (e.g. PRISM) spectra, provides direct evidence of AGN activity in these systems.

astro-ph.GA