arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 991 records · Page 55Linked to original sources

AEGIS: Audio Endogenous Guarding via Internal Signals Against Large Audio-Language Model Jailbreaks

Large audio-language models (LALMs) expand language models to process and interpret audio, but also expose them to heterogeneous audio jailbreaks. We ask whether successful jailbreaks reflect failures to recognize harmful intent or failures occurring after such recognition. Layer-wise probing reveals the latter: risk-related information remains decodable from intermediate representations, yet the internal risk signal fails to translate into refusal in later-layer processing. We identify this discrepancy as the risk-to-refusal gap. Building on this finding, we propose AEGIS, a detect-then-intervene defense whose mid-layer risk gate selectively activates downstream safety adapters. Across six LALMs and three heterogeneous audio jailbreak benchmarks, AEGIS reduces the average unsafe rate from 17.9% to 0.4%, while causing only a marginal increase in over-refusal on benign inputs. These results establish selective internal intervention as an effective path toward more robust refusal in LALMs. The code is available at https://github.com/azzzzliao/aegis-audio-defense.

cs.SD↗

Convexity of inverse spectral Green kernels and nondegeneracy of Robin centers in convex domains

Let $Ω\subset\mathbb{R}^N$ be a bounded convex domain and let $A=-Δ_D$ be the positive Dirichlet Laplacian. For every real $s>0$ with $N>2s$, we prove that the function \[ (x,y)\longmapsto K_{s,Ω}(x,y)^{-1/(N-2s)}, \] where $K_{s,Ω}$ is the Green kernel of $A^{-s}$, extends continuously by zero to the diagonal and is jointly convex on $Ω\timesΩ$. We also prove that the associated regular part extends real analytically across the diagonal and that the corresponding Robin function is strictly convex and diverges at the boundary. Consequently, for every real $s>0$ satisfying $N>2s$, there is a unique Robin center. This applies in particular to the spectral fractional Dirichlet Laplacian and to all integer-order Navier polyharmonic operators. If $Ω$ is of class $C^{2,\vartheta}$, $0<\vartheta<1$, we further establish second-order rigidity. For $0 0$ throughout $Ω$. For integer orders, we introduce a finite-part doubling principle across real spectral orders. The argument is based on a two-term expansion of truncated square energies together with the spectral identity $A^{-σ}A^{-σ}=A^{-2σ}$. It follows that, for every $p\in\mathbb N$ with $N>2p$, the unique Navier polyharmonic Robin center is nondegenerate; if $N>3p$, the Hessian is positive definite throughout the domain. The boundary regularity required by these second-order results is independent of the polyharmonic order.

math.AP↗

Temporal Regression-Based Model-Free Sensorless Control of Permanent Magnet Synchronous Motor

To address the widespread sensitivity of surface-mounted permanent magnet synchronous motor (SPMSM) sensorless control to motor parameters, this paper proposes a temporal regression-based model-free sensorless control (TFC) method. First, voltage integrals and current increments over consecutive short intervals are stacked to construct a finite window regression, in which the unknown stator inductance appears as a common scalar coefficient. Second, a projection operator constructed from the stacked current increments eliminates the inductance term, and a least-squares formulation is developed to reconstruct the rotor flux vector. Meanwhile, the analysis of the projected regression and current-flux geometry establishes a two-dimensional direction vector whose components share a common amplitude containing the stator resistance and flux linkage. This amplitude cancels during position extraction. By setting the resistance reference to zero, the proposed method estimates the position without specifying the stator resistance, inductance, or flux linkage. Finally, experimental results verify the effectiveness of the proposed TFC method.

eess.SY↗

Sufficiently Reduced Distributional Regression

We propose Sufficiently Reduced Distributional Regression (SRDR), a generative method that combines conditional distribution estimation with nonlinear sufficient dimension reduction (SDR). It builds on a characterization of sufficiency through strictly proper scoring rules: a dimension reduction is sufficient if and only if predicting the response from the reduced covariates incurs no loss in expected score relative to the full covariates. Sufficient dimension reduction thus becomes a risk minimization problem. SRDR jointly trains a dimension reduction map and a generative prediction model by minimizing the energy score, which can be estimated by sampling without density evaluation or adversarial training. The framework extends to multi-environment data and to classification. We prove that the estimated conditional distributions converge in energy distance to the true ones, which implies that the learned representation is asymptotically sufficient. In simulations and applications to CT slice localization, superconductivity, and digit classification, SRDR recovers low-dimensional sufficient structure and matches or outperforms state-of-the-art nonlinear SDR methods in representation quality and predictive performance.

stat.ME↗

PHOSA: Photorealistic 3D Sign Avatar Modeling and Benchmark

In this work, we focus on photorealistic sign avatar modeling, which is crucial for effective communication with the Deaf community and is characterized by complex hand gestures and nuanced facial expressions. To this end, we introduce MVSign, the first multi-view Chinese sign language dataset co-designed with Deaf experts, featuring diverse gestures and rich annotations. For precise SMPL-X annotation, we develop a hybrid fitting pipeline that produces accurate body, hand, and facial parameters and can also be applied to the monocular setting. Building on MVSign, we propose a decoupled sign avatar representation that isolates body, head, and hand components to capture complex articulations, together with a motion-aware sampling strategy to handle motion blur and balance gesture diversity. Extensive experiments demonstrate that our method achieves high-fidelity visual results on MVSign, particularly in detailed hand and facial regions, and generalizes well to in-the-wild monocular sign language videos. Project page: https://naaapi.github.io/PHOSA.

cs.CV↗

Zero-forward Kerker scattering via synthesized complex-frequency excitation

All objects illuminated with light inevitably cast a shadow -- a universal phenomenon encapsulated in the fundamental property that all passive systems with plane-wave illumination feature a non-zero forward scattering amplitude. Since Kerker's landmark 1983 paper, considerable effort has been devoted to overcoming this limitation and achieving the elimination of forward scattering -- an objective now widely known as the zero-forward Kerker scattering. However, this objective is fundamentally restricted to active systems, requiring either physical gain in materials or virtual gain in the excitation source. Despite advances in active materials science and non-Hermitian photonics, the experimental realization of zero-forward Kerker scattering still remains an open challenge. Here, we show, both theoretically and experimentally, how zero-forward Kerker scattering can be effectively synthesized by a weighted superposition of readily accessible real-frequency responses. Our synthetic recipe builds on creating an artificial pole at a complex frequency that prevails over inherent scattering poles. These findings not only unlock a realistic experimental framework for non-Hermitian light-matter interactions, but also hold technological relevance for non-invasive sensing and imaging.

physics.optics↗

A Faster Algorithm for Fewer Vertex-Disjoint Paths Parameterized by Treewidth

The $k$ vertex-disjoint paths problem asks whether, given a graph $G$ and $k$ pairs of vertices $(s_1,t_1)$, \ldots, $(s_k,t_k)$, $G$ has $k$ pairwise vertex-disjoint paths connecting $s_i$ and $t_i$ for all $1\leq i\leq k$. If $G$ is undirected, then this problem is NP-complete, but there exist FPT algorithms parameterized by $k$.Since these algorithms involve an extremely large function on $k$, algorithms for restricted graphs have also been investigated. In particular, a $2^{2tw\log tw+O(tw)}\cdot n$ time algorithm for undirected graphs with $n$ vertices and treewidth $tw$ is proposed by Scheffler (Technical Report 396, TU Berlin, '94), and it is proved by Lokshtanov, Marx, and Saurabh (SIAM J. Comput. '18) that, under the ETH, there exists no $2^{o(pw\log pw)}\cdot n^{O(1)}$ time algorithm for either directed or undirected graphs with pathwidth $pw$ and for $k=Ω(pw^4)$. It has not been known whether the lower bound also holds for a smaller $k$. In this paper, we prove that, for both the directed and undirected cases, there is an algorithm faster than Lokshtanov et al.'s lower bound for $k=tw^{o(1)}$ by proposing a $2^{O((tw+k)\log k)}\cdot n$ time algorithm. Besides, we prove a lower bound that, under the SETH, there exists no $(2-ε)^{pw\log pw}\cdot n^{O(1)}$ time algorithm for directed graphs and for a general $k$. This lower bound is tight because, with slight modifications, Scheffler's algorithm runs in $2^{pw\log pw+O(pw)}\cdot n$ time also for directed graphs.

cs.DS↗

The lepton flavor universality including the $b\rightarrow c l ν$ process in the $U(1)_X$SSM

This paper calculates the lepton flavor universality including the $b\rightarrow c\ell ν$ process under the U(1) extension of the minimal supersymmetric standard model ($U(1)_X$SSM), and provides numerical analysis and summary. The numerical results are used to create one-dimensional and multi-dimensional plots to analyze the impacts of various parameters on the ratios $(\frac{R_{J/ψ}}{R^{SM}_{J/ψ}}$, $\frac{R_{{D}_s}}{R^{SM}_{D_s}}, \frac{R_{{D^*}_s}}{R^{SM}_{D^*_s}},\frac{R_{Λ_c}}{R^{SM}_{Λ_c}})$. The results show that several parameters have a certain effect on the ratios. Although it can't match the experimental data very well, with the right combination of parameter values, it can give better numerical results than the Standard Model. This shows that the $U(1)_X$SSM is very helpful for studying $b\rightarrow c\ell ν$ processes.

hep-ph↗

ProofGap: Benchmarking Step-Level Formal Reasoning with Local Obligations Derived from Natural-Language Solutions

Existing formal mathematics benchmarks, such as miniF2F, ProofNet, and PutnamBench, primarily evaluate models on constructing complete formal proofs for challenging problems. Because success is measured at the theorem level, these benchmarks offer limited insight into models' step-level formal reasoning. Evaluating this capability separately enables finer-grained diagnosis of model limitations than theorem-level evaluation alone. To fill this evaluation gap, we introduce ProofGap, a fine-grained benchmark for step-level formal reasoning. ProofGap is constructed through a natural-language proof-processing pipeline that decomposes each reasoning step into one or more aligned proof gaps. Applying this pipeline to natural-language solutions to 3,015 exercises in B. P. Demidovich's Problems in Mathematical Analysis yields 26,116 gaps. The benchmark focuses on mathematical analysis, a domain that remains challenging for current models. By supplying the local context and target explicitly, gap completion isolates local formal proof construction from end-to-end proof composition, enabling more precise localization of model failures. Natural-language solutions serve as the provenance of these obligations, while the benchmark task itself starts from an already formalized local context and goal. Beyond benchmarking, the same pipeline may support future proof-verification systems, provided that semantic translation and sequential proof composition are handled reliably.

cs.PL↗

Gaussian Process Modeling of Time Series

Gaussian processes (GPs) provide a flexible nonparametric framework for modeling time series through appropriately chosen kernel functions. This chapter introduces the basic formulation of Gaussian processes, commonly used kernels, GP regression, hyperparameter estimation, and model evaluation using in-sample and out-of-sample criteria. Applications to stationary, quasi-periodic, and seasonal time series illustrate how individual and composite kernels can represent different forms of temporal variation. Additive kernels also provide interpretable decompositions into latent components such as trend, smooth local variation, and seasonality, while product kernels allow more complex dependence structures to be constructed. Finally, Gaussian process state-space models (GP-SSMs) are briefly introduced, and a nonlinear example demonstrates how a GP transition model can be combined with particle filtering and smoothing for latent-state estimation.

stat.ME↗

A new lower bound for the Schur-Siegel-Smyth trace problem

We prove that lambda_SSS >= 1.80220 for the smallest limiting trace-to-degree ratio of totally positive algebraic integers, improving the bound 1.80203 obtained by Orloski, Talebizadeh Sardari and Smith. The proof exhibits an explicit probability measure on [0,8] together with eighteen integer polynomials, and verifies their dual inequality on the whole of [0,infinity) by interval bisection. The certificate assumes nothing about the data it is built from: an error in that data can only weaken the bound, never invalidate it.

math.NT↗

Adjoint Reidemeister torsion from hyperbolic gluing equations

Motivated by the study of asymptotics of quantum invariants, Dimofte and Garoufalidis introduced a power series associated to a suitable ideal triangulation of a cusped hyperbolic $3$-manifold. They proved that its constant term can be written in terms of Neumann-Zagier data and the complex shape parameters of ideal tetrahedra, and conjectured that it equals the adjoint twisted Reidemeister torsion. On the other hand, in the study of asymptotics of Turaev-Viro type invariants of cusped $ 3$-manifolds, the authors, together with Liu, Sun and Yang, found that the one-loop terms of their asymptotic expansions could be written in terms of Gram matrices and decorated edge lengths of ideal tetrahedra. In this paper, we prove the conjecture of Dimofte and Garoufalidis and relate the one-loop term appearing in the Turaev-Viro type invariant to the adjoint twisted Reidemeister torsion.

math.GT↗

A low-temperature entropy source for on-chip true random number generation: universal robustness beyond device quality

True random number generation is a critical capability for fault-tolerant quantum computing at millikelvin temperatures. Yet existing Josephson-junction-based TRNGs all rest on a widely accepted but untested assumption: that reliable entropy extraction requires precisely controlled device parameters. Here we show that this assumption does not always hold. We demonstrate a counterintuitive finding: a single current-biased Josephson junction, regardless of its parameter quality, can serve as a cryptographic-grade true random number generator. To establish the universality of this conclusion, we deliberately selected the most extremely deviated devices from fabrication, with critical currents three orders of magnitude away from theoretical predictions and $I_cR$ products an order of magnitude above conventional values, as the ultimate stress test. Even under these extreme conditions, the raw Shannon entropy reaches 0.9981~bit (99.8\% of the theoretical maximum), with a min-entropy of 0.9271~bit. Using a square-wave pulsed-bias scheme, we tune the switching probability to $P\approx0.5$. After SHA-256 post-processing, the bitstreams pass all 15 NIST SP 800-22 tests under a conservative $m=3$ criterion that is more demanding than the standard recommendation, and this certification holds across the entire 100-700~mK operating window of a dilution refrigerator.

quant-ph↗

Absolute continuity and dimension conservation for self-similar sets and measures

Let $A\subset\mathbb{R}^d$, $d\ge3$, be a self-similar set whose defining rotations generate a dense subgroup of $\mathrm{SO}(d)$. For every integer $1\le k<\dim_{\mathrm{H}} A$, we prove that its orthogonal projections $π_V A$ have uniformly positive $k$-dimensional Lebesgue measure, and that their fibres have Hausdorff dimension $\dim_{\mathrm{H}} A-k$ at Lebesgue-almost every point of the projected image. This follows from an absolute-continuity theorem for equicontractive self-similar measures $μ$ with the same rotation hypothesis and finite $t$-energy for some $t>k$. Every projected measure $(π_V)_*μ$ has a density $f_V$ satisfying $\int_{\{f_V>M\}}f_V\,d\mathcal{L}_V^k\lesssim e^{-c(\log M)^{1/3}}$ as $M\to\infty$, uniformly in $V$. Under strong separation, the conditional measures on the fibres are almost surely exact dimensional of dimension $\dim_{\mathrm{H}}μ-k$. The key novelty is to apply Varjú's $L^2$ estimate under dense rotations to $L^1$ smoothing increments of martingale-difference type similarly as Fourier decay is studied. This method requires no uniform spectral gap assumption.

math.DS↗

Constant-Probability Witness Isolation Implies $\mathrm{NP}\subseteq\mathrm{P/poly}$

Valiant and Vazirani isolate a satisfying assignment of a circuit with probability $Ω(1/n)$. Dell, Kabanets, van Melkebeek, and Watanabe showed that success above $2/3$ implies $\mathrm{NP}\subseteq\mathrm{P/poly}$ and asked about the range in between. We show that every positive constant already implies the collapse: if a randomized nonuniform polynomial-size pruning procedure succeeds with probability $ε$ on affine circuit inputs with at most $2^{\lfloor 2/ε\rfloor}$ satisfying assignments, then $\mathrm{NP}\subseteq\mathrm{P/poly}$. Success $10/\log L$ on affine inputs with at most $L^{1/3}$ satisfying assignments suffices, where $L$ is the description length, and on inputs with one or two satisfying assignments the threshold $2/3$ drops to $3/5$. No cryptographic assumption is used, and the procedure may read the entire circuit. The proof compiles a pool of circuits into one circuit whose satisfying assignments are indexed by tags in $\mathbb{F}_2^d$. Each member is assigned an affine region of tag space, and if one member is unsatisfiable, the satisfying set shrinks to that member's region. Because regions may overlap and have different dimensions, the collapse reduces to a combinatorial bound: no set of tags meets more than a $2/d$ fraction of an equally weighted family of affine subspaces of all dimensions below $d$ in exactly one point. This regional counting cannot go below order $1/\log L$. The range between $Θ(1/n)$, achieved by affine hashing, and $O(1/\log n)$ remains open.

cs.CC↗

Noether Symmetry in a Non-Metricity Theory with a Boundary Term: Exact Solutions and Bayesian Cosmological Constraints

We investigate the cosmological implications of a power-law $f(Q, B)$ gravity model, where $Q$ denotes the non-metricity scalar and $B$ is the associated boundary term. We employ the Noether symmetry approach to identify the admissible functional form of the gravitational Lagrangian and to obtain the corresponding conserved quantities and exact cosmological solutions. Using the exact solution, we find that for this model parameter condition $α+β>\frac{3}{2}$, the model exhibits the accelerating phase of the Universe. Furthermore, we have also discussed the model with recent observational data using a Bayesian Markov Chain Monte Carlo analysis. We consider different combinations of Cosmic Chronometers (CC), Pantheon+ \& SH0ES, and DESI DR2 observations to constrain the model parameters and the Hubble constant $H_0$. The inferred value of $H_0$ depends on the adopted dataset combination, ranging from a higher value for the CC+Pantheon+ combination to a lower value for CC+DESI DR2, with the latter yielding $H_0=67.5^{+1.4}_{-1.6}\,\mathrm{km\,s^{-1}\, Mpc^{-1}}$, consistent with the Planck 2018 CMB-preferred value within the standard $Λ$CDM framework. The inferred \(H_0\) is strongly dataset dependent, spanning the range between the early-Universe and locally calibrated determinations. We further examine the posterior covariance and correlation matrices to identify parameter degeneracies and assess the robustness of the MCMC constraints. The background evolution also exhibits late-time accelerated expansion and remains compatible with the observational constraints. These results demonstrate that the $f(Q, B)$ model provides a viable alternative description of late-time cosmology.

gr-qc↗

The strong (non-induced) Turán numbers

In this paper we introduce and explore the following new graph invariant: For a graph $G$ on $k$ vertices, $G \neq K_k$, let $st(n,G)$ denote the maximum number of edges in a graph of order $n$ which does not contain any subgraph on $k$ vertices strictly containing $G$. A basic relation to classical Turán numbers is developed via the following: For $G$ on $k$ vertices, let $D(G) = \{ H : |H| = |G|, H = G + e \}$. Using this notion we prove that $ex(n,G) \leq st(n,G) = ex(n, D(G) ) \leq \min \{ ex(n,H) : H \in D(G) \}$ holds for all $n \geq |G|$. The family $D(G)$ happened to be smoothly amenable to the use of classical extremal results, and in many cases allows us to get asymptotically sharp estimates as well as exact values of $st(n,G)$. From the many results proved here we state the following as an illustration. (1) If $χ(G) \geq 3$ and $χ(D(G)) = χ(G)$, then $st(n,G) = (1+o(1))ex(n,K_{χ(G)})$. (2) If $χ(D(G)) = χ(G) +1$, then $G$ is a complete $χ(G)$-partite graph and $st(n,G) = ex(n,K_{χ(G) +1})$ for $n$ sufficiently large. (3) For $k$ odd, $k\geq 5$, $st(n,C_k) = ex(n,C_k) = ex(n,K_3)$ for $n$ sufficiently large. (4) If $T$ is a tree of order $q$ with diameter $k \geq 2$ and $q \geq k+1 \geq 3$, then $ex(n, \{C_3,...,C_{k+1}\}) \leq st(n,T) \leq ex(n, \{C_3,...,C_{k+1}\}) + (q-1)n$. Many results concerning even cycles, theta graphs, dense bipartite graphs and graphs of the form $G = G^* \cup tK_1$ are obtained, moreover the value of $st(n,G)$ is computed for all graphs on at most 4 vertices.

math.CO↗