arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,027 records · Page 57Linked to original sources

Linear-Time FPT Algorithm for Surface Disjoint Paths via Surface Cutting

We study the \textsc{$k$-Disjoint Paths} problem on a graph embedded on a surface with bounded Euler genus. Given a graph $G$ with $n$ vertices and $k$ vertex pairs embedded on a surface of Euler genus $g$, we present a $2^{O(k^2+g^2)}n$-time algorithm that computes $k$ pairwise vertex-disjoint paths connecting the given vertex pairs if such paths exist. Our approach relies on the decomposition of $G$ into $O(k+g)$ planar subgraphs while bounding the complexity of the boundaries between these subgraphs. This approach enables the use of techniques for compressing linkages in planar graphs. Moreover, our techniques yield two kernels of size polynomial in $k$, $g$, and the treewidth of the graph, and of size $2^{O(k+g)}$. These results extend recent advances on \textsc{$k$-Disjoint Paths} on planar graphs [Cho et al. SODA 2023] and [Włodarczyk and Zehavi FOCS 2023] to surface-embedded graphs.

cs.DS↗

Monte Carlo sampling of first-order QED processes in laser and pulsar plasmas

Monte Carlo sampling of strong-field quantum electrodynamics processes underpins simulations of high-intensity laser experiments and of astrophysical compact-object magnetospheres. Sampling an event requires the total rate of the process together with the cumulative probability that determines how energy is partitioned between the produced particles. Simulations typically tabulate both in advance and invert the tabulated probability numerically. Here we replace this procedure with elementary-function approximations for synchrotron radiation and the nonlinear Breit--Wheeler process. For each process, we approximate the auxiliary function that sets the total rate, as well as the cumulative probability, with Padé approximants chosen so that the inversion reduces to a quartic equation. This yields the sampled quantum parameter---electron $χ_e$ or photon $χ_γ$---in closed form. The approximations and the particle spectra sampled from them agree with the exact results to within $1\%$. The procedure requires no lookup tables, no interpolation, and no numerical root finding, and can be inserted directly into radiative particle-in-cell codes.

physics.comp-ph↗

Long-term behaviour of refracted Lévy processes in a half-line

We study the long-term behaviour of refracted Lévy processes within a half-line, with killing on exiting the domain and at state-dependent rate inside it. We identify the decay rate of the killed semigroup together with the associated invariant function and measure, and we obtain convergence at exponential rate to a Yaglom limit. The proofs make use of R-theory and Lyapunov function techniques.

math.PR↗

GCUL: Ambiguity Identification in Text Emotion Classification via Cluster-Guided Learning

Selective classification enables a model to abstain from predictions on uncertain instances, but existing approaches typically reject them through confidence scores, predefined coverage constraints or instance-level distance measures. These approaches may overlook the collective geometric structure of difficult samples in learned representation spaces. We propose Guided Clustering-based Uncertain Learning (GCUL), a geometric-guided selective classification framework that identifies misclassified and ambiguous instances as a potential confusion attractor in the representation space. GCUL uses a three-phase procedure to initialize, cluster, and explicitly relabel this uncertain region, allowing the rejection boundary to emerge from the underlying representation geometry rather than from a prescribed rejection rate. We further derive a selectivity score and a geometric sufficient condition that characterizes when rejection can provide positive operational utility, enabling pre-deployment feasibility assessment. GCUL improves DistilBERT accuracy from 89.37 percent to 94.98 percent with less than 9 percent rejection. Beyond accuracy, our selectivity score correctly pre-detects the only dataset (GoEmotion) where all baselines fail, and controlled simulations yield 6.1 percent Type-I and 0 percent Type-II errors, validating the sufficient condition's conservatism. These results suggest that collective representation geometry provides a useful alternative perspective for selective prediction.

stat.ML↗

Grammatical "grandmother neurons" are rare in LLMs

Understanding how Large Language Models (LLMs) encode linguistic structures remains a fundamental challenge in interpretability research. While diagnostic classifiers (or "probes") are widely used for this task, they face significant methodological criticism: training auxiliary classifiers introduces capacity confounds and calibration issues, often making it difficult to distinguish the model's intrinsic representations from the probe's ability to learn the task. To address these limitations, we introduce a probe-free framework for localizing linguistic selectivity at the individual neuron level. Leveraging the controlled contrasts of linguistic minimal pairs, we propose a Neuron Separability Index (NSI), a metric that directly quantifies how reliably single neurons differentiate grammatical from ungrammatical constructions without parameter updates. Applying NSI across 68 linguistic paradigms and seven checkpoints reveals three main patterns: 1) raw separability reaches near-peak levels earlier for morphological and syntactic distinctions than for syntax-semantics interface and conceptual distinctions. 2) after permutation normalization, single-unit selectivity is sparse, weak, and narrowly tuned: only a small fraction of units are sensitive to an average paradigm, and strongly selective "grandmother neurons" are rare. 3) whole-vector linear separability, single-neuron selectivity, and behavioral competence are largely dissociated, and targeted ablations further separate activation selectivity from causal reliance.

cs.CL↗

Hyperbolic Multimodal Continual Learning: A Closest-Admissible Solution

Existing continual-learning methods protect parameters, replayed examples, or Euclidean feature subspaces. When applied to hyperbolic multimodal models, they do not explicitly preserve the Lorentz geometry that jointly encodes within-modality similarity, cross-modal correspondence, and semantic hierarchy; sequential updates can therefore retain task scores while still distorting previously learned relations. We address this gap with Hyperbolic Multimodal Continual Learning (HMCL). We show that preserving the old multimodal geometry amounts to restricting all modalities to one shared hyperbolic isometry, which induces a family of admissible first-order parameter changes. We formulate a joint closest-admissible (CA) correction that retains the shared rotation best matching the candidate modal updates; its minimal-rotation (MR) special case fixes this rotation to zero. Both variants correct the displacement realized by AdamW, and task anchoring bounds within-task accumulation while preserving learning freedom. Across a unified 16-task classification-retrieval stream with three hyperbolic backbones, HMCL improves final performance and backward transfer over sequential fine-tuning and four continual-learning baselines; HMCL-CA gives the highest Overall score on every backbone. A modality-extended stream confirms the retrieval gains. Representation analyses find 81.2 to 95.5 percent less radial, angular, cross-modal, and paired-distance drift; ImageNet-WordNet results show better semantic ancestry and radial hierarchy.

cs.CV↗

FlowAtom: Atom-Based Evidence Aggregation for Multi-Label Website Fingerprinting

Identifying the set of monitored websites in mixed encrypted traffic is challenging because an individual flow often provides only partial evidence of website identity. To address this challenge, we propose FlowAtom, which constructs shared prototypes, called Atoms, from flow representations without website labels. Specifically, FlowAtom pretrains a flow encoder on external unlabeled traffic and aggregates Atom responses across flows within each observation window into a fixed-dimensional, permutation-invariant representation for monitored website-set prediction. Across Direct HTTPS, Trojan, and VMess, FlowAtom achieves micro-F1 scores of 97.82%, 94.43%, and 93.92% in closed-world evaluation, respectively, and consistently outperforms the evaluated baselines in open-world evaluation on windows containing monitored visits. The code is available at https://github.com/aimafan123/FlowAtom.

cs.LG↗

Filtered deformations of Lie groupoids

Let $G \rightrightarrows G^{(0)}$ be a Lie groupoid, $\mathrm{A} G \rightarrow G^{(0)}$ its Lie algebroid and $X_1, \dots, X_r$ a family of sections of $\mathrm{A} G$ satisfying a Lie bracket generating condition of Hörmander type. We aim to build a pseudodifferential calculus allowing to study a Helffer-Nourrigat's conjecture on the groupoid $G$; in particular, we want differential operators of the form $\sum_{i = 1}^r X_i^2$ to have an invertible symbol. In this article we achieve the geometrical part of this construction by defining a "weighted" version of the deformation to the normal cone $\mathrm{DNC}(G, G^{(0)}) \rightrightarrows G^{(0)} \times \mathbb{R}_+$. Heuristically, we deform $G$ around $G^{(0)}$ with a "zoom" parameter $t \in \mathbb{R}_+$ by stretching $G$ by $t$ in the directions of the sections $X_i$, $t^2$ along $[X_i, X_j]$, $t^3$ along $[X_i, [X_j, X_k]]$ etc. In the case where $G = M \times M$ we recover a construction of Mohsen, and when the structure is equiregular we recover a construction of van Erp-Yuncken.

math.DG↗

Diverse Representation in Approval-Based Committee Voting

The study of approval-based committee (ABC) voting has so far focused predominantly on proportional representation. The canonical notion of diverse representation, based on the Chamberlin--Courant score, counts the number of voters with at least one representative in the committee, making it an individualistic notion. We develop a more comprehensive theory that instead requires the representation of many groups of voters, grounding it in the justified representation (JR) axiom, which we strengthen in two directions. First, we study the existing axioms of Strong JR (SJR) and Semi-Strong JR (SSJR), which consider the same cohesive groups as JR but demand stricter representation. We show that neither can be optimized efficiently on general domains (unless P=NP), but both can be on the Candidate Interval domain. Second, to capture the unique and defining opinions of a group, we introduce Distinctive Representation (DR) and its local optimization variant, Local DR. We address satisfiability and computation time, and show (Local) DR to be distinct from known proportionality and diversity axioms. Experiments on real-world and synthetic data show Local DR performs well on multiple empirical measures of diversity.

cs.GT↗

Where LLM Graders Succeed and Break: Evidence from Two Computer-Science Exams

One long-form exam in a large course costs hundreds of grader-hours, and qualified graders are scarce; LLM graders are a tempting alternative. To show its pitfalls we grade a practical Computer Vision exam ($570$ dual-graded students) under $171$ configurations spanning closed and open-weights models; the best reaches mean absolute error $1.64/35$, below the $2.61/35$ two human graders achieve against each other. The catch is the prompt: a short ''strict grader'' preamble drives $14$ of $17$ open-weights models out of the graded band ($\text{MAE} \ge 8$), three stopping grading altogether. The damage traces to the preamble's two credit-withholding sentences, not to tone or model scale; one of them, ''never give partial credit'', alone makes two of three probed models stop grading. The closed flagships of three vendors shift calibration under it but stay in the band. In $162$ further configurations on a second, independent Machine Learning exam from another course ($1{,}038$ dual-graded students), the preamble worsens ten models, moving three out of the band into collapse and one into refusal, yet improves seven whose neutral prompts over-mark: the vulnerability replicates, but its direction is exam-specific. Light LoRA fine-tuning repairs it: one adapter on the two exams' pooled $\sim 3{,}900$ graded examples brings five small open models to parity or better with a human grader in agreement with the grader pair, and sensitivity to the three harsh personas nearly vanishes ($\le 0.32$ MAE). We release the anonymised dataset, full ablation grid, and grading, fine-tuning and analysis pipelines.

cs.CL↗

A Study of the Limits of Collaborative DCT-Based Image Denoising via Interpretable Neural Networks

Image denoising remains a fundamental problem in image restoration, with applications in photography, biomedical, and scientific imaging. Modern deep neural networks achieve strong performance by learning powerful image priors, but often rely on large black-box models with limited interpretability. In contrast, DCT-based sliding-window and collaborative filtering methods such as BM3D offer clear algorithmic structure, but depend on handcrafted and non-differentiable operations. This work studies how far such structured collaborative filtering principles can be pushed when reformulated as trainable models. We introduce DeepBM3D, a compact fully differentiable architecture that combines non-local patch grouping, DCT-domain filtering, and multi-stage refinement within a BM3D-inspired pipeline. Lightweight convolutional feature extractors guide patch grouping, while filtering is performed through learned Wiener weights in the DCT domain. Experiments show that DeepBM3D improves over classical and hybrid baselines, remains competitive with FFDNet at low and moderate noise levels, and performs particularly well on repetitive textures.

cs.CV↗

FDR-Controlled Variable Selection for Generalized Linear Models and Cox Regression with Virtual Dummies

In genomics, imaging and clinical studies, only a few of many candidate predictors are often nonlinearly associated with a response that may be, e.g. binary, categorical, a count or a censored event time. The Terminating-Random Experiments (T-Rex) selector is a scalable variable selection method that controls the false discovery rate (FDR) by letting synthetic null variables (dummies) compete with the real predictors. While the FDR control theory embraces more general settings, to date, the T-Rex selector has been specified only for linear models. We propose a memory-efficient selection procedure with FDR control for generalized linear models and Cox regression by extending the recently developed virtual dummy construction to score-based forward selection for Bernoulli, Poisson, multinomial and Cox responses. The virtual-dummy-based selection path remains equal in distribution to explicit augmentation, so FDR control carries over under the same assumptions. Simulations confirm this equivalence and the power gained by correct model specification. Real-world applicability is illustrated on simulated genotypes and on cancer survival data.

stat.ME↗

PQLS: A Quasilinear Gyrokinetic Transport Solver with a Bayesian Saturation-Rule Closure

Quasilinear models make gyrokinetic turbulent-transport predictions sufficiently fast for integrated modelling, but their predictive capability is limited by two factors: the physical and geometrical applicability of the linear solver, and the validity of the saturation rule used to close the model. We present the Predictive Quasilinear Solver (PQLS), a quasi- linear gyrokinetic transport solver formulated in general magnetic geometry. Its implementation as an eigenvalue solver retains electromagnetic and collisional effects, provides access to dominant and subdominant modes and is differen- tiable with respect to all plasma parameters. Linear benchmarks against GENE reproduce the growth rates, frequencies, and eigenfunctions. We additionally formulate the saturation-rule closure as a Bayesian inference problem that distin- guishes uncertainty in its fitted coefficients from the residual model-form uncertainty. The approach is demonstrated by calibrating the SAT3 rule on PQLS quasilinear weights against published nonlinear CGYRO cases. In addition to improving the robustness of the calibration, the new method also quantifies the uncertainty in each of the fit coefficients. Such uncertainty is propagated through transport calculations to produce error-aware profiles that are compared to the ones obtained from the full gyrokinetic simulation, showing excellent agreement.

physics.plasm-ph↗

Nonorthogonal variational quantum simulation for quantum chemistry

Dynamical simulation of quantum many-body systems is a central task in quantum chemistry and requires efficient wavefunction representations. Although tensor-network and neural-network quantum states have achieved considerable success in ground-state calculations, entanglement growth hinders their application to quantum dynamics. Quantum computing may offer a route to quantum advantage, but algorithms such as Trotterization generally require fault-tolerant quantum computers, whereas near-term hardware supports only limited circuit sizes. Variational quantum simulation (VQS) instead represents the evolving state with a single shallow, fixed-size parameterized quantum circuit (PQC). Here, we introduce nonorthogonal variational quantum simulation (NOVQS), which applies linear combinations of parameterized quantum states to real- and imaginary-time evolutions. We design a shallow, hardware-friendly ansatz tailored to second-quantized electronic-structure Hamiltonians, together with resource-efficient protocols for measuring the matrices and vectors in the parameter equations of motion. Error analysis and resource estimation are also provided. Numerical simulations of hydrogen chains and the nitrogen molecule demonstrate that a collection of shallow, or even single-layer, PQCs can match or outperform a much deeper PQC in VQS. In particular, we identify a trade-off between circuit number and depth. NOVQS therefore provides a flexible route to enhancing wavefunction expressivity under circuit-depth constraints, making it a promising framework for quantum dynamics simulations on near-term quantum processors.

quant-ph↗

From Hits to Tracks: A BERT-based Tracking Model for Track Reconstruction in Drift Chambers

Track reconstruction in drift chambers is essential for momentum measurement and particle identification at electron-positron colliders. While Transformer architectures have transformed many sequence-processing domains, their application to tracking in high energy physics is still being explored. We present a model that combines a BERT encoder with a Transformer decoder to perform hit-to-track association through autoregressive sorting. The model is evaluated on the DCTracks open dataset that provides realistic drift chamber simulations with varying particle types, momenta, track multiplicities, and noise conditions. Across single-track, two-track, and multi-track samples, the model achieves high hit and track efficiencies while keeping the rates of clones and fakes very low. It also works well in the reconstruction of displaced vertices. These results show BERT-based sequence-to-sequence models as a promising approach for track reconstruction in low-background, precision-oriented experiments.

hep-ex↗

Stability of electrokinetic Couette states in the Poisson--Nernst--Planck--Navier--Stokes system

In this paper, we study a two-dimensional Poisson--Nernst--Planck--Navier--Stokes system in a periodic channel driven by wall motion and an imposed tangential electric field, with two ionic species that may have unequal diffusivities. At the electroneutral Couette state, a diffusivity-weighted ionic energy combines with the triangular structure of the linearized operator to yield exponential linear stability for fixed \((A,E_0)\) and all positive diffusivities. Small-data nonlinear exponential stability is then obtained on a complex interpolation space adapted to the graph domain of the generator by combining analytic-semigroup smoothing with quadratic estimates. For unequal diffusivities, the nonzero streamwise modes of the linearized ionic subsystem satisfy a bounded-channel enhanced-dissipation estimate at rate \(D_{\min}^{1/3}|A|^{2/3}\) under an explicit strong-shear condition. When \(D_+=D_-\), a Fourier-mode factorization separates the common Couette advection--diffusion operator from a contractive drift--reaction semigroup, so the same scalar mixing rate is retained without any smallness condition on the electrostatic coupling. We also construct an exact Poisson--Boltzmann/electroosmotic Couette family for prescribed wall potentials. For sufficiently small wall-potential amplitude, its linearized generator is a small graph-domain perturbation of the electroneutral generator, which yields linear and nonlinear exponential stability on the same interpolation scale.

math.AP↗

A Simple Gripper Interface for Simulator-Agnostic Cloth Manipulation

This paper presents a grasping model for cloth manipulation specifically tailored to ease the deployment of robotic control methods. The model is robust, fast and easy to implement avoiding at the same time contact and friction considerations between the gripper and the cloth in favor of simple positional constraints. The gripper is described by its pose, jaw state, and an attached grasping volume. Two kinds of grasping volumes are considered: an axis-aligned box to simulate a pinch grasping and a square pyramidal volume to simulate point grasping. When the gripper closes, the discrete cloth positions lying inside this volume are selected, stored in the local gripper frame, and then transported with the gripper motion. A simple squeezing step is also included to progressively move the selected cloth positions toward the center of the grasping region, avoiding an instantaneous displacement at closure. The model can be used in any simulator as it only requires access to discrete cloth positions and a mechanism for imposing target positions as constraints. We implement our grasping model in conjunction with a constraint-based inextensible cloth simulator, where grasping is implemented as moving positional equality constraints coupled with stretch, shear, collision, and table contact projection steps. The same gripper trajectory is applied on a robot arm to fold a real piece of cloth, serving as a simple bridge between simulation and physical cloth manipulation and showcasing the realism and practicality of our idealized grasping model.

cs.RO↗

SkinAgent AI: A Safety-Grounded Multimodal Agentic Framework for Non-Diagnostic Skincare Support

Consumer-facing skincare AI must coordinate visual evidence, product information, tool use, and user-facing actions within explicit evidence and safety boundaries. This study evaluates SkinAgent AI, a non-diagnostic multimodal framework that combines visual concern routing with grounded and auditable LLM-based orchestration. The architecture includes routing for Acne, Pores, and Wrinkles; photograph-based skin-type estimation; count-informed ordinal acne-severity support; typed tools; database-grounded recommendation and action functions; deterministic safety, privacy, and evidence checks; approval before state-changing actions; and structured trace and replay mechanisms. Visual-model performance and system-level agent behavior were evaluated separately. Across three seeds, the skin-condition routing model achieved 99.84% +/- 0.07% accuracy. Skin-type estimation achieved 88.85% accuracy, while count-informed acne-severity support achieved 84.59% accuracy with a quadratic weighted kappa of 0.9076. On a locked but non-independent 240-case system benchmark, intent accuracy was 80.00%, exact tool-set match was 62.92%, and strict task completion was 47.08%. No violations or successful cross-user leakage events were observed in the finite safety and privacy test suites. Tool-selection errors, incomplete grounding of product attributes, and unreliable failure fallback nevertheless remained. These findings support the feasibility of bounded, database-grounded, and traceable agent orchestration for non-diagnostic skincare assistance. They do not establish clinical readiness, external generalization, formal privacy guarantees, or universal safety. Independent validation, expert assessment, robustness and fairness testing, and prospective evaluation in real-world settings remain necessary.

cs.AI↗