arXiv Science⌕ Search

SEARCH · arXiv Science

Search arXiv Science

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,621 records · Page 90Linked to original sources

STcubeOperator: A Framework for Analyzing Spatiotemporal Event Data

The analysis of spatiotemporal event data is essential for informed decision-making in domains such as disaster response, conflict analysis, or intelligence investigations. However, the complexity and interdependence of spatial, temporal, and multiple thematic attributes pose significant challenges for both analysis and visualization. While space-time cubes (STCs) present a powerful integrated visualization technique to analyze this kind of data, existing approaches often lack support for complex exploratory workflows, thus limiting the ability to derive meaningful insights. We address this gap by introducing STcubeOperator, a novel framework that models analysis tasks through space-time cube operations, considering them in context of visualizations, interactions, and computational choices, and implement them in an interactive visual analytics environment. By expressing analysis tasks as a sequence of multiple elementary operations--such as filtering, chopping, and flattening--our approach enables analysts to dynamically explore data from different perspectives. We further provide an open-source prototype implementing the operations in a 3D interactive environment to facilitate task-based exploratory analysis of spatiotemporal event data. We demonstrate the applicability of our framework with a case study based on real-world data on strategic and military operations in the Russia-Ukrainian War, showing its capabilities to reveal spatiotemporal patterns. An expert user study (n=8) shows how specific tasks can be solved with our framework, highlights the versatility of our approach, and provides valuable insights on which operations experienced analysts utilize in practice.

cs.HC↗

Permutations from ranking independent random variables

Let $X_1,\ldots,X_n$ be almost surely distinct independent real-valued random variables. Let $σ$ be the random permutation of $\{1,\ldots,n\}$ such that $X_{σ(1)}<X_{σ(2)}<\cdots< X_{σ(n)}$. We show that the set of laws of $σ$, as the laws of $X_1,\ldots,X_n$ vary, has semialgebraic dimension \[ \sum_{k=2}^n {n \choose k}(k-1)! \] as a subset of the $(n!-1)$-simplex. This establishes a conjecture of Babson, Duchin, Iseli, Poggi-Corradini, Thurston, and Tucker-Foltz who proved that the dimension is upper bounded by the above expression. We prove their conjecture by constructing a family of finitely supported laws for $X_1,\ldots,X_n$ that provides the correct dimension. We also give an alternative proof of the upper bound using a theorem of Radford. Furthermore, we show that the Mallows law on permutations cannot arise as a law of $σ$ for $n\geq 4$, answering a question of Ed Crane.

math.PR↗

Automated Assembly Instruction Generation from CAD Models Using Grounded Large Language Models: A Human-in-the-Loop Framework

Assembly documentation is a downstream manufacturing artifact that is still usually authored by interpreting CAD models by hand. Structured product data and large language models are both available, yet studies of CAD interpretation, assembly sequence planning, instruction writing, and human oversight have largely proceeded separately. This paper formulates CAD-grounded assembly instruction generation: the production of natural-language assembly procedures constrained by structured engineering information extracted from CAD models. The proposed framework maps a STEP assembly to a typed ProductGraph intermediate representation, derives a precedence order by deterministic topological sorting, realizes each step as language conditioned only on selected graph context, attaches per-step visual documentation, and applies rule-based and model-assisted checks. PDF export remains disabled until a human reviewer resolves every quality flag. The case study establishes endto-end feasibility on a built-in six-part reference assembly: the pipeline preserves a reported assembly order and carries quantity, material, and torque into an exported manual page. Generalization and geometric validation remain open empirical questions. The contribution is an architecture that separates engineering state, deterministic reasoning, grounded language realization, verification, and human release.

cs.LG↗

Emergent SSH physics and localization in a cavity--atom system beyond the rotating-wave approximation

We investigate a cavity--atom system beyond the rotating-wave approximation and recast its dynamics as an effective tight-binding problem in the excitation-number basis. The same experimentally relevant light--matter platform exhibits two distinct regimes controlled by the photon number. In the large-photon-number regime, the hopping amplitudes become nearly uniform and the effective lattice reduces to a dimerized SSH-type chain subject to a linear energy gradient. The spectrum then organizes into one or two Wannier--Stark ladders, giving rise to controllable Bloch and Bloch--Zener oscillations and quantitatively explaining the observed revival patterns. In the small-photon-number regime, the intrinsic $\sqrt{\bar n+(\cdots)}$ dependence produces a pronounced hopping deformation, namely dimerized couplings that increase along the excitation-number lattice. This deformation reshapes the density of states, enhances collective localization, and generates an energy-resolved spatial bias of the eigenstates. The corresponding real-time dynamics displays direction-dependent anomalous diffusion together with Bloch-like revivals and beat phenomena in selected parameter windows. Our results show that the same cavity--atom platform can realize both emergent ladder spectra and hopping-induced localization in excitation-number space.

quant-ph↗

Inference in Panel SVARs with Two-Way Dependence

We develop inference for heterogeneous panel vector autoregressive (VAR) models and their structural impulse response functions, where the error terms are dependent in the cross-sectional and time dimensions (two-way dependence). For proxy-identified structural VARs, we first adapt mean-group estimation and construct a closed-form pooled identification. Considering the reduced-form VAR dynamics, residual covariances, and structural parameters, we derive a joint central limit theorem under joint limits in the cross-sectional and time dimensions. We then propose a recursive-design panel moving-block bootstrap that resamples the estimated error terms in (i) the temporal, (ii) the cross-sectional, or (iii) both dimensions jointly, and prove consistency of the joint panel-block scheme under two-way dependence. Simulations show coverage close to the nominal level for the joint scheme but severe undercoverage for cross-sectional resampling.

econ.EM↗

Forms of LLM-Integrated Applications from LLM-Chats to Autonomous AI Agent System

Large language models (LLMs) are increasingly embedded as components in software systems, marketed under labels such as chatbot, copilot, retrieval-augmented generation, workflow, coding agent and AI agent. Whether these labels denote genuine architectural forms or serve as branding has not been assessed systematically. In the sources surveyed, labels do carry architectural content, most clearly in vendor usage: copilot denotes a router-worker architecture operating a host application under step-by-step user confirmation, while the more recent shift to the label agent coincides with AI-planned multi-step execution of which the user sees only the outcome. The coding agents of four major providers share one architecture, a reason-and-act loop delegating to subagents. This survey describes seven recurring forms---LLM chats, custom agents, retrieval-augmented generation (RAG), AI-enhanced workflows, copilots, coding agents, and, in part, agentic RAG---in a common vocabulary of agents and tools. Each is characterized along four structural dimensions (agentic RAG only partially): the architectural pattern, the control of execution and the point of user intervention, the number of agent calls per task, and tool use. An illustrative corpus of 22 systems from research publications and vendor documentation grounds the descriptions and shows where they reach their limit.

cs.CL↗

Reach-Stabilize Control of Control-Affine Systems with Unknown Affine Parameters

This paper considers the reach-stabilize prob- lem for a class of nonlinear control-affine systems with unknown parametric uncertainties, where the system states must remain in a safe set at all times, or enter a safe set in finite time and then remain in it, and converge to a goal point. An estimator is designed that generates a parameter estimate and a computable, nonincreasing bound on the estimation error from a known initial error bound, without requiring persistence of excitation. The bound defines margins that are added to a control barrier and a control Lyapunov function condition, which are then enforced in a quadratic program for efficient control design. It is shown that, for the closed-loop system, the safe set is forward invariant when the system starts within the safe set and is finite-time reachable from outside the safe set with a recovery-time bound that does not depend on the unknown parameter, and that the goal point is exponentially stable. In two numerical examples, with the initial state outside and inside the safe set, the proposed control policy, a baseline oracle control policy that uses the true parameter, both recover or remain in the safe set and converge to the goal point, while another baseline control policy that uses a fixed, incorrect, parameter does not remain

math.OC↗

Can Decision Models Understand Stance? Evaluating Jev Against General-Purpose LLMs

Stance detection requires identifying an author's attitude toward a given target, sometimes based on conversational context. Jev, a specialized decision model designed for structured decision-making, offers an alternative to general-purpose large language models (LLMs). In this work, we evaluate Jev on two stance detection datasets, VAST (English texts) and ZS-CSD (Chinese conversations), comparing it with four general-purpose LLMs and two fine-tuned models. Results show that Jev achieves competitive performance on VAST, matching GPT-5.6 and outperforming the other general-purpose LLMs. However, it falls behind stronger LLMs on ZS-CSD, particularly in distinguishing favor from against. Further analysis suggests that this limitation may be related to understanding reply relationships and stance direction rather than conversation length alone. These findings highlight both the potential and limitations of Jev for stance detection.

cs.CL↗

The differentiation and integration operator on weighted spaces of analytic functions in the weight sequence setting

We investigate the differentiation and integration operator defined on resp. between weighted spaces of analytic functions in the special situation when the weight is given in terms of a weight sequence. More precisely, in this case the weight is defined via the associated weight function which is a crucial object in the ultradifferentiable setting and the information can be then expressed purely in terms of the underlying weight sequence. Due to the additional structure this setting is exploited in order to construct (counter-)examples, to compare conditions on weights in the different weighted frameworks, and to study and to understand in a more precise qualitative way required and admissible growth restrictions on the weights.

math.FA↗

Fluid deformable surfaces with variable thickness - a Surface Shallow-Water-Helfrich model

Epithelial tissues play a fundamental role in morphogenesis. Mechanically they can be viewed as thin soft materials exhibiting a solid-fluid duality. The Surface Navier-Stokes-Helfrich model accounts for these properties by combining bending and surface hydrodynamics. In order to account for varying cell thickness we incorporate a thickness field in the spirit of a shallow-water equation, but defined on the (self-)evolving surface. This replaces the inextensibility constraint of the two-dimensional fluid and with it the conservation of surface area by an incompressibility constraint of the thin film fluid allowing for changes in surface area. We develop a numerical scheme based on surface finite elements, perform convergence tests, demonstrate the impact on shape evolution of closed surfaces with a constant enclosed volume and discuss implications on modeling morphogenesis.

cond-mat.soft↗

Large-Scale Partition-Based RIS Beamforming For Uplink RIS-Equipped Multi-User Systems: Asymptotic Analysis

Combining a reconfigurable intelligent surface (RIS) with a receive antenna array is a promising low-complexity architecture for multi-user uplink reception, but its performance analysis for more than two users has remained an open problem: the zero-forcing (ZF) signal-to-interference-plus-noise ratio (SINR) no longer admits an explicit, low-dimensional closed-form expression, and its distribution is analytically intractable for design purposes. This paper addresses this gap for a K-user, K-antenna uplink system in which a large-scale, L-element RIS, partitioned into K user-dedicated sub-surfaces, precedes ZF reception at the base station. Through an asymptotic analysis in which the sub-surface sizes grow without bound, we show that the orthogonal projector underlying the ZF SINR converges to a rank-one matrix aligned with the desired user's channel. We use this convergence to derive a closed-form asymptotic approximation for the average per-user SINR that depends only on deterministic channel parameters. Treating this expression as a tractable design objective, we prove that equal partitioning is approximately sum-rate-optimal at leading order regardless of path-loss asymmetry across users. We also develop a low-complexity greedy pairwise-transfer search that refines the partition beyond this leading-order optimum. Monte Carlo simulations across a range of system and RIS sizes confirm that the closed-form SINR tracks the exact simulated rate closely once the RIS is large relative to the number of users. The greedy search also yields consistent, if modest, sum-rate gains over equal partitioning, validating the theory as both an accurate performance predictor and a practical design tool.

eess.SP↗

Uniform Height Gaps in Arbitrary Characteristics

We prove uniform height gaps for subvarieties of abelian varieties in arbitrary characteristic, extending the new gap principle of Gao-Ge-Kühne. Our proof goes through the theory of adelic curves and globally valued fields, and proves the Bogomolov conjecture for an arbitrary globally valued field. This gives a different way to obtain uniform Bogomolov-type results, differing from the approaches of Dimitrov-Gao-Habegger-Kühne and Yuan. First, we prove the Bogomolov conjecture over globally valued fields for divisors. We then reduce the case of general subvarieties to the case of divisors by an induction argument. Specializing to the case of global function fields, we obtain a new proof of the geometric Bogomolov conjecture following the strategy of Gubler and Yamaki.

math.NT↗

RobustLDS: Learning linear dynamical systems under adversarial corruptions

We consider the problem of learning linear dynamical systems under adversarial contamination from a single trajectory of length $T$. While identification of linear dynamical systems itself is well-studied, the problem of robust system identification under adversarial contamination is relatively less explored. In this work, we study the setting where a fraction of the $T$ observations are contaminated by adversarial outliers. We propose different estimators based on relaxations of least-trimmed squares along with an alternating minimization algorithm. Furthermore, we also propose two estimators which exploit the group-sparsity (through penalization/hard-constraints) of the outliers. For the estimator with group-sparse penalty, we derive non-asymptotic error bounds which establish its robustness to outliers. We also show empirically that the proposed estimators work well in practice.

stat.ML↗

Beyond Visual Enhancement: Adaptive Multi-Context Steering to Mitigate LVLM Hallucinations

Hallucination remains a significant challenge in Large Vision-Language Models (LVLMs). Existing training-free methods generally mitigate hallucinations through contrastive decoding or visual enhancement, often increasing the relative influence of visual evidence during generation. This raises a fundamental question: Can LVLMs dynamically regulate the contributions of different context sources to suppress hallucinations? In this work, we investigate and quantify how LVLMs coordinate multiple context sources during decoding and examine how this intrinsic behavior can guide hallucination mitigation. We find that LVLMs exhibit an intrinsic vision-attending tendency that can guide adaptive visual steering, while textual contexts can also contribute to hallucination mitigation. Motivated by these findings, we propose AIMS (Adaptive Information Multi-source Steering), a lightweight training-free framework that adaptively coordinates visual, prefilled textual, and generated contexts during decoding. Specifically, AIMS constructs compact prototypes for the three context domains and estimates their affinities with the current query to determine head-wise steering weights. The resulting multi-source steering direction is applied to the query representation, enabling adaptive context integration without additional model training or auxiliary forward passes. Extensive experiments across multiple LVLMs and decoding strategies demonstrate that AIMS effectively mitigates object hallucination while maintaining competitive general-purpose multimodal capabilities.

cs.CV↗

Cost-Aware Mixture-of-Experts Coordination for Model Markets

Existing model marketplaces typically trade and select individual models as indivisible units, limiting their ability to exploit complementarities among heterogeneous experts. This paper proposes an MoE-based model market framework that lifts Mixture-of-Experts from a model-level learning architecture to a market-level coordination mechanism. In this framework, brokers use gating networks to coordinate multiple heterogeneous experts and deliver a composite model service. We formalize the market participants, service workflow, expert cost structure, and a welfare objective that combines predictive utility with heterogeneous execution costs. We then derive a cost-aware gating mechanism and market-aware training objective, and introduce a cost-adjusted revenue allocation rule that distributes residual revenue according to realized expert participation and execution cost. We also establish basic theoretical properties of the allocation rule, including budget balance, participation monotonicity, and cost sensitivity. Experiments over five random seeds on fifteen tabular and image benchmarks use independently trained and frozen neural and tree-based experts together with latency-derived execution costs. MoE Market achieves the highest mean welfare on all fifteen datasets and a lower mean expected cost than Standard MoE in every case, while maintaining competitive predictive performance. The allocation experiments further demonstrate systematic sensitivity to expert participation and cost, together with substantially lower computational overhead than exact Shapley allocation. These results suggest that MoE can serve as a market-level coordination principle for collaborative, cost-aware, and economically grounded model marketplaces.

cs.DB↗

The Fiedler dimension of networks of networks: from fractal to small-world architectures

Is the relaxation of a modular network dictated by its modules or by the network that connects them? The answer fixes the time scales of diffusion, consensus, and synchronization, all set by the Fiedler eigenvalue. We show that, for bundled networks, the two levels act one after the other: a random walker must first escape from its module and then spread over the network that connects them. The equilibration time is the sum of these two times, the time needed to reach the base from inside a fiber and the relaxation time of the base, slowed down by the mass of the fibers. The consequences are unexpected. However large the modules are, a small-world core always imposes its own Fiedler dimension; the modules survive only as a logarithmic correction, which slows down equilibration and hides the true exponent up to sizes far beyond any real network. Effective dimensions measured on finite modular systems can therefore be systematically biased. We analytically work out all combinations of finite-dimensional and small-world bases and fibers, and compare them with numerically exact spectra of representatives of each category.

cond-mat.stat-mech↗

Comprehensive study of massively overlapping cascades in common elemental metals

Massively overlapping cascades simulations were carried out in 21 elemental metals: Be, Al, Ti, V, Cr, Fe, Co, Ni, Cu, Zr, Nb, Mo, Rh, Pd, Ag, Hf, Ta, W, Pt, Au and Pb. These elements have simple FCC, BCC and HCP lattice structures. For each element, 2000 cumulative 5 keV cascades were simulated using molecular dynamics. The overlapping cascades were simulated using various classical analytical interatomic potentials as well as with machine-learning interatomic potentials for many of the elements. The general conclusion is that results from massively overlapping cascades simulations are very sensitive to the choice of interatomic potential across the periodic table. Furthermore, we see that FCC metals all form stacking fault tetrahedra due to irradiation, and also typically include interstitial type Shockley partial and Frank type dislocations. In BCC materials typically form interstitial 1/2$\langle$1 1 1$\rangle$ dislocations, but vacancy type 1/2$\langle$1 1 1$\rangle$ dislocations were also observed in V and Nb. HCP materials tend to form complex dislocation structures mainly consisting of a-type dislocations that can be of both vacancy or interstitial type. Correlations between saturated defect concentrations and fundamental underlying material properties are explored.

cond-mat.mtrl-sci↗

The Hawking and Penrose singularity theorems for low regularity Finsler spacetimes

We prove the Hawking and Penrose singularity theorems for Lorentz--Finsler structures of mixed horizontal/vertical regularity $C^{(1,2)}$ in the $C^1$ weighted case and $C^{(1,3)}$ in the unweighted case. At this regularity, geodesics need not be uniquely determined by their initial conditions, and the (weighted) Ricci curvature must be interpreted distributionally. Our approach relies on a two-stage approximation procedure: first, a homogeneity-preserving convolution, which is of independent interest in the broader semi-Riemann--Finsler setting, and second, a causality-adapted approximation à la Chruściel--Grant. In the special case of $C^1$ Lorentzian metrics, our results extend the $C^1$ singularity theorems of Graf to the $C^1$ weighted case.

math.DG↗