arXiv ScienceSearch

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 145 records · Page 8Linked to original sources

A Phase-Field Study of Desiccation Crack Pattern Maturation under Drying-Wetting Cycles

The characteristic intersection angle of the desiccation crack relaxes from near \ang{90} toward \ang{120} under repeated drying--wetting cycles. However, the theoretical understanding of this relaxation is insufficient, especially the modeling of the drying--wetting cycles. Here we introduce a phase-field model of desiccation fracture, extending the model proposed in previous studies by adding crack healing and a scar effect left by past cracks. By repeating drying--wetting cycles in a finite element simulation, we find that the angle distribution develops a growing peak near \ang{120} as the cycle number increases, consistent with experiments. The standard deviation of the intersection angle from \ang{120} relaxes exponentially with a characteristic time of about 2.85 cycles. These results are consistent with experiments, except that the characteristic time is slightly smaller than the experimental value. Crack energy dominates the total energy and also relaxes exponentially with nearly the same characteristic cycle as the angle relaxation. This decay is driven mainly by a shortening of the effective crack length rather than a change in effective fracture toughness.

cond-mat.soft

Computational Trade-Offs Between Newton and Picard Solvers for Mean Field Game PDE Systems

We study computational trade-offs between two solvers for the same semi-implicit finite-difference discretization of forward-backward partial differential equation (PDE) systems arising in mean field games (MFGs). The Picard method uses an outer fixed-point iteration that alternates a forward Fokker-Planck solve and a backward Hamilton-Jacobi-Bellman solve. The Newton method instead applies Newton's method directly to the coupled nonlinear space-time system. Across one- and two-dimensional MFG benchmarks, we find that the Picard method has much lower wall-clock cost when it converges, but may fail at sufficiently low viscosity and may require strong damping under temporal shocks. With parameter continuation, the Newton method is more robust in these regimes, at the cost of larger coupled linear systems. We relate these trade-offs to the residuals, Jacobian blocks, and sparsity structures produced by separable, local nonseparable, and nonlocal Hamiltonians. We also test a hybrid continuation strategy on a two-dimensional double-well MFG benchmark, using Picard iterations for the inexpensive part of the viscosity descent before switching to Newton continuation.

math.NA

Tensor Decomposition Based Mixed-Field Sensing for XL-MIMO AFDM Systems

Integrated sensing and communications enabled by extremely large-scale MIMO (XL-MIMO) and affine frequency division multiplexing (AFDM) is a highly promising paradigm for vehicular networks. However, the near-field spherical wavefront distortions induce severe non-linear parameter coupling, while the highly dynamic scattering environments exacerbate mismatch errors. To address these critical challenges, this paper proposes a novel tensor-based sensing scheme for XL-MIMO AFDM systems. First, the received signals are reformulated into a tensor, followed by an efficient decomposition approach that exploits the inherent Vandermonde structure of the factor matrices. This allows parameters to be directly estimated from the decomposed matrices, effectively avoiding inter-parameter coupling. Subsequently, a symmetric decoupling and real-domain manifold optimization algorithm is proposed for angle of arrival estimation, circumventing the high-dimensional searches typically induced by near-field effects. Furthermore, a baseband reconstruction and analytical gradient-based algorithm is developed to perform delay-Doppler estimation in the continuous parameter domain, fundamentally eradicating the grid-mismatch errors inherent in high-mobility scenarios. With these decoupled factors, the remaining unknown angle of departure can be readily extracted. Extensive simulation results demonstrate that the proposed scheme achieves orders-of-magnitude improvements in delay-Doppler accuracy and eliminates the error floors in angular estimation that severely bottleneck state-of-the-art baselines.

eess.SP

UniRec: Cross-stage Multi-Task Fusion with Preference Alignment for Cascaded Recommender Systems

Industrial recommender systems use cascaded stages with different objectives, feature spaces, and latency constraints. Optimizing pre-ranking and ranking separately can create cross-stage inconsistency: upstream models may filter out items preferred by downstream rankers, and independently tuned downstream fusion can offset upstream improvements. Existing multi-task fusion methods focus on multi-objective fusion within the ranking stage, and cross-stage methods typically only add a downstream score factor to upstream ranking. Joint optimization of fusion modules across both stages remains largely unexplored. We propose UniRec, a Unified Cross-stage Recommendation Fusion model. First, the two fusion agents partially share input embeddings and are trained in a single computation graph, so gradients from either stage propagate through the shared representation and influence the other. Second, we introduce a dual-axis preference alignment objective: a vertical cross-stage consistency term transfers downstream pairwise preferences to the upstream fusion score, and a horizontal compact aggregation term reorganizes dozens of pairwise objectives over heterogeneous prior signals into bidirectional preference evidence. Third, we find that unconstrained end-to-end fusion optimization can exploit imbalances in item attribute distributions, over-concentrating on high-reward regions at the cost of other objectives. We therefore add an attribute group-relative regularization that computes advantages within attribute groups and normalizes the policy over the same groups, so uniformly promoting an entire high-reward group yields no optimization gain. Offline, UniRec consistently outperforms single-stage fusion and cross-stage coordination baselines. Online A/B tests show a 0.616\% gain in app usage duration. UniRec is fully deployed on the Kuaishou platform.

cs.IR

Uncovering the Origin of Slow Rotators among Intermediate-Mass Stars in the Star-Forming Cluster Trumpler 14

Intermediate-mass (about 1.5-8 M$_{\odot}$) stars exhibit a wide range of rotation rates and have gained attention for their roles in the extended main-sequence turnoffs (eMSTO) seen in some young and intermediate-age clusters (ages $< 2$ Gyr). Although rapid rotation is expected due to their radiative envelopes, the presence of slow rotators remains puzzling. In this study, we examine disk and X-ray signatures among the intermediate-mass members of the star-forming cluster Trumpler 14. Of the 118 intermediate-mass members, 24 are considered disk-bearing on the basis of their infrared spectral indices derived from spectral energy distributions. The majority of these reside outside the heavily irradiated cluster core and are all slow rotators ($\lessapprox 100$ km s$^{-1}$) except for three fast-rotating Class III objects. Additionally, in the same sample of 118, 37 are X-ray sources, and 22 having available $v \sin i$ data show a systematically slower rotation than X-ray quiet sources. Our findings suggest that, while young stellar disks in intermediate-mass stars contribute to early spin-down, high-energy processes traced by X-ray emission during this phase could be an additional channel for angular-momentum loss in these stars.

astro-ph.SR

There is no maximal $K$-degree

The Kolmogorov complexity of a string characterize how complex it is to describe the string. If every prefix of a real $x$ is more complex to describe than every prefix (of the same length) of real $y$, then it is seen as $x$ is more complex to describe than $y$. It is wondered if there is a real $x$ so that no other reals are strictly more complex (to describe) than $x$. The behavior of Kolmogorov complexity functions generated by reals (namely $n\mapsto$ the minimal description length of the real) is quite chaos. Therefore, it is widely believed that there are many reals that are maximally complex to describe. For instance, it is conjectured that all random enough reals have maximal $K$-degree. In this paper, it is shown that there is no real with maximal $K$-degree. Actually, for almost all real $x$, we can uniformly computably find another real whose $K$-degree is strictly above $x$.

math.LO

A Tale of Two Gauges: Effective Field Theory for Relativistic Behavior of Cosmological Axions

In this work, we present a formalism to model the relativistic behavior of axions. The relativistic behavior of axions is surprisingly difficult to model precisely, as it involves oscillations on timescales much shorter than the Hubble timescale. To overcome this challenge, one typically resorts to some form of effective treatment, focusing only on the time-averaged description of the exact oscillations. Salehian, Namjoo & Kaiser provide a systematic framework for such treatment, based on the effective field theory formalism. While the aforementioned study was formulated for axion perturbations in the Newtonian gauge with no anisotropic stress, we extend the formalism to the synchronous gauge that is more conventionally used for numerical implementation in a realistic cosmological setting. Unlike their work, however, we propose a fluid interpretation in which the axion field can be identified as a perfect fluid at all times, both in the exact and effective regimes. Moreover, we present the effective field theory for the Newtonian gauge with non-zero anisotropic stress, making the original formulation more general and useful for scenarios where the matter content of the universe is multi-component. These results lay the theoretical foundation for a companion paper where we discuss how the axion field should be incorporated alongside other species in common cosmological Boltzmann solvers.

astro-ph.CO

(Re)constructing Accurate Axion Oscillations

The cosmological evolution of ultralight axions typically involves rapid oscillations on the timescale of the inverse mass, making it challenging to resolve the dynamics at both the background and perturbation levels. In this study, we numerically implement a novel approach to this problem based on an effective field theory (EFT) -- that is, for the first time, capable of reconstructing the relativistic oscillations intrinsic to the axion field. Compared to other available techniques, the reconstruction of the rapid oscillations is unique to the EFT approach, which describes the axion effective field through a wavefunction representation with relativistic corrections, rather than through effective fluid variables. From a computational standpoint, our approach archives a high level of accuracy -- up to a subpercent agreement with the exact solution -- while remaining relatively fast compared to alternative methods. As such, it opens up new possibilities for high-precision predictions for axion searches with future cosmological experiments.

astro-ph.CO

Search for the process $e^+e^-\to f_1(1285)$ at the SND detector

In the experiment with the SND detector at the VEPP-2000 $e^+e^-$ collider, a search is performed for the direct production of the $C$-even $f_1(1285)$ resonance in $e^+e^-$ collisions. The analysis is based on data with an integrated luminosity of about 200 pb$^{-1}$, accumulated in the center-of-mass energy range of 1.14--1.46 GeV, of which about 72 pb$^{-1}$ were recorded near the maximum of the $f_1(1285)$ resonance. The $f_1(1285)$ production cross section at the resonance maximum $\sigma(e^+e^-\to f_1)=(31\pm 13\pm 2)$ pb and the branching fraction $B(f_1(1285)\to e^+e^-)=(3.5\pm 1.4\pm 0.3)\times 10^{-9}$ have been measured. The significance of the observation of the $e^+e^-\to f_1(1285)$ process is $2.5\sigma$. Since the significance is low, we also present the upper limits at the 90% confidence level: $\sigma(e^+e^-\to f_1)<48\mbox{ pb}$ and $B(f_1(1285)\to e^+e^-)<5.4\times 10^{-9}$.

hep-ex

EMMI: Edge Multi-Modal Intelligence for Communication-Efficient MLLM Inference via Fused Representation Compression

Recent advances in multimodal large language mod- els (MLLMs) have opened new opportunities for edge intelligence by enabling reasoning across heterogeneous sensor modalities, such as vision, text, and telemetry data. However, deploying these capabilities on resource-constrained edge platforms remains challenging due to the substantial computational, memory, and communication demands of modern MLLMs. Rather than transmitting raw sensor observations or partitioning neural networks at intermediate layers, Edge Multi-Modal Intelligence (EMMI) communicates a compact representation between edge devices and server resources, enabling communication-efficient edge MLLM inference. To achieve this, EMMI performs modality-specific encoding, cross-modal representation fusion, and learned compression at the edge, transmitting only a compact latent representation to server-side resources for high-capacity MLLM reasoning. This representation-centric design reduces communication overhead, preserves local data privacy, and provides a fixed-size interface between heterogeneous edge devices and server-side MLLMs. Evaluation on a representative multimodal benchmark demonstrates that EMMI can reduce the communication payload by 32x while maintaining comparable downstream accuracy, resulting in up to a 3.4x reduction in estimated end-to-end inference latency under bandwidth-constrained edge conditions.

cs.LG

Gait-Dependent Effects on Quadruped Locomotion for Load-Carrying using Passive Mechanism

Passive mechanical interfaces offer a lightweight alternative to actuated manipulators for quadruped payload carrying, but their impedance directly couples the payload dynamics with the locomotion pattern. This paper analyzes how passive-arm stiffness-damping selection affects payload-carrying locomotion under different gait and payload conditions. We compare damped and underdamped passive-arm impedance configurations in simulation during flat-ground locomotion. For crawl gaits, where the support polygon remains well defined, the results show that underdamped impedance increases passive-joint oscillations and can reduce the ZMP margin with respect to the support polygon. Trot is retained as a dynamic excitation case for the passive arm, but it is not used for direct ZMP-margin stability comparison. The results are summarized in gait-payload-stiffness-damping maps, where ZMP-margin reduction is evaluated for crawl gaits and trot is retained only as a passive-arm excitation case.

cs.RO

Grounding Agent Memory: Environment-Probing Curation for Enterprise Agents

Persistent memory is entering production-oriented agent platforms to help long-horizon agents accumulate experience across sessions. Yet a post-task curator agent restricted to completed trajectories can preserve errors, overgeneralize partial evidence, or retain stale knowledge. We introduce environment-probing curation, a deployment-compatible extension that gives an existing asynchronous curator agent least-privilege, read-only world tools to check, scope, and refresh candidate memories. It requires no model retraining and leaves the task agent, retriever, memory representation, and production write authority unchanged. In a production-like GitHub Copilot (GHCP) harness built on its SDK, we compare stateless execution, full in-context learning, GHCP + Mem, and GHCP + Mem (w/ Env Probing) on CLBench database exploration and 90 adapted APEX management-consulting tasks. On CLBench, probing raises pass rate from 39% to 73% and pass-discounted reward from 8.60 to 22.60 while reducing queries from 8.8 to 4.7 per question and task-agent cost from \$3.38 to \$1.68. Across six APEX worlds, all 18 memory-versus-baseline mean reward comparisons are positive and task-agent tool calls fall by 16--75%; probing gives the best task-agent reward gain per dollar in five worlds. Probing also attains higher mean reward than GHCP + Mem on both Sonnet 4.6 and Opus 4.7 without schema drift. Environment probing therefore turns existing agent-memory curation into an environment-informed, auditable process while preserving a compact task-time interface.

cs.AI

Fork Where the Model Changes Its Mind: Belief-Shift Branching for Tree-Structured Reinforcement Learning

Tree-structured rollouts give critic-free reinforcement learning with verifiable rewards (RLVR) step-level credit: fork a chain at an intermediate point, and sibling outcome differences estimate step value. Each fork adds sampling cost, so realistic budgets typically allow only a few forks per chain. A fork placed where the outcome is already largely settled yields siblings that mostly agree and provide almost no credit signal; hence, for a given tree size, where forks are placed largely determines how much step-level RL can gain. Most existing mainstream methods place forks by structure, such as fixed lengths, midpoints, and delimiters, or by next-token entropy. We formalize fork placement as locating the \emph{pivots} of the chain's value curve, where the expected outcome turns. We propose \emph{belief-shift branching}: read the model's answer belief at candidate boundaries and fork just before the step where consecutive beliefs diverge most. Three instantiations, none needing step-level supervision, span access levels: a black-box probe, a logit-lens depth profile, and a learned activation direction, which is fit offline and therefore used only in the validation before RL training. The signal only \emph{places} forks, and the probe costs about $1\%$ of step compute on mathematics and under $5\%$ on code when it runs inside the rollout engine. In that validation, against Monte-Carlo value curves, a belief-shift signal ranks first in each of the eight model$\times$benchmark panels, ahead of entropy, structural, and LLM-judge baselines. In RL across three model families and two domains, belief-shift forking leads every mathematics aggregate, on OLMo-3-7B by $+2.6$ aggregate and $+2.9$ on AIME 2026 over the strongest baseline, and sweeps every OLMo code column, by $+6.5$ on LiveCodeBench-medium.

cs.AI

Fusion Estimation in Multi-sensor Systems for Data Packets with Disrupted Identities

In this paper, we explore the problem of fusion estimation for a multi-sensor system where the identity of the data packet received by each sensor may be disrupted or incorrect due to confusion in device identity allocation, communication protocol defects, or the lack of a clear sensor identifier. This can result in a random shuffle of the data components during the fusion estimation process, compromising the performance of the fusion estimation. To address this issue, we introduce the concepts of permutations and symmetry groups to describe this phenomenon as data packet permutation. We construct statistics to simplify the information set, developing two algorithms: a Bayesian approach, which performs fusion using posterior arrangement probabilities, and a greedy approach, which effectively improves estimation performance by guessing the likely data arrangement. We compare these two algorithms and demonstrate that both are expectation error-bounded. We improve algorithms for information-scarce scenarios. By employing the expectation-maximization algorithm, we fill in the prior information of data arrangement where the correct convergence is proven. Finally, we present numerical simulations to validate our results.

eess.SY

The information geometry of large language models is shared, learned, and controllable

Large language models learn similar behaviours, yet it remains unclear what structure they share or how to change one behaviour without disturbing others. The Fisher-Rao geometry of next-token probabilities connects these questions: behaviour determines this geometry up to output-preserving symmetries, whereas activation geometry depends on coordinates. Across transformer, state-space and recurrent models, output geometries agree more strongly than activation geometries, and shared geometry supports semantic-category transfer. Agreement with human word choices increases with predictive accuracy, scale and training, and improves further after model-only calibration. Token probabilities and read-out geometry jointly predict the spectrum and its effective dimension. Controlled language assignments show that geometry follows the language law across architectures. Pretraining corpus statistics predict held-out fact acquisition without recalibration, while randomised experiments show that deeper evidence substantially delays acquisition across every tested architecture and evidence construction. Finally, the geometry prescribes minimum-disturbance local interventions, predicts their relative cost, and supports reusable control: updates learned on donor prompts transfer to unseen prompts while better preserving behaviour on reference prompts than Euclidean control. The same geometric correction improves steering, editing, attribution, dictionary learning and fine-tuning.

cs.LG

Polynomial preserving recoveries of edge element method on Cartesian grids for the time-harmonic Maxwell equations with large wave number

This paper considers the lowest-order first type N\'{e}d\'{e}lec edge element method (EEM) on Cartesian grids for the three-dimensional time-harmonic Maxwell equations with a large wave number. New polynomial preserving recovery (PPR) operators are proposed for the curl of the edge element solution and for the solution itself, respectively. Under the condition that $\kappa^3 h^2 C_{\mathrm{sol}}$ is sufficiently small, second-order superconvergence estimates are proved for both the recovered curl and the recovered solution, where $\kappa$ is the wave number, $h$ is the mesh size, and $C_{\mathrm{sol}}$ is a stability constant associated with the Maxwell solution operator. In particular, the analysis shows that the proposed PPR procedures cannot mitigate the well-known pollution effect inherent to the EEM. To reduce the pollution error, we further propose a new continuous interior penalty edge element method (CIP-EEM) that incorporates an additional normal-jump penalty term. It is shown that by appropriately choosing the penalty parameters, the new CIP-EEM can improve the phase error by two orders in $\kappa h$. Numerical experiments are presented to confirm the theoretical superconvergence results and to demonstrate that the CIP-EEM can effectively reduce the pollution error in the high-frequency regime.

math.NA

MOSAIC: Query-Aware Exploration Policy Adaptation for GraphRAG

Graph Retrieval-Augmented Generation (GraphRAG) can connect evidence distributed across a corpus graph, but most systems use largely shared exploration procedures across queries. This creates a structural mismatch: direct facts may need compact local neighborhoods, comparisons need balanced coverage of multiple targets, and mediated questions may require deeper paths through weakly related connectors. We present Mosaic, a training-free framework that formulates GraphRAG retrieval as a per-query control problem. An LLM analyzer converts query-specific evidence requirements into a bounded policy over seed selection, graph traversal, stopping, and evidence selection, while the corpus graph, indexes, scoring functions, grounding procedure, and answer generator remain shared. On GraphRAG-Bench, Mosaic achieves query-weighted Answer Correctness of 76.97 on Medical and 64.33 on Novel, improving over the strongest previously reported overall results by 5.13 and 4.43 points. On Medical, it reaches 95.1 Evidence Recall and 86.1 Context Relevancy. Controlled comparisons on an identical graph and generator show that no fixed narrow, medium, or wide policy is consistently optimal; Mosaic improves by 9.96 points over the strongest canonical fixed policy. Relative to Fixed Wide, it evaluates 81.9% fewer paths and retains 47.2% fewer evidence items. Transfer experiments on HotpotQA, MuSiQue, and 2WikiMultiHopQA further show that the policy interface can be applied without benchmark-specific retriever training.

cs.AI