arXiv Science⌕ Search

arXiv subjects

Search papers

Search indexed arXiv papers on artificial intelligence, large language models, computer vision and robotics. Read source abstracts and follow links to arXiv.

At least 1,243 records · Page 69Linked to original sources

Evidence for Distributed Fault Energetics and Their Impact on Deformation in a Chemically Complex Alloy

Chemically complex alloys feature intrinsically heterogeneous local chemical environments and, consequently, fluctuations in local fault energetics. However, experimentally quantifying their relationship remains challenging, leaving the role of this distributed energy landscape in deformation mechanisms incompletely resolved. Here, we develop a distribution based framework linking experimentally measured stacking fault widths to deformation relevant apparent fault energy, revealing a distributed local fault-energy landscape in CrCoNi. The framework reveals the stabilizing role of energy fluctuations and captures an upward shift in the apparent fault energetics, from negative values toward zero following heat treatment, which atomistic simulations associate with the emergence of L12 type chemical short-range order. Using one-dimensional kinetic Monte Carlo simulations, supported by electron microscopy observations, we further show that history-dependent changes in the local fault-energy landscape bias the competition among stacking faulting, nano-twinning and HCP transformation in CrCoNi. Our results provide an experimentally anchored, distribution-based framework for understanding deformation in chemically complex alloys as it evolves within a distributed fault-energy landscape shaped by local chemical order.

cond-mat.mtrl-sci↗

Aperture: Merge-Consistent Rotary States for Compressed Tokens

Token compression combines content from several positions, yet rotary position embeddings usually assign the merged token one coordinate. We ask what positional information must survive later merges. Aperture stores Fourier moments of the token's weighted support at the model's rotary frequencies. We prove that these moments have minimal real dimension among continuous states sufficient for the selected expected rotary interactions. Represented mass makes updates additive; attention normalisation remains a separate readout choice. Uniform intervals give a centre rotation times a sinc gain. We characterise when centres determine interval widths and construct matched examples where they do not. Numerical checks verify the weighted-support implementation. In trained temporal readers, compression transfer varies with gain calibration and feature placement. In a prespecified native video question-answering comparison, stored support reaches $65.63\%$ accuracy versus $67.12\%$ for the deployed merging rule. These results separate exact positional preservation under compression from downstream benefit.

cs.AI↗

Decoding Affective Nuances: Enhancing MLLMs via Hierarchical Emotion Reasoning and Contrastive Discriminative Pruning

While multimodal large language models (MLLMs) have demonstrated exceptional capabilities in objective understanding tasks, their performance in affective reasoning still falls significantly short of human standards. We attribute it to a central capability gap: MLLMs are difficult to reliably distinguish semantically proximal emotions based on fine-grained visual evidence, which could be decoupled as two limitations: 1) Insufficient Attribution. The global reasoning paradigm of conventional MLLMs severely dilutes fine-grained emotion cues, where subtle emotional states are usually implicitly encoded, thereby generating emotional misjudgments in complex scenarios. 2) Insufficient Discrimination. Existing methods could only identify regions generally associated with emotions, which fails to distinguish discriminative regions between semantically similar emotions, leading to ambiguous emotion judgements. To overcome these limitations, we present a training-free inference-time optimization framework, named Decoding Affective Nuances (DAN). Specifically, we propose a Hierarchical Emotional Reasoning Chain (HERC) that enhances the insufficient attribution by harmonizing fine-grained scene/object-level cues and performing a soft-gated reasoning. Furthermore, to discriminate between semantically proximal emotions, we design a Contrastive Discriminative Visual Pruning (CDVP), which isolates discriminative visual tokens to reason the final emotion category by computing the absolute discrepancy between the attention distributions of similar emotions. Performances on several benchmarks demonstrate that DAN significantly improves discrimination for affective nuances without consuming additional training resources, especially achieving +10.47% improvements with Qwen3-VL-8B-Instruct on WebEmo25 dataset that contains 25 fine-grained emotion categories.

cs.CV↗

The Editor Has Read-Only Access: Correctness Signals in Diffusion Language Models

Diffusion language models generate code by repeatedly updating a partially masked sequence. We ask whether their internal activations encode code correctness and whether that information can improve generation. Across six diffusion models, linear probes distinguish passing from failing attempts, with the strongest reads generally appearing beyond the early layers. Controls using small semantic mutations support a connection to correctness rather than surface style alone. In comparisons with model confidence, probe point estimates offer no consistent advantage. Adding a probe-derived direction to the residual stream does not yield a dependable improvement in the tested steering settings, while the opposite direction degrades performance. We distinguish these observations from claims about statistical significance or a general inability to steer. Supplementary methods, archived results, and code document the tested interventions and the limits of their statistical calibration and reproducibility.

cs.SE↗

Scale-Invariant Manipulability Shape Tracking Across Heterogeneous Manipulators

When transferring manipulability across systems with different sizes and kinematic structures, matching absolute ellipsoid scale may be unnecessary when the goal is to reproduce orientation and semi-axis length ratios. Full-matrix tracking, however, penalizes both shape and absolute-scale differences, even when only shape matching is required. We therefore propose a scale-invariant manipulability shape-tracking method that treats matrices differing only by a positive scalar factor as equivalent and uses their unit-determinant representatives. We derive the differential of the unit-determinant shape representative and an orthonormal coordinate representation of the tangent tracking residual under the affine-invariant Riemannian metric (AIRM). The resulting scale-invariant objective is integrated with position and end-effector direction tasks in a constrained joint-velocity quadratic program. Simulations with four heterogeneous robots evaluate robot-to-robot and human-to-robot transfer. On three followers, the proposed method achieves endpoint shape distances of 9.30 x 10^-5 without scale tuning. With robot-specific target scales tuned during motion, the Full method retains endpoint axis-ratio errors of 0.19-0.31 on KR500 and UR20. For human reaching with concurrent tasks, the proposed method yields dual force shapes elongated along X like the human target on all four robots, with endpoint position errors of 2.4-5.6% of reference arm length versus up to 75% for the Full method tracking the original human ellipsoid.

cs.RO↗

TaRL: Learning General and Physical Rewards from Tactile Demonstrations

Contact-rich manipulation requires robots to sequence precise contacts, maintain stable grasps, and apply directed forces. Reinforcement learning (RL) can acquire such behaviors automatically, but its performance hinges on reward design: sparse rewards reduce the learning efficiency, while dense rewards are hard to specify. Visual reward learning addresses this by inferring rewards from action-free demonstrations. Because it conditions only on visual observations, it fails to capture rewards beyond visual goals. We propose Tactile Reward Learning (TaRL), a framework that learns rewards from tactile demonstrations. TaRL takes a sequence of tactile deformation maps as input, and regresses task-completion progress from both successful and failed demonstrations. Because TaRL captures local robot-object interaction, it provides informative feedback to learn firm grasps and correctly directed forces; meanwhile, it is robust to changes in scene layout such as object position. We evaluate TaRL on four manipulation tasks in simulation and two in the real world. Used as a shaping reward, it substantially improves both sample efficiency and final success rate, raising success on Nut threading from 34% to 56% in simulation and on cube pickup from 37% to 97% in the real world. Combining tactile with visual rewards improves performance further. TaRL also generalizes across object instances: trained on box placement and directly deployed to can placement, it significantly improves policy learning on the new task. Project page is available at https://embodiedai-ntu.github.io/tarl.

cs.RO↗

Testing Diagonal Anisotropy Based on Single Realisation of Spherical Random Field

Spherical data appear in various applications, including cosmology and the Earth sciences. A standard statistical model for such data employs isotropic spherical random fields. The isotropy assumption may be unrealistic for many real-world datasets. Also, in various applications such as cosmology, only a single realisation of the field is observed, making isotropy impossible to assess through repeated sampling. This paper considers an alternative anisotropic model, referred to as diagonal anisotropy. To test isotropy against diagonal anisotropy, we develop score and cumulative-sum-type tests. Their asymptotic properties and Monte Carlo-based alternatives suitable for moderate sample sizes are established. The performance of the proposed methods is illustrated by numerical studies via applications to simulated data and to actual Planck cosmic microwave background radiation observations. The tests are also applied to three Planck SMICA cosmic microwave background radiation maps and demonstrate how results can vary for different data releases and map construction procedures.

stat.ME↗

Regularized policy gradient with learned mixtures of Gaussians for games with continuous actions

Most successes of superhuman game-playing algorithms are in games with discrete actions, yet in auctions, robotics, sports, or trading, actions are nearly continuous. Prior techniques either rely on expert-designed discretizations or are sample inefficient. We present a scalable policy-gradient algorithm for large sequential games with continuous or mixed discrete and continuous actions. It combines magnetic mirror descent with a mixture of Gaussians reparametrization, trained via self-play. We show that it approximates equilibrium in games where gradient descent fails. In sequential games, it outperforms neural fictitious self-play and matches or outperforms the final strategies of policy space response oracles with 3.5--5.5$\times$ fewer samples. In heads-up no-limit Texas hold'em, it performs on par with Slumbot.

cs.MA↗

Harnessing Large Language Models to Compile Task-Relevant Context into Bayesian Optimisation

Incorporating rich task-relevant context, such as domain knowledge and external observations, is a key capability yet remains challenging for Bayesian optimisation (BO). Recently, practitioners have started to use large language models (LLMs) to generate and execute BO programs through coding harnesses. In such emerging practices, the posterior belief is shaped not only by Bayesian inference but also by LLM-generated model and data artefacts, offering a flexible route for task context to enter BO as executable code. To study whether and how LLMs can be harnessed to compile diverse contextual signals for BO, we formulate LLM-compiled BO as generalised-context decision making. We propose HarBO, a BO-specialised harness that compiles generalised context into the core artefacts of standard BO through a validated multi-stage workflow. Our theory analyses the regret under imperfect compilation and the effect of adding new context. Across synthetic functions and real-world benchmarks, we find that LLM harnesses can effectively compile context into standard BO, achieving competitive performance with specialised LLM-embedding-based and direct LLM-in-the-loop BO methods. General coding harnesses can be effective in familiar domains such as hyperparameter optimisation, but fall short in unfamiliar, context-rich domains. Together, these results establish LLM harnesses as a promising, but not automatically reliable, route for making rich task context usable in BO.

cs.LG↗

GitHarness: Git Init Your Harness Working Memory for Perpetual User Requirements

LLM-based agents increasingly collaborate with users on long-horizon tasks, accumulating evidence, code, and drafts through extensive search, reasoning, and execution. As users inspect these results, they may supply missing information requirement completion, introduce new requirements requirement elicitation, or revise existing ones requirement shift. These changes often affect only part of the accumulated work, yet agents may carry forward obsolete information or turn local revisions into global rewrites. Existing approaches clarify current intent without determining how prior work should change, or reuse execution histories under a fixed objective. We address this gap by formulating dynamic-requirement collaboration as joint requirement tracking and local update. We introduce GitHarness, a pluggable Git-style framework that organizes requirement states and their corresponding harness work states into a branchable version history. A trainable Git Agent resolves requirement changes and selects a semantically compatible historical state. A unified version interface then restores that state and creates a new branch, enabling the underlying harness to exclude obsolete information, inherit compatible work, and focus execution on affected parts. The Git Agent is trained through interface-level black-box reinforcement learning, with downstream harnesses and task-execution models kept fixed. We also construct MTAgentBench, a verifier-preserving benchmark covering mathematical reasoning, text-to-SQL, agentic search, software engineering, and research synthesis. Experiments demonstrate strong task performance alongside effective requirement tracking, preservation of valid work, and efficient execution.

cs.MA↗

TDOA-Based Online Target Tracking with Simultaneous Sensor Pairing and Relocation

We propose a target tracking method for mobile sensor networks (MSNs) based on time difference-of-arrival (TDOA). Target tracking using an MSN requires reorganizing sensor pairs and relocating sensor positions to obtain TDOAs with higher quality. Existing studies only consider either one while fixing the other, which may limit the tracking accuracy. We address this limitation by simultaneously solving sensor pairing and relocation. We formulate the problem as a maximization of the determinant of the Fisher information matrix, and decompose the problem into two subproblems to solve them alternately. Sensor pairing is solved by a mixed-integer second-order cone program algorithm, and sensor relocation is solved by majorization-minimization. Experimental results show that the proposed method obtains lower tracking error than existing designs with a practical computation time.

eess.SP↗

Influence Ranking Improvement via Link Addition in Social Networks

Social media platforms increasingly rely on influential users for information dissemination in domains such as marketing and political campaigns. As the value of being recognized as influential grows, users may have incentives to strategically enhance their influence. In this study, we investigate whether and to what extent the influence ranking of a target node can be improved by adding a limited number of outgoing links in the context of influence maximization (IM). We formulate the Link Selection Problem for Influence Ranking Improvement (LSP-IRI) as the problem of selecting a fixed-size set of additional outgoing links from a target node in order to maximize its rank improvement. To examine this problem, we consider two representative heuristic strategies: a greedy method that directly optimizes rank improvement and a computationally efficient random search method. We conduct experiments on four real-world networks ranging from thousands to hundreds of thousands of nodes. The results show that even a few added links can substantially improve IM-based influence rankings. In particular, the greedy method improves the ranks of nodes initially ranked around 50 by several tens of positions on average with only three added links, sometimes moving them into the top 10. The random search method achieves smaller gains under strict budgets but reduces computation time by up to approximately 98% compared with the greedy method and becomes effective when larger budgets are allowed. These findings show that IM-based influence rankings are sensitive to limited local structural modifications and highlight a trade-off between ranking improvement and computational efficiency.

cs.SI↗

Analysis of Delay Differential Equations Using the Offset Linear Canonical Transform

Motivated by the work of Ohira and the advantages of the offset linear canonical transform (OLCT) over the Fourier transform (FT), this paper proposes an OLCT based framework for solving a class of delay differential equations. By exploiting the operational properties of the OLCT, the original delay differ- ential equation is transformed into a Volterra-type delay integral equation in the transform domain and solved numerically using Brunners method of steps . An explicit analytical solution is derived for a special case to investigate the effect of the delay parameter. A unified transform-domain formulation is also established by relating the proposed OLCT approach to its Fourier transform counterpart. Numerical and graphical results demonstrate the accuracy and effectiveness of the proposed method and validate it through comparisons with the Fourier transform based formulation of Ohira . The proposed framework may provide a promising foundation for solving non linear delay differential equations involving transcendental terms, as both the OLCT and the resulting Volterra integral equation possess the essential properties required to handle non linearities and transcendental terms.

math-ph↗

Hard-Trace First-Order Residual Learning for Coupled Stokes-Brinkman-Darcy Flow

Stokes-Brinkman-Darcy (SBD) models connect free flow and porous flow through a Brinkman layer of finite thickness. The proposed method for solving these coupled equations consists of a mixed first-order formulation, shared interface traces and hard traction constraints, and a pressure correction based on boundary mass balance. Independent stresses, Darcy flux and pressure-gradient auxiliaries avoid second derivatives of network outputs. Boundary liftings and shared traces impose exterior and interface kinematic conditions exactly, while a stress map enforces native Brinkman-Darcy force balance. An integral momentum relation determines the pressure correction after training, without reference-pressure labels. Across two manufactured solutions and three permeabilities, mean corrected upper-pressure relative $L^2$ errors range from 0.87% to 2.58%; correction reduces both upper-pressure $L^2$ errors and the full Brinkman-pressure $H^1$ error in every run. Paired controls assess the contributions of interface sharing and traction enforcement. A non-manufactured filtration problem tests mixed-boundary adaptation against an independently refined finite-element reference, with all four runs satisfying the prescribed field and physical-consistency criteria.

math.NA↗

RAE-PPG: Duration-Grounded Retain-and-Extend Pretraining for PPG Foundation Models

Signal features derived from photoplethysmography (PPG) require different signal durations to characterize. Existing PPG foundation models treat duration as a pretraining or evaluation condition rather than using the different durations required by PPG features to organize self-supervision. We hypothesize that self-supervision should expand with signal duration, allowing a single encoder to progressively acquire additional features while preserving and reusing earlier learning. We introduce Retain-and-Extend PPG (RAE-PPG), which trains a single Transformer encoder successively on 10 s, 30 s, and 240 s inputs, adding supervision for signal features supported by each longer observation. The encoder is partitioned into duration-specific parameter groups, allowing later stages to reuse earlier groups while updating only the group assigned to the current stage. Selected earlier targets are reused to supervise later stages, encouraging the corresponding features to remain accessible in longer-input representations. Direct decoding from the final encoder shows that earlier features remain recoverable from longer-input representations, while later-stage features show higher mean decoding performance at their introduction durations. Controlled comparisons further show that prior-stage learning provides a better basis for learning newly introduced features at both transitions. Across 18 tasks from eight datasets, the final frozen encoder achieves the best observed score on 12 tasks compared with five existing PPG foundation models.

cs.LG↗

Epistemic Typing as a PostgreSQL Table Access Method: Adversarial Conflict Resolution Under Confidence Forgery and Sybil Coordination

We describe KNDB, a PostgreSQL 18 table access method (TAM) that types every row with an engine-assigned epistemic kind (MEASURED, INFERRED, or DERIVED) and resolves per-slot conflicts inside every write-time heapam callback. Rows land as ordinary heap tuples; seven of the 44 TAM callbacks are overridden (tuple_insert, multi_insert, tuple_update, tuple_delete, tuple_insert_speculative, tuple_complete_speculative, relation_toast_am), the other 37 delegate to heap; we provide a completeness argument over the interface as a paper artefact. This paper reports the engineering behind that decision and the adversarial evaluation that motivated it. On a confidence-forgery workload where an attacker asserts INFERRED writes with confidence in [0.95,1.0] against honest MEASURED writes with confidence in [0.5,0.9], KNDB beats a confidence-only baseline by 63 percentage points on the Book-Author fusion dataset and 92.7 points on the Zheng crowdsourcing dataset. Both wins are proven load-bearing on the kind axis by a source-rebuild disable-and-test in which the lattice is neutralised and the win vanishes. Against four truth-discovery baselines (TruthFinder, CRH, CATD, ACCU) reimplemented from the original equations and validated to within 0.3 percentage points of the published numbers, KNDB is competitive below a per-dataset density-saturation cell and dominant at or above it. We formalise the cell as k* ~ rho_alg * h_top, where h_top is per-slot top honest surface-form support, and validate the prediction within +/-20% on Book-Author and +/-30% on Zheng. Because the kind axis is assigned by the engine from independent metadata and cannot be forged at write time, KNDB's k* is unbounded. The paper is honest about where KNDB loses: CRH and ACCU outperform KNDB below saturation on Zheng, and KNDB scores zero on three temporal knowledge-editing benchmarks whose ground truth is last-writer-wins.

cs.DB↗

Dynamical friction vs. subhalo heating in Cold Dark Matter haloes

The orbital evolution of massive stellar systems embedded in dark matter haloes is governed by competing processes. Dynamical friction removes orbital energy, whereas fluctuations in the gravitational field generated by dark matter subhaloes inject energy through stochastic heating. We investigate the balance between these two mechanisms using analytical arguments and numerical experiments. We show that, for a given host halo and subhalo population, there exists a critical mass $M_{\rm crit}$ at which heating and friction balance. Clusters with masses $M_{\rm cl}>M_{\rm crit}$ lose orbital energy and sink towards the centre of the halo, whereas those with $M_{\rm cl}<M_{\rm crit}$ gain energy and expand outwards. At $M_{\rm cl}\simeq M_{\rm crit}$, the two processes balance, leading to a cluster population that, on average, neither sinks nor expands within the host halo. In CDM haloes, the critical mass is primarily controlled by the upper end of the subhalo mass function. Since massive subhaloes are intrinsically rare, the balance between heating and friction shows substantial halo-to-halo scatter. For dark matter haloes in the mass range associated with dwarf spheroidal galaxies (dSphs), we find $M_{\rm crit}\gtrsim 10^5M_\odot$, comparable to the masses of globular clusters in the Fornax dSph. Stochastic heating by dark substructure may therefore significantly delay the orbital decay of globular clusters and help alleviate the Fornax timing problem.

astro-ph.GA↗

Where Does Randomness Matter in Neural Cellular Automata?

Stochastic cell updates are often used throughout the life of a neural cellular automaton (NCA), from backpropagation through time to final rollout. This leaves two questions entangled: does update randomness help learn a useful rule, and must that randomness remain at execution? We separate training and evaluation update modes in controlled Growing NCA experiments, then vary the states shown during training. Under the standard constant-rate persist recipe, asynchronous training passes the short-horizon quality test in 10/10 runs, compared with 3/10 synchronous runs. All ten asynchronous models also retain the target for 4,096 steps under deterministic evaluation. For a scalar translation-invariant lattice, we derive an exact mean-square criterion: random masking can damp mean modes, but it also injects variance, and a mean-only test misclassifies four non-marginal settings. Finally, among 30 models that all pass the same reconstruction test, eight of ten grow-trained models become off-target at 4,096 steps, while all persist and regenerate models retain the target; damage recovery separates persist from regenerate. The results distinguish optimization reliability, execution mode, and task-specific behavior instead of treating them as one stability property.

cs.LG↗