arXiv ScienceSearch

arXiv subjects

Xiyuan Li

Publications and source records attributed to Xiyuan Li.

9 recordsLinked to original sources

RT-SHCUA: Real-Time Self-Hosted Computer-Use Agent for UAV Control

Natural-language control offers a promising interface for unmanned aerial vehicles (UAVs), but directly applying self-hosted computer-use agents (SHCUAs) to UAV control introduces a structural mismatch. SHCUAs are designed for interactive host-side tool use, where delayed agent iterations are often acceptable. UAV control, however, is coupled with continuously changing physical states, strict timing constraints, safety risks, and security accountability. A stale, unauthorized, or tampered agent decision may therefore lead to unsafe or untraceable vehicle behavior. This paper proposes a real-time and security-oriented restructuring of SHCUA-based UAV control. Instead of allowing an SHCUA to directly issue flight commands, we transform its outputs into contract-bound UAV skill invocations with explicit timing, state, authority, fallback, and evidence semantics. Based on this abstraction, we design an architecture that separates semantic reasoning from onboard execution and security/safety enforcement. Slow cloud or edge reasoning is used for mission understanding, while onboard components validate and dispatch only timely, authorized, and state-consistent skills. Security-critical enforcement points can be protected by TEE-style or microcontroller isolation mechanisms without moving the full language agent or high-frequency flight-control loop into trusted components. Prototype evaluation shows that RT-SHCUA maintains bounded task-level responsiveness while supporting degraded handling, trusted admission, and auditable evidence preservation for SHCUA-mediated UAV actions.

cs.CR

Benchmarking Quantum Software Testing with Scalable Quantum Programs

Quantum software testing (QST) checks whether quantum programs behave according to their intended specifications. A key requirement for QST research is a benchmark that supports rigorous empirical evaluation on programs that are testable and better reflect current software development practices. However, existing studies heavily rely on small hard-coded or circuit-level benchmarks, while available quantum programs are scattered across repositories without clear selection criteria, which limits fair comparison and systematic reproducibility. To this end, we present Qolumbina, a benchmark infrastructure for controlled QST experiments on scalable quantum programs. Qolumbina curates 40 programs from open-source repositories, turns them into test-ready subjects through systematic selection, refactoring, specifications, test case examples, unit tests, and standardized interfaces. We also propose QST-oriented criteria to characterize quantum programs along functionality, output behavior, development complexity, and quantum-specific execution complexity. Using these criteria, our empirical study shows that Qolumbina covers diverse testing-relevant properties and supports scalability analysis beyond fixed-size circuit benchmarks. Through controlled experiments with two recent QST approaches, we demonstrate the feasibility of using Qolumbina for execution-cost and fault-detection studies, and highlight backend-dependent effects that can influence QST result interpretation.

cs.SE

Constraining Host-Level Abuse in Self-Hosted Computer-Use Agents via TEE-Backed Isolation

Self-hosted computer-use agents (SHCUAs), such as OpenClaw, combine natural-language interaction with direct access to host-side resources, including browsers, files, scripts, system commands, and external communication channels. While useful for automating real tasks, this capability also creates a host-level abuse surface: a legitimately deployed agent may be steered toward unsafe operations through malicious messages, indirect prompt injection, unsafe skills, or tampering along the host-side control path. We argue that such risks cannot be addressed by ad hoc blocking rules alone, because the security criticality of an operation depends jointly on its action type, target object, execution context, and potential effect. This paper presents an operation-centric model for risk-based confinement of SHCUA operations. The proposed design keeps ordinary functionality on the constrained REE path, while protecting security-critical classification, authorization, binding, evidence generation, and selected execution-control decisions inside a cloud-native TEE-backed trusted operation plane. We instantiate the architecture on OpenClaw using Intel TDX as the primary trusted backend, with remote terminal-side trusted components verifying TDX-audited commands before constrained local execution. The evaluation shows that the design can block unsafe or policy-disallowed operations before execution, preserve ordinary functionality for allowed workloads, and provide auditable evidence with deployment-dependent overhead.

cs.CR

MemFine: Memory-Aware Fine-Grained Scheduling for MoE Training

The training of large-scale Mixture of Experts (MoE) models faces a critical memory bottleneck due to severe load imbalance caused by dynamic token routing. This imbalance leads to memory overflow on GPUs with limited capacity, constraining model scalability. Existing load balancing methods, which cap expert capacity, compromise model accuracy and fail on memory-constrained hardware. To address this, we propose MemFine, a memory-aware fine-grained scheduling framework for MoE training. MemFine decomposes the token distribution and expert computation into manageable chunks and employs a chunked recomputation strategy, dynamically optimized through a theoretical memory model to balance memory efficiency and throughput. Experiments demonstrate that MemFine reduces activation memory by 48.03% and improves throughput by 4.42% compared to full recomputation-based baselines, enabling stable large-scale MoE training on memory-limited GPUs.

cs.DC

MoFa: A Unified Performance Modeling Framework for LLM Pretraining

The exponential growth in LLM scales, with parameters soaring from billions to trillions, has necessitated distributed pretraining across large clusters comprising thousands to tens of thousands of devices. While hybrid parallelization strategies enable such pretraining, the vast combinatorial strategy space introduces significant optimization challenges. Traditional manual tuning methods incur prohibitive trial-and-error costs, and existing performance modeling approaches exhibit critical limitations: they fail to comprehensively account for prevalent optimization features and ignore the substantial overhead imposed by essential fault tolerance mechanisms like checkpoint recovery in long-duration pretraining. To address these gaps, we propose MoFa, a novel pretraining performance modeling framework that unifies multi-dimensional optimization features and fault tolerance. MoFa incorporates an enhanced cost model to accurately capture the effects of key optimizations and integrates a fault tolerance model based on historical cluster reliability data. Besides, a MoFa-based tuning system is developed to explore optimal pretraining performance and potential bottlenecks in various scenarios. Extensive modeling evaluations demonstrate that MoFa can achieve high prediction accuracy across various scenarios. In addition, through comprehensive tuning experiments, our framework systematically reveals the key factors influencing pretraining performance under different configurations, which provides solid a priori guidance for LLM pretraining system design and deployment.

cs.DC

Machine Learning based Glitch Veto for inspiral binary merger signals using Linear Chirp Transform

Transient non-Gaussian noise artifacts commonly known as glitches remain a major challenge in gravitational wave (GW) detection because they can mimic genuine compact binary coalescence signals and increase the false-alarm rate of detection pipelines. Accurate discrimination between astrophysical signals and instrumental glitches is therefore essential for improving the reliability of GW observations. In this work, we investigate the Linear Chirp Transform (LCT) as a feature extraction technique for glitch classification. Unlike the conventional Fourier transform, the LCT incorporates an additional chirp-rate parameter $\gamma$, enabling improved representation of signals with time-varying frequencies. Applying the LCT to GW strain time series produces three-dimensional chirp-volume spectrograms spanning time, frequency and chirp-rate dimensions, providing richer information than conventional time-frequency spectrograms. The dataset consists of confirmed compact binary coalescence events and glitch samples from the O1-O4 observing runs of the LIGO detectors at Hanford and Livingston. For classification, we employ a hybrid deep learning architecture combining convolutional neural networks (CNNs), gated recurrent units (GRUs) and an attention mechanism. The CNN layers extract local spectro-temporal features, the GRUs model correlations across chirp-rate slices and attention pooling highlights the most informative regions. The proposed framework achieves high classification performance on training and validation datasets, demonstrating that chirp-domain representations provide highly discriminative information for distinguishing merger signals from glitches. These results highlight the potential of combining chirp-based signal processing with deep learning to improve glitch mitigation in current and future GW observatories.

gr-qc

Hourglass Magnetic Field of a Protostellar System

An hourglass-shaped magnetic field pattern arises naturally from the gravitational collapse of a star-forming gas cloud. Most studies have focused on the prestellar collapse phase, when the structure has a smooth and monotonic radial profile. However, most observations target dense clouds that already contain a central protostar, and possibly a circumstellar disk. We utilize an analytic treatment of the magnetic field along with insights gained from simulations to develop a more realistic magnetic field model for the protostellar phase. Key elements of the model are a strong radial magnetic field in the region of rapid collapse, an off-center peak in the magnetic field strength (a consequence of magnetic field dissipation in the circumstellar disk), and a strong toroidal field that is generated in the region of rapid collapse and outflow generation. A model with a highly pinched and twisted magnetic field pattern in the inner collapse zone facilitates the interpretation of magnetic field patterns observed in protostellar clouds.

astro-ph.SR

The Role of r-Modes in Pulsar Spin-down, Pulsar Timing, and Gravitational Waves

We investigate the role of r-mode oscillations in pulsar spin-down and their implications for gravitational wave emission and pulsar timing analysis. Using a non-linear differential framework that includes r-mode contributions, we derive time-dependent solutions for rotational frequency and period evolution. These expressions are validated using observational data from the Crab pulsar with high precision. By analytically fitting braking indices and spin-down coefficients, we link measurable pulsar properties to gravitational wave signatures. Furthermore, we present closed-form expressions for neutron star compactness and tidal deformability using Lambert W and Lambert-Tsallis functions, enabling model-independent inferences from r-mode gravitational wave frequencies. Our results show that incorporating r-modes significantly improves the accuracy of spin-down models and continuous wave detectability, particularly through the inclusion of high-order frequency terms. This framework supports the modeling of timing residuals, glitch quantification, and gravitational wave constraints. Our findings have direct relevance for data analysis in ongoing and future gravitational wave observatories.

astro-ph.HE

A Joint-Chirp-Rate-Time-Frequency Transform for BBH Merger Gravitational Wave Signal Detection

Low-latency detection of Binary Black Hole (BBH) and Binary Neutron Star (BNS) merger Gravitational Wave (GW) signals is essential for enabling multi-messenger observations of such systems. The merger GW signals have varying frequencies and are contaminated by non-stationary noises. Earlier studies of non-templated merger signal detection techniques used traditional Fourier transform-based time-frequency decomposition methods for spectrogram generation, which have had difficulties identifying rapid frequency changes in merger signals with heavy background noise. To address this problem, we introduce the Joint-Chirp-rate-Time-Frequency Transform (JCTFT), in which complex-valued window functions are used to modulate the amplitude, frequency, and phase of the input signal. In addition, we outline the techniques for generating chirp-rate-enhanced time-frequency spectrograms from the results of a JCTFT. We demonstrate an average of 14 improved merger detectability among simulated detector signals with Signal-to-Noise Ratios between 6 and 10 using the InceptionV3 image classification neural network compared to the same network trained with Q-transform spectrograms. The JCTFT is a general transformation technique that can be applied to existing and third-generation GW detector signals. Further studies will aim to improve the efficiency and performance of the JCTFT.

gr-qc