arXiv ScienceSearch

arXiv subjects

Ping Wang

Publications and source records attributed to Ping Wang.

At least 19 recordsLinked to original sources

Study of dynamical dark energy using DESI DR2 BAO, CMB, and Type Ia Supernovae data

Dark Energy Spectroscopic Instrument (DESI) Collaboration reported an evidence for dynamical dark energy model using baryon acoustic oscillation (BAO) measurements from DESI Data Release 2 (DR2), cosmic microwave background (CMB), and three Type Ia Supernovae (SNe) data (Pantheon+, Union3, and DES-SN5YR) recently. The significances are from $2.8\sim4.2\sigma$ depending on which SNe sample is used, and the highest significance is obtained based on the DES-SN5YR data. But more recently, Dark Energy Survey (DES) Collaboration recalibrated the DES-SN5YR data, and a significance of only $3.2\sigma$ is obtained based on the calibrated data, which is called DES-Dovekie. In this work, we perform a comprehensive analysis of five representative dynamical dark energy parametrization models, including Chevallier-Polarski-Linder (CPL), Barboza-Alcaniz (BA), exponential (EXP), logarithmic (LOG), and Jassal-Bagla-Padmanabhan (JBP), using DESI DR2 BAO, Planck CMB, ACT DR6, and SNe data from Pantheon+, Union3, and DES-Dovekie. Similarly, a significance of around $3\sigma$ is obtained based on the DES-Dovekie data. And the significances are $2.1\sim3.7\sigma$ for different SNe samples and different parametrization models, indicating an evolutionary dark energy dynamics. Furthermore, a phantom-like behavior in the past and has transitioned into quintessence-like behavior today for the dynamical dark energy is preferred.

physics.gen-ph

Contrast-Free Autonomous Navigation of Untethered Endovascular Microrobots Using Single-Plane Fluoroscopy

Reliable three-dimensional (3D) navigation of magnetically actuated untethered microrobots remains a major barrier to clinical translation. X-ray fluoroscopy is the standard real-time imaging modality for endovascular procedures, but single-plane fluoroscopy provides only a two-dimensional (2D) projection, eliminating depth information and complicating autonomous navigation. Recovering this information through biplane imaging or repeated contrast-enhanced angiography increases procedural complexity, radiation exposure, or contrast burden. Here, we introduce VISTA (Virtual Integration for Spatial Tracking and Autonomy), a digital twin framework enabling contrast-free autonomous navigation under single-plane fluoroscopy. VISTA reconstructs vascular anatomy as a 3D digital twin, discretizes the vessel centerline into navigation milestones, and assigns the detected 2D robot position to the nearest projected milestone. Consecutive milestones define the local vessel orientation used to generate magnetic actuation commands, converting single-plane fluoroscopic observations into topology-constrained navigation states without requiring contrast injection during navigation. VISTA is demonstrated across anatomically distinct vascular phantoms under continuous flow and within the inferior vena cava of a live rat in vivo. Compared with conventional fluoroscopic human-in-the-loop control, VISTA reduced navigation time by up to 62%, corrective actuation commands by up to 98%, and radiation exposure by up to 57%. These results establish VISTA as a digital twin-guided framework for contrast-free autonomous navigation of untethered endovascular microrobots using widely available single-plane fluoroscopy.

cs.RO

From Passive Response to Proactive Correction: Enhancing LLM Robustness Against Input Fact Perturbations

Large language models (LLMs) frequently produce confident yet factually incorrect responses when user inputs contain misleading premises, a phenomenon we attribute to fact perturbations in the input. Existing approaches to hallucination mitigation typically assume reliable user inputs, overlooking how such factual errors can actively mislead model reasoning. To address this vulnerability, we propose DEDUCE, a three-stage framework that transforms LLMs from passive responders into proactive error correctors. DEDUCE operates in three stages: (1) detect errors through fine-grained fact extraction and verification; (2) devise correction strategies via multi perspective deliberation; and (3) correct misconceptions while delivering reliable answers. We also present MisFactQA, a dataset containing factual errors of varying degrees, and propose new metrics for evaluating model robustness. Experiments on TruthfulQA, FalseQA, and our MisFactQA benchmark demonstrate that DEDUCE significantly improves both accuracy and error correction capability. Consistent gains across Qwen, LLaMA, and Gemma families confirm its effectiveness and scalability.

cs.CL

CRAFT: LLM-Based Iterative Refinement for Temporal Reasoning over Clinical Narratives

Understanding the temporal progression of symptoms in clinical narratives is critical for disease monitoring, safety surveillance, and causality assessment. Clinical narratives, however, rarely provide explicit temporal anchors. Current approaches to temporal information reasoning focus predominantly on pairwise relation classification across multi-visit and timestamp-rich records, leaving the reconstruction of structured symptom trajectories from individual anchor-sparse reports largely unaddressed. We propose CRAFT, an LLM framework that pairs a generator with a constraint-based verifier to iteratively produce and refine stage-wise symptom timelines through targeted feedback. We conduct evaluation on MedTempo, a new benchmark of 5,347 vaccine adverse-event narratives spanning three COVID-19 vaccine types, with expert-validated temporal stage annotations for 3,166 reports. Experiments across four LLM backbones demonstrate that CRAFT consistently improves temporal ordering accuracy, with ablation analysis isolating the contribution of generator and verifier components across model capability levels.

cs.CL

GS$^{2}$CI: Robust Gaussian Splatting For Snapshot Compressive Imaging via Large Vision Model Priors

Snapshot Compressive Imaging (SCI) offers an efficient solution for high-speed video acquisition and, under exposure-time camera--scene relative motion, multi-view scene capture by compressing temporal or spatial information into a single 2D measurement. While recent studies have explored SCI for 3D scene reconstruction, existing methods struggle with significant challenges due to information loss, limited viewpoint diversity, and the computational burden of jointly optimizing 3D representations and camera poses. In this work, we propose a novel framework that reconstructs high-quality 3D scenes from a single SCI measurement by leveraging 3D Gaussian Splatting (3DGS) and the powerful priors of large-scale vision foundation models (VFMs). Our primary reconstruction combines measurement-derived 3D VFM initialization with SCI-aware Gaussian optimization. After coarse-stage convergence, an auxiliary 2D VFM provides pseudo-view supervision at synthesized viewpoints for local appearance refinement. To further address the instability caused by ambiguous SCI supervision during 3DGS optimization, we introduce Opacity-Guided Splitting and Growth Regulation (OSGR), an SCI-specific densification strategy that augments split candidates using local opacity statistics, discourages loss-compensating opacity inflation through mean-opacity regulation, and bounds representation growth with explicit candidate-ratio and Gaussian-count constraints. Extensive experiments across multiple benchmarks demonstrate that our method achieves the strongest overall performance, combining leading reconstruction quality and robustness to viewpoint variation with competitive computational efficiency.

cs.CV

A Dual Evaluation for Music Transcription

Automatic music transcription systems produce sheet music that can be read and played back. We argue that these two targets call for complementary evaluations of notation similarity to a reference score and playback similarity to the original performance, respectively. Our study considers notation similarity metrics from the optical music recognition literature and a wide range of playback-similarity methods validated through a listening study across over 100 participants and 230 piano recordings covering 23 works, 30 performers, and six composers. We find, fortuitously, that the playback similarity metric that correlates best with human judgments, CLEWS, is also the cheapest to run. We also find that the two evaluation dimensions favor different systems among a collection of 24 pipelines formed by pairing eight audio-to-MIDI models with three MIDI-to-score converters, with the latter component systematically determining the favored objective. The complementarity between metrics also holds when adding to the pool Rubato, a new end-to-end system that offers substantially improved notation similarity while remaining competitive, though not the best, on playback similarity.

cs.SD

Erodible bed turbulence modulation driven by transition between longitudinal and transverse bedforms at varying Shields numbers

The mechanism of turbulence modulation in particle-laden flow over erodible beds remains an open question. Using particle-resolved direct numerical simulations, this study realises a longitudinal-to-transverse bedform transition by varying the Shields number, revealing non-monotonic modulation of near-wall turbulence. At low Shields numbers, streamwise sediment ridges generate form-induced streaks that produce a distinct secondary peak in the premultiplied energy spectra, exceeding the conventional near-wall turbulent peak and enhancing the turbulent kinetic energy. As the Shields number increases, saltation intensifies and disrupts these structures, causing the secondary peak to vanish in the streamwise direction and weaken in the spanwise direction, thereby suppressing turbulence. Proper orthogonal decomposition of the bed surface reveals a redistribution of modal contribution from a single dominant mode to higher-order modes, with longitudinal features persisting as remnants, directly linking bedform evolution to turbulence modulation.

physics.flu-dyn

BayesSeg: A Bayesian Optimization Framework for State Segmentation of Electricity Consumption Time Series

In Non-Intrusive Load Monitoring (NILM), adaptive segmentation of electricity consumption time series is critical for appliance recognition. However, prevailing methods face challenges including heuristic parameter tuning, boundary sensitivity, and metric saturation. This paper proposes BayesSeg, a unified framework integrating time-series segmentation, multidimensional evaluation, and automatic parameter optimization. The segmentation layer employs a dual steady-state criterion based on the tail value and mean of preceding subsequences, combined with a sequential extraction and complement-set parsing strategy, to achieve precise unsupervised partitioning of steady-state and transition-state segments. The evaluation layer maps segmentation results to binary state sequences and formulates a composite metric integrating an event-level F1 score (event_F1) with Normalized Mutual Information (NMI). The event_F1 quantifies switching-event precision and recall via tolerance matching, while NMI captures global structural consistency, jointly overcoming the boundary sensitivity and limited discriminability of point-wise metrics. In the optimization layer, the composite score serves as the objective function for Bayesian optimization, which constructs a TPE surrogate model for efficient global parameter-space exploration. Experiments on the SustDataED2 dataset demonstrate that Bayesian optimization requires only ~100 objective evaluations to locate a parameter region within 0.35% deviation of the exhaustive grid-search optimum. The framework achieves a weighted composite score of 0.7149 and an event_F1 of 0.9340 while reducing optimization latency from ~5300 seconds to under 1 second, a speedup exceeding 5700x. BayesSeg automates segmentation configuration and provides a scalable, efficient solution for time-series analysis in NILM and related domains.

cs.AI

Intelligent Multi-UAV Navigation in ITNTNs: A Hierarchical LLM Approach

The deployment of high-speed Uncrewed Aerial Vehicles (UAVs) in 3D aerial highways necessitates robust coordination of physical flight kinematics and multi-tier network handovers. While Deep Reinforcement Learning (DRL) offers rapid tactical control, it lacks the zero-shot strategic reasoning required to quickly adapt to dynamic Integrated Terrestrial and Non-Terrestrial Networks (ITNTNs). Conversely, Large Language Models (LLMs) excel at semantic reasoning but suffer from high inference latency, rendering them unsuitable for real-time aerodynamic control. To bridge this gap, we propose a novel Hierarchical LLM-driven control framework. A massive cloud-based LLM deployed on a High-Altitude Platform Station (HAPS) manages slow-timescale global load balancing, while lightweight edge-LLMs on individual UAVs translate local observations into tactical sub-goals. These sub-goals guide a fast-timescale physical DRL controller to execute collision-free, handover-aware trajectories. Simulation results demonstrate that our agentic architecture significantly reduces collision rates and improves aggregate system throughput compared to existing baselines.

cs.RO

JustDiag!: A Diagnostic Justification Engine for Accountable Root Cause Analysis

Large language models can produce fluent root cause analyses, but fluent final answers alone are insufficient evidence for accountability in high-stakes operations. In real incident response, engineers need to know what evidence supported a diagnosis, which alternatives were considered, where contradictions remained, and whether the system resolved the case or preserved uncertainty. We address this gap with JustDiag, a diagnostic justification engine for RCA that maintains an explicit process state over evidence, findings, competing hypotheses, conflicts, and next checks. We evaluated the system on 66 real-world incidents using a two-layer protocol that separately scores final-answer quality and process quality. Relative to a matched control without diagnostic justification, JustDiag achieved stronger outcome and process scores, while accepting slightly lower terminal completion due to more calibrated non-closure. These results suggest that accountable RCA requires explicit diagnostic justification artifacts and process-aware evaluation, not only fluent final answers.

cs.SE

Deep Probabilistic Unfolding for Quantized Compressive Sensing

We propose a deep probabilistic unfolding model to address the classical quantized compressive sensing problem that leverages an unfolding framework to enhance the reconstruction accuracy and efficiency. Unlike previous unfolding methods that apply L2 projection to measurements, we derive a closed-form, numerically stable likelihood gradient projection, which allows the model to respect the true quantization physics, turning the hard quantization constraint into a soft probabilistic guidance. Furthermore, an efficient, dual-domain Mamba module is specifically designed to dynamically capture and fuse the multi-scale local and global features, ensuring the interactions between the distant but correlated regions. Extensive experiments demonstrate the state-of-the-art performance of the proposed method over previous works, which is capable of promoting the application of quantized compressive sensing in real life.

cs.CV

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control and connectivity remains a fundamental challenge. In this paper, we study a multi-UAV system operating in an integrated terrestrial and non-terrestrial network (ITNTN) comprising terrestrial base stations and high-altitude platform stations (HAPS). We consider a three-dimensional (3D) aerial highway scenario where UAVs must adapt their motion to ensure collision avoidance, efficient traffic flow, and reliable communication under dynamic and partially observable conditions. We first model the problem as a hierarchical multi-objective partially observable Markov decision process (H-MO-POMDP), capturing the coupling between control and communication objectives. Based on this formulation, we propose a large language model (LLM)-driven hierarchical multi-rate control framework. At the global level, an LLM-based controller on the HAPS performs long-term planning for load balancing and handover decisions. At the local level, each UAV employs a hybrid controller that integrates a slow-timescale LLM for high-level spatial reasoning with a reinforcement learning agent for faster UAV-to-infrastructure (U2I) communication and motion control. We further develop a high-fidelity 3D simulation platform by integrating the gym-pybullet-drones environment with 3GPP-compliant RF/THz channel models. Numerical results demonstrate that the proposed framework significantly outperforms state-of-the-art baselines, achieving a 14% increase in transportation efficiency and a 25% improvement in telecommunication throughput. Additionally, it achieves a 23% reduction in physical collision rates, demonstrating strong handover stability and zero-shot generalization in dynamic scenarios.

cs.AI

Finite-Size Gradient Transport in Large Language Model Pretraining: From Cascade Size to Intensive Transport Efficiency

We introduce a finite-size gradient-transport framework for real language-model training, based on five observables $(D,z,\beta,\delta,v_{\mathrm{rel}})$ that separate cascade size, duration, absolute transport, and intensive transport efficiency. We analyze direct raw-gradient measurements from Pico-LM across four scales and 125 aligned steps, together with a five-scale Pythia companion dataset built from 153 aligned checkpoint-difference update fields. The same algebraic closure holds in both families, and both share a near-unity cascade-size backbone, but they occupy distinct transport regimes: Pico-LM shows positive duration scaling and negative intensive-efficiency scaling, whereas Pythia remains near the $D=1$ baseline with only weak positive efficiency scale dependence. Randomized-field controls give nearly matched null floors in the intensive and duration channels, indicating that the contrast reflects different real departures from a shared null skeleton rather than different null calibrations. The families also differ in stepwise power-law compressibility: Pico-LM retains clean duration and efficiency power laws, whereas Pythia preserves the size backbone but shows weaker one-slope compressibility in those channels. External performance associations are correspondingly channel-level, carried mainly by $v_{\mathrm{rel}}$ and normalized cascade duration, while $D(t)$ acts as a shared size backbone without a significant exponent-level performance association. These results support a reusable transport measurement framework without claiming a universal fixed point or a first-principles derivation of neural scaling laws.

cs.LG

GRM Scientific Pipeline

The Gamma-Ray Monitor (GRM) is a key payload of the Space-based multiband astronomical Variable Objects Monitor (SVOM) mission, which is designed to detect gamma ray bursts (GRBs) within the energy range of 15 keV to 5 MeV. The GRM Instrument Center (GRM\_IC) features real-time data processing through the X-band, enabling rapid response of high-energy GRB events. The system employs an event-driven architecture and distributed design, achieving efficient processing and real-time monitoring of massive observational data. Through comprehensive data production processes and scientific data product management, the system achieves efficient production of scientific data products of the L1B / C level through the submission of jobs to the task scheduling system. Through modular architecture design and automated processing workflow, the GRM data processing system realizes precise conversion and scientific analysis of GRB detection data, providing robust technical support for future system upgrades and cross-platform collaboration.

astro-ph.IM

The Gamma-Ray Monitor onboard the SVOM satellite

The Gamma-Ray Monitor (GRM) is a key scientific payload onboard the Space-based Multi-band Variable Object Monitor (SVOM) satellite, designed specifically for the detection and study of gamma-ray bursts (GRBs). Launched into a 625 km low-Earth orbit on 22 June 2024, GRM serves as a large-area, wide-field-of-view instrument capable of observing the hard X-ray and soft gamma-ray emissions in the energy range of 15 keV to 5 MeV. Its primary scientific objectives include: promptly triggering and localizing GRBs (with particular sensitivity to short-hard GRBs), measuring spectral and temporal properties of bursts, monitoring charged particle fluxes in orbit. GRM successfully detected its first GRB (GRB 240627B) on 27 June 2024, and has since maintained a detection rate of more than 100 GRBs per year. Cross-instrument comparisons with detectors such as GECAM and Fermi/GBM have validated the performance and data quality of GRM. This paper provides a comprehensive overview of GRM instrument design, reliability verification through ground testing, in-orbit triggering and localization algorithms, performance calibration, and preliminary in-orbit results, demonstrating its capability as a versatile gamma-ray all-sky monitor.

astro-ph.IM

Study on the detector energy response of SVOM/GRM

The SVOM mission is specifically designed to for the detection and localization of Gamma-Ray Bursts (GRBs) and subsequent follow-up observations. Among the four telescopes installed on the SVOM satellite, the Gamma-Ray Monitor (GRM) plays a crucial role in capturing the prompt emission of GRBs due to its wide field of view (FOV) and broad energy range. Accurate determination of the detector's energy response is vital for analyzing GRM data, particularly considering the significant impact of the atmospheric albedo effect on this response. This research focuses on deriving the detector's energy response and establishing a calibration database for the GRM, with particular emphasis on investigating the atmospheric albedo effect. The study shows that the contribution of albedo photons to the detector's effective area depends strongly on the orientation of the GRD line of sight (LoS) relative to Earth and on the incident direction of the GRB. When the GRD LoS is anti-Earth oriented, the albedo effect is minimal, with the highest proportion of albedo effective area accounting for approximately 10% of the total effective area. This occurs when the incident angle of the GRB is nearly perpendicular to the LoS. Conversely, if the GRD LoS is not pointing away from Earth and the GRB arrives from angles greater than about 90$^{\circ}$, the albedo component can become predominant, contributing up to around 100% of the total effective area. This is especially pronounced in the 8-20 keV range, where the direct effective area drops to zero due to the large GRB injection angle. Our results show that, it is necessary for GRM to consider the atmospheric albedo effects in detector response, otherwise the spectral and localization analyses will result in biased measurements.

astro-ph.HE

The trigger and localization system of SVOM-GRM

The Space multi-band Variable Object Monitor (SVOM) is an astronomical satellite jointly developed by China and France, primarily focused on the detection of gamma-ray bursts (GRBs) and transient sources. The SVOM satellite was launched on 22nd June, 2024 with four payloads installed onboard. As one of payload, GRM comprises 3 gamma-ray detectors (each detector has an effective area of approximately 200~cm$^{2}$) with distinct pointing directions, enabling the temporal and spectral measurements as well as localization of GRBs in the energy range of 15-5000 keV. This article firstly introduces the on-board localization algorithm design for GRM and presents preliminary test results. Then, leveraging abundant ground-based computational resources, a joint fitting method for spectral and localization analysis using Monte Carlo Markov Chain (MCMC) is implemented. In contrast to the on-board localization algorithm, the on-ground MCMC method comprehensively considers the influence of spectral characteristics, thereby mitigating systematic biases. Finally, a systematic analysis based on this method is provided, highlighting the localization and spectral measurement capabilities of GRM. The preliminary localization analysis result for the on-board detected GRB 240629A by both GRM and Fermi/GBM shows that the localization result (error$\sim$4.14$^{\circ}$) of GRM is consistent with the Fermi/GBM result.

astro-ph.IM

RACER: Retrieval-Augmented Contextual Rapid Speculative Decoding

Autoregressive decoding in Large Language Models (LLMs) generates one token per step, causing high inference latency. Speculative decoding (SD) mitigates this through a guess-and-verify strategy, but existing training-free variants face trade-offs: retrieval-based drafts break when no exact match exists, while logits-based drafts lack structural guidance. We propose $\textbf{RACER}$ ($\textbf{R}$etrieval-$\textbf{A}$ugmented $\textbf{C}$ont$\textbf{e}$xtual $\textbf{R}$apid Speculative Decoding), a lightweight and training-free method that integrates retrieved exact patterns with logit-driven future cues. This unification supplies both reliable anchors and flexible extrapolation, yielding richer speculative drafts. Experiments on Spec-Bench, HumanEval, and MGSM-ZH demonstrate that RACER consistently accelerates inference, achieving more than $2\times$ speedup over autoregressive decoding, and outperforms prior training-free methods, offering a scalable, plug-and-play solution for efficient LLM decoding. Our source code is available at $\href{https://github.com/hkr04/RACER}{https://github.com/hkr04/RACER}$.

cs.CL