arXiv ScienceSearch

arXiv subjects

Xia Zhou

Publications and source records attributed to Xia Zhou.

At least 19 recordsLinked to original sources

Globally regular charged black holes in non-polynomial quasi-topological gravity with Born-Infeld electrodynamics

We construct exact static, spherically symmetric charged solutions in four-dimensional non-polynomial quasi-topological gravity coupled to Born-Infeld electrodynamics. We focus on the model $h(p)=p/(1+\ell^{2}p)$, whose vacuum branch develops a curvature singularity at a finite radius. We show that Born-Infeld nonlinearities can remove this singularity within a finite region of parameter space, yielding globally regular geometries with an asymptotically flat exterior and a finite-curvature AdS-type core. The regular sector contains both horizonless configurations and RBHs, separated by a degenerate-horizon boundary. We further identify a continuous branch of regular black holes with a triple-degenerate inner horizon and a simple outer event horizon, satisfying $κ_-=0$ and $κ_+\neq0$. These results provide a converse example to cases in which introducing charge spoils the regularity of a black hole that is regular in vacuum. In the present model, Born-Infeld electrodynamics instead removes the finite-radius singularity of a gravitational branch that is singular in vacuum and supports globally regular charged geometries, including regular black holes with nontrivial inner-horizon structure.

gr-qc

Video-Based Palm-Vein Authentication under Challenging Conditions

Palm-vein biometrics are increasingly used for secure, contactless authentication. Yet real-world deployment exposes them to surface noise (sweat, dirt), illumination and motion variation, and temperature-driven changes in vascular visibility, which remain underexplored for lack of data captured under such conditions. To study these effects, we introduce the Columbia University Palm-vein (CUP) dataset, to our knowledge the first public video-based palm-vein dataset. CUP records every palm under four surface conditions (a clean baseline, warm, wet, and dirty) and pairs each subject with physiological and demographic metadata. On it we benchmark twenty-one recognizers spanning static, video, and multi-frame aggregation architectures. Models that verify reliably on clean palms lose most of their accuracy on dirty ones, and the mean equal error rate (EER) roughly quadruples. We recover much of that robustness along both axes of the capture. Temporally, a consensus over the few frames the sensor already returns cancels transient corruption; spatially, a test-time matcher that adds no learned parameters fuses the global cosine with a saliency-steered region-level optimal transport that routes the comparison around corrupted regions. The full design leads on every surface of CUP in EER, TAR@FAR=0.01, and Rank-1, at 4.3M parameters and 3.1 GFLOPs, a fraction of the video models' cost. Attached to four frozen state-of-the-art backbones it cuts their mean EER by 29-37% without retraining, and on four public single-image datasets the regional matching alone still helps. A preliminary audit across ten demographic and physiological traits finds two warm-condition gaps, along body water and gender, that survive multiple-comparison correction. CUP will be released for non-commercial research use at https://github.com/MobileX-CU/CUP_v1 upon publication.

cs.CV

Distinct Modes of Pulsar Glitch Activity Revealed by the Waiting-Time--Amplitude Morphology

Pulsar glitches display diverse behaviors. They are commonly labeled as "Vela-like" or "Crab-like," yet these qualitative categories do not provide reproducible quantitative boundaries. We constructed an objective classification scheme based on the joint distribution of backward waiting time $Δt_{-}$ and fractional glitch amplitude $Δν/ν$. A sample of 349 glitches from 22 pulsars was compiled by cross-matching the Jodrell Bank and ATNF glitch databases. For each source, we used two-dimensional kernel density estimation (KDE) together with highest-density regions (HDRs) to derive geometric descriptors, including the HDR area, anisotropy ratio, and out-of-HDR fraction. Hierarchical clustering of the KDE--HDR descriptors revealed two morphology classes independent of visual inspection. In this convention, the Compact class is associated with small HDR areas, elevated anisotropy, and limited peripheral occupancy. The Extended class is associated with broader HDR occupancy and stronger low-density extensions. We found that bootstrap resampling and hyperparameter sensitivity tests confirmed this partition for most sources. Nevertheless, two prolific pulsars (PSR J0537-6910 and PSR J1341-6220) were found to occupy intermediate positions. Logistic regression identified the mean backward waiting time $\langleΔt_{-}\rangle$ as the strongest class predictor (in-sample $\mathrm{AUC}=0.88$; leave-one-out $\mathrm{AUC}_{\mathrm{CV}}=0.59$, reflecting the limited sample size). Additionally, a parallel analysis using $Δt_{+}$ recovered the same two-class structure but with reduced stability. The data suggest that the Compact class is consistent with near-complete reservoir depletion, whereas the Extended class reflects partial, avalanche-like releases. As a result, short-term measurements of $G$ in Extended-class pulsars may underestimate the true crustal superfluid fraction.

astro-ph.HE

Probing strange quark matter objects with future space-based gravitational wave detectors DECIGO and BBO

The Strange Quark Matter (SQM) hypothesis posits that objects composed of SQM could exist across a wide mass range, from strange planets (SPs) to strange stars (SSs). It has been proposed that gravitational waves (GWs) emitted by inspiraling SS-SP systems may be detectable by ground-based GW observatories such as advanced LIGO and the Einstein Telescope. Nevertheless, such a system may undergo an extended period of orbital evolution in a close configuration before entering the inspiraling phase. During this time, it can generate continuous GW signals at frequencies ranging from milli-hertz (mHz) to deci-hertz (dHz). The detailed characteristics of these GWs have not yet been thoroughly explored. In this study, we delve into the continuous GW features of SS-SP systems, with a focus on exploring the physically viable parameter space. We compared the GW signals emitted by these systems to the sensitivity curves of next-generation space-based GW detectors like the Deci-hertz Interferometer Gravitational wave Observatory (DECIGO) and the Big Bang Observer (BBO). Our analyses demonstrate that both the DECIGO and BBO detectors are capable of detecting continuous GWs from SS-SP systems across a broad parameter space. These GWs carry important information for testing the SQM hypothesis, as well as for advancing our understanding of supernovae and compact star merger processes.

astro-ph.HE

Mammal: Supporting Breastfeeding Monitoring Through Computational Garments with Inter-Body Sensing

Breastfeeding provides critical insight into infant feeding competence and physiological health, yet objective monitoring remains difficult due to the intimate and internal nature of feeding. We present Mammal, a caregiver-worn computational garment that unobtrusively monitors breastfeeding without attaching sensors to the infant. Mammal leverages inter-body signal transmission through natural mouth-to-breast contact to capture infant cardiac and feeding-related acoustic signals on the caregiver's body. Using novel algorithms to detect latch onset, infer infant electrocardiogram (ECG), and identify suck and swallow events from inter-body signals, Mammal estimates latch duration, in-feeding heart rate, suck-swallow-breathe (SSB) ratio, and milk intake. In a user study with 10 caregiver-infant dyads, Mammal achieves a mean absolute percentage error (MAPE) of 5.56% for latch duration, a mean absolute error (MAE) of 3.61 bpm for infant heart rate estimation, a mean absolute error of 0.12 for SSB ratio estimation, and a mean relative error of 15.76% for milk intake, with participants reporting high comfort and wearability.

cs.HC

Detectability of continuous gravitational waves from planetary-mass companions orbiting compact stars

Binary systems with ultrashort-period planetary-mass companions are expected to radiate continuous gravitational waves (GWs). However, earlier studies found that the detectability of such systems by the Laser Interferometer Space Antenna (LISA) is unlikely. In this study, we investigate the detectability of GWs from planetary-mass companions orbiting pulsars (PSRs) or white dwarfs (WDs) whose fundamental parameters, essential for calculating GW properties, have been measured. We compare the GW signals from our sample with the sensitivity curves of space-based GW detectors. We find that fourteen sources achieve a signal-to-noise ratio (\(\text{S/N}\)) of \(\gtrsim 5\) within four years of observations. Among these, three sources have PSR primaries (2S 0918-549 b, 4U 0513-40 b, and 4U 1543-62), and eleven systems possess WD primaries (BW Scl b, CP Eri b, CR Boo b, EF Eri b, GP Com b, GW Lib b, SDSS J0926+3624 b, SDSS J1507+5230 b, SMSS J1606-1000 b, SRGeJ0453 b, and WZ Sge b). We note that their detectability is less probable with near-term missions such as LISA, TianQin, and Taiji. Nevertheless, they could be detected by more advanced, future-generation observatories, such as the Deci-hertz Interferometer Gravitational wave Observatory (DECIGO) and the Big Bang Observer (BBO). This offers the potential to investigate the formation and evolution of ultrashort-period planetary-mass companions around compact stars through joint GW and electromagnetic surveys.

astro-ph.HE

AD-R1: Closed-Loop Reinforcement Learning for End-to-End Autonomous Driving with Impartial World Models

End-to-end models for autonomous driving hold the promise of learning complex behaviors directly from sensor data, but face critical challenges in safety and handling long-tail events. Reinforcement Learning (RL) offers a promising path to overcome these limitations, yet its success in autonomous driving has been elusive. We identify a fundamental flaw hindering this progress: a deep seated optimistic bias in the world models used for RL. To address this, we introduce a framework for post-training policy refinement built around an Impartial World Model. Our primary contribution is to teach this model to be honest about danger. We achieve this with a novel data synthesis pipeline, Counterfactual Synthesis, which systematically generates a rich curriculum of plausible collisions and off-road events. This transforms the model from a passive scene completer into a veridical forecaster that remains faithful to the causal link between actions and outcomes. We then integrate this Impartial World Model into our closed-loop RL framework, where it serves as an internal critic. During refinement, the agent queries the critic to ``dream" of the outcomes for candidate actions. We demonstrate through extensive experiments, including on a new Risk Foreseeing Benchmark, that our model significantly outperforms baselines in predicting failures. Consequently, when used as a critic, it enables a substantial reduction in safety violations in challenging simulations, proving that teaching a model to dream of danger is a critical step towards building truly safe and intelligent autonomous agents.

cs.CV

DriveCombo: Benchmarking Compositional Traffic Rule Reasoning in Autonomous Driving

Multimodal Large Language Models (MLLMs) are rapidly becoming the intelligence brain of end-to-end autonomous driving systems. A key challenge is to assess whether MLLMs can truly understand and follow complex real-world traffic rules. However, existing benchmarks mainly focus on single-rule scenarios like traffic sign recognition, neglecting the complexity of multi-rule concurrency and conflicts in real driving. Consequently, models perform well on simple tasks but often fail or violate rules in real world complex situations. To bridge this gap, we propose DriveCombo, a text and vision-based benchmark for compositional traffic rule reasoning. Inspired by human drivers' cognitive development, we propose a systematic Five-Level Cognitive Ladder that evaluates reasoning from single-rule understanding to multi-rule integration and conflict resolution, enabling quantitative assessment across cognitive stages. We further propose a Rule2Scene Agent that maps language-based traffic rules to dynamic driving scenes through rule crafting and scene generation, enabling scene-level traffic rule visual reasoning. Evaluations of 14 mainstream MLLMs reveal performance drops as task complexity grows, particularly during rule conflicts. After splitting the dataset and fine-tuning on the training set, we further observe substantial improvements in both traffic rule reasoning and downstream planning capabilities. These results highlight the effectiveness of DriveCombo in advancing compliant and intelligent autonomous driving systems.

cs.CV

The stochastic gravitational wave background from QCD phase transition in the framework of higher-order GUP

This work studies the impact of a new higher-order generalized uncertainty principle (GUP) on the stochastic gravitational wave background (SGWB) associated with a QCD-scale first-order phase transition. Assuming a strongly first-order transition at the QCD-scale as a phenomenological benchmark, the analysis shows that the sign and magnitude of the dimensionless deformation parameter $β_0$ play a crucial role. For negative $β_0$, the thermodynamic quantities of the radiation fluid develop a maximal temperature beyond which entropy and pressure vanish, and the SGWB spectrum exhibits divergent behavior at high temperatures, so this branch is discarded as phenomenologically inconsistent. For positive $β_0$, the higher-order GUP shifts the SGWB peak frequency towards lower values and slightly enhances the peak energy density, with the size of the effect controlled by $β_0$. For natural values $β_0=\mathcal{O}\left( 1 \right)$ the corrections at QCD temperatures are strongly suppressed, whereas larger benchmark values still compatible with existing experimental and cosmological bounds can induce appreciable shifts in the SGWB spectrum. A future detection of a QCD-scale first-order SGWB would therefore allow the framework developed here to be used to translate the measured signal into constraints on the higher-order GUP parameter, providing an indirect probe of quantum gravity effects.

gr-qc

CorrectAD: A Self-Correcting Agentic System to Improve End-to-end Planning in Autonomous Driving

End-to-end planning methods are the de facto standard of the current autonomous driving system, while the robustness of the data-driven approaches suffers due to the notorious long-tail problem (i.e., rare but safety-critical failure cases). In this work, we explore whether recent diffusion-based video generation methods (a.k.a. world models), paired with structured 3D layouts, can enable a fully automated pipeline to self-correct such failure cases. We first introduce an agent to simulate the role of product manager, dubbed PM-Agent, which formulates data requirements to collect data similar to the failure cases. Then, we use a generative model that can simulate both data collection and annotation. However, existing generative models struggle to generate high-fidelity data conditioned on 3D layouts. To address this, we propose DriveSora, which can generate spatiotemporally consistent videos aligned with the 3D annotations requested by PM-Agent. We integrate these components into our self-correcting agentic system, CorrectAD. Importantly, our pipeline is an end-to-end model-agnostic and can be applied to improve any end-to-end planner. Evaluated on both nuScenes and a more challenging in-house dataset across multiple end-to-end planners, CorrectAD corrects 62.5% and 49.8% of failure cases, reducing collision rates by 39% and 27%, respectively.

cs.CV

DriveLiDAR4D: Sequential and Controllable LiDAR Scene Generation for Autonomous Driving

The generation of realistic LiDAR point clouds plays a crucial role in the development and evaluation of autonomous driving systems. Although recent methods for 3D LiDAR point cloud generation have shown significant improvements, they still face notable limitations, including the lack of sequential generation capabilities and the inability to produce accurately positioned foreground objects and realistic backgrounds. These shortcomings hinder their practical applicability. In this paper, we introduce DriveLiDAR4D, a novel LiDAR generation pipeline consisting of multimodal conditions and a novel sequential noise prediction model LiDAR4DNet, capable of producing temporally consistent LiDAR scenes with highly controllable foreground objects and realistic backgrounds. To the best of our knowledge, this is the first work to address the sequential generation of LiDAR scenes with full scene manipulation capability in an end-to-end manner. We evaluated DriveLiDAR4D on the nuScenes and KITTI datasets, where we achieved an FRD score of 743.13 and an FVD score of 16.96 on the nuScenes dataset, surpassing the current state-of-the-art (SOTA) method, UniScene, with an performance boost of 37.2% in FRD and 24.1% in FVD, respectively.

cs.CV

RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation

Synthetic data is crucial for advancing autonomous driving (AD) systems, yet current state-of-the-art video generation models, despite their visual realism, suffer from subtle geometric distortions that limit their utility for downstream perception tasks. We identify and quantify this critical issue, demonstrating a significant performance gap in 3D object detection when using synthetic versus real data. To address this, we introduce Reinforcement Learning with Geometric Feedback (RLGF), RLGF uniquely refines video diffusion models by incorporating rewards from specialized latent-space AD perception models. Its core components include an efficient Latent-Space Windowing Optimization technique for targeted feedback during diffusion, and a Hierarchical Geometric Reward (HGR) system providing multi-level rewards for point-line-plane alignment, and scene occupancy coherence. To quantify these distortions, we propose GeoScores. Applied to models like DiVE on nuScenes, RLGF substantially reduces geometric errors (e.g., VP error by 21\%, Depth error by 57\%) and dramatically improves 3D object detection mAP by 12.7\%, narrowing the gap to real-data performance. RLGF offers a plug-and-play solution for generating geometrically sound and reliable synthetic videos for AD development.

cs.CV

Plasma lens with frequency-dependent dispersion measure effects on fast radio bursts

Radio signals propagating through inhomogeneous plasma media deviate from their original paths, producing frequency-dependent magnification effects. In this paper, after reviewing the classical plasma-lensing theory, we have found a fundamental contradiction: the classical model assumes that the distribution of lensing plasma medium is related to the frequency-independent image position; however, our analysis demonstrates that both the image position ($θ(ν)$) and dispersion measure (DM$(ν)$) are inherently frequency-dependent when signals traverse a structured plasma medium. We have been able to resolve this paradox by developing a framework that explicitly incorporates frequency-dependent dispersion measures (DMs) following power-law relationships ($\rm DM\propto ν^γ$). Our analysis shows that the signal magnification decreases systematically with decreasing frequency, offering a plausible explanation for the frequency-dependent peak flux densities observed in fast radio bursts (FRBs), particularly in the case of the repeating FRB 180814.J0422+73. Our results suggest these FRBs could originate from the magnetized compact star magnetospheres. By considering these plasma-lensing effects on the sub-pulses of an FRB across different frequencies, we have the ability to more accurately investigate the intrinsic properties of FRBs via precise measurements of radio signals.

astro-ph.HE

Set Phasers to Stun: Beaming Power and Control to Mobile Robots with Laser Light

We present Phaser, a flexible system that directs narrow-beam laser light to moving robots for concurrent wireless power delivery and communication. We design a semi-automatic calibration procedure to enable fusion of stereo-vision-based 3D robot tracking with high-power beam steering, and a low-power optical communication scheme that reuses the laser light as a data channel. We fabricate a Phaser prototype using off-the-shelf hardware and evaluate its performance with battery-free autonomous robots. Phaser delivers optical power densities of over 110 mW/cm$^2$ and error-free data to mobile robots at multi-meter ranges, with on-board decoding drawing 0.3 mA ($97\%$ less current than Bluetooth Low Energy). We demonstrate Phaser fully powering gram-scale battery-free robots to nearly 2x higher speeds than prior work while simultaneously controlling them to navigate around obstacles and along paths. Code, an open-source design guide, and a demonstration video of Phaser is available at https://mobilex.cs.columbia.edu/phaser.

cs.RO

Combating Falsification of Speech Videos with Live Optical Signatures (Extended Version)

High-profile speech videos are prime targets for falsification, owing to their accessibility and influence. This work proposes VeriLight, a low-overhead and unobtrusive system for protecting speech videos from visual manipulations of speaker identity and lip and facial motion. Unlike the predominant purely digital falsification detection methods, VeriLight creates dynamic physical signatures at the event site and embeds them into all video recordings via imperceptible modulated light. These physical signatures encode semantically-meaningful features unique to the speech event, including the speaker's identity and facial motion, and are cryptographically-secured to prevent spoofing. The signatures can be extracted from any video downstream and validated against the portrayed speech content to check its integrity. Key elements of VeriLight include (1) a framework for generating extremely compact (i.e., 150-bit), pose-invariant speech video features, based on locality-sensitive hashing; and (2) an optical modulation scheme that embeds $>$200 bps into video while remaining imperceptible both in video and live. Experiments on extensive video datasets show VeriLight achieves AUCs $\geq$ 0.99 and a true positive rate of 100% in detecting falsified videos. Further, VeriLight is highly robust across recording conditions, video post-processing techniques, and white-box adversarial attacks on its feature extraction methods. A demonstration of VeriLight is available at https://mobilex.cs.columbia.edu/verilight.

cs.CV

Physics of Strong Magnetism with eXTP

In this paper we present the science potential of the enhanced X-ray Timing and Polarimetry (eXTP) mission, in its new configuration, for studies of strongly magnetized compact objects. We discuss the scientific potential of eXTP for quantum electrodynamic (QED) studies, especially leveraging on the recent observations made with the NASA IXPE mission. Given eXTP's unique combination of timing, spectroscopy, and polarimetry, we focus on the perspectives for physics and astrophysics studies of strongly magnetized compact objects, such as magnetars and accreting X-ray pulsars. Developed by an international Consortium led by the Institute of High Energy Physics of the Chinese Academy of Sciences, the eXTP mission is expected to launch in early 2030.

astro-ph.HE

Dense Matter in Neutron Stars with eXTP

In this White Paper, we present the potential of the enhanced X-ray Timing and Polarimetry (eXTP) mission to constrain the equation of state of dense matter in neutron stars, exploring regimes not directly accessible to terrestrial experiments. By observing a diverse population of neutron stars - including isolated objects, X-ray bursters, and accreting systems - eXTP's unique combination of timing, spectroscopy, and polarimetry enables high-precision measurements of compactness, spin, surface temperature, polarimetric signals, and timing irregularity. These multifaceted observations, combined with advances in theoretical modeling, pave the way toward a comprehensive description of the properties and phases of dense matter from the crust to the core of neutron stars. Under development by an international Consortium led by the Institute of High Energy Physics of the Chinese Academy of Sciences, the eXTP mission is planned to be launched in early 2030.

astro-ph.HE

Discovery of the anti-glitch in PSR J1835$-$1106

We report the detection of an anti-glitch with a fractional frequency change of $Δν/ν=-3.46(6)\times10^{-9}$ in the rotation-powered pulsar PSR J1835$-$1106 at MJD 55813, based on timing observations collected with the Nanshan 26-m and Parkes 64-m radio telescopes from January 2000 to July 2022. A comparison of the average pulse profiles within $\pm300$ d of the event reveals no significant morphological changes. We also estimate the angular velocity lag between the normal and superfluid components at the time of the glitch, showing that one of the superfluid glitch models is incompatible with PSR J1835$-$1106 due to its insufficient spin-down rate and angular velocity lag. The wind braking scenario offers a viable alternative, consistent with the observed spin-down behavior, glitch amplitude, and post-glitch recovery. High-cadence, high-sensitivity monitoring of similar events is essential to distinguish between internal (superfluid) and external (wind-related) glitch mechanisms.

astro-ph.HE