arXiv Science⌕ Search

arXiv subjects

Ping Chen

Publications and source records attributed to Ping Chen.

At least 37 records · Page 2Linked to original sources

Multiwavelength Analysis of Six Luminous Fast Blue Optical Transients

We present multiwavelength observations and analysis of six luminous fast blue optical transients (LFBOTs) discovered in Zwicky Transient Facility (ZTF) survey data. We identified these LFBOTs from their fast light-curve evolution ($t_{1/2}\leq 12 $d), blue colors at peak brightness ($g-r\leq-0.5 $mag), a visible host galaxy, high optical luminosity ($M_g<-20$), and an X-ray or radio detection. With the exception of AT2024aehp (ZTF24abygbss), these transients exhibit peaks in their $10\,$GHz radio light curves at $t_{\text{rest}} \approx 50-100$ d, with peak radio luminosities ranging from $10^{38}-10^{40}$ erg s$^{-1}$. Modeling the radio emission as synchrotron radiation indicates a fast ($v=0.1-0.3c$) shock in a dense ($n_e\approx10^{3}-10^{4}$ cm$^{-3}$) medium. The X-ray emission varies by $\approx2$ orders of magnitude in luminosity ($10^{42}-10^{44}$ erg s$^{-1}$) at $t_{\text{rest}}\sim20 $d. Analysis of the host-galaxy photometry and spectroscopy for each transient shows that they are predominantly nonnuclear (a few kpc offset) with star-forming host galaxies of stellar masses $10^{9}-10^{11} ,M_\odot$. Unlike all other LFBOTs to date, AT2024aehp exhibited a luminous ($M<-19 $mag) plateau in the optical light curve; spectra during this plateau phase showed a featureless blue continuum. The $6-15$ GHz radio emission of AT2024aehp brightened by over an order of magnitude from $t_{\text{rest}} \approx70 $d to $t_{\mathrm{rest}} \approx130 $d. The mostly consistent radio behavior between optically selected LFBOTs implies a similar circumburst medium, leading us to prefer a progenitor scenario in which mass is lost in a consistent way shortly prior to the terminal event, such as a massive star merging with a compact object.

astro-ph.HE↗

When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

While enabling effective collaboration on complex tasks, LLM-based Multi-Agent Systems (MAS) face critical security challenges due to vulnerabilities at the agent and interaction levels. Most existing MAS security defenses are built upon two core assumptions: semantically-explicit malicious attacks and explicit graph-based modeling of the MAS topology and agent-level interactions. In practice, real-world attacks are becoming more semantically stealthy, while MAS execution is typically asynchronous without the temporal alignment assumed by graph-based propagation models. To address these limitations, we propose AcMAS, an activation-based framework for malicious-behavior detection in MAS. By analyzing internal reasoning states in the activation space of local agents, AcMAS detects even stealthy attacks in a synchronization-robust fashion, without relying on explicit interaction graphs. Moreover, our activation analysis provides critical signals to guide AcMAS in restoring the functionality of compromised agents, rather than the disruptive agent isolation commonly used by the state-of-the-art methods. Comprehensive evaluation demonstrates that AcMAS significantly outperforms graph-based baselines against stealthy attacks, by +0.22 F1 in synchronous settings (0.94 vs. 0.72) and by +0.55 F1 in asynchronous settings (0.93 vs. 0.38), with generalization across diverse open-source LLM backbones, attack intensity, and MAS scale.

cs.CR↗

Prior-Anchored Debiasing for Long-Tailed Multi-Organ Pathology Report Generation

Automated pathology report generation from Whole Slide Images (WSIs) has attracted increasing attention in digital pathology. However, existing methods are predominantly developed under single-organ settings, overlooking the multi-organ scenarios encountered in clinical practice, where organ types typically follow a long-tailed distribution. To address this gap, we identify two critical biases: (1) visual representation bias, where the encoder favors head-class patterns over tail-class discriminative features, and (2) textual decoding bias, where the decoder overfits to head-class narrative patterns, yielding diagnostically unreliable outputs for tail-class organs. To mitigate these two biases, we propose a novel Prior-anchored multi-Organ pathology report Generation framework (PriOrGen). Specifically, a Visual-Prototype Anchored Bottleneck module leverages the information bottleneck principle with learnable anchor representations to selectively retain diagnostically relevant visual information while filtering out head-biased redundancy. Secondly, a Meta-Report Anchored Bank module constructs an organ-specific meta-report anchored bank and retrieves organ-faithful textual priors to steer the decoder away from head-class narrative patterns. Extensive experiments on a multi-organ pathology dataset demonstrate that our method effectively mitigates long-tail biases and achieves superior report generation performance across both head and tail organ categories compared to state-of-the-art methods.

cs.CV↗

Decoding the Early-Time Light Curves of Type Ia Supernovae. II. Population Parameters of One Thousand ZTF Supernovae

Early-time light curves of Type Ia Supernovae (SNe Ia) encode critical information about their progenitor systems. We characterize the rise of normal SNe Ia using a volume-complete sample of 972 events from the Zwicky Transient Facility Data Release 2, an order of magnitude larger than any previous dataset for similar analyses. Fitting light curves up to $30\%$ of peak flux with a power-law model under a hierarchical Bayesian framework, we provide robust population-level constraints on the rise time ($t_\mathrm{rise}$; $μ=18.55\pm0.08$ days, $σ=1.42\pm0.07$ days), rise index ($α$; $μ=2.10\pm0.04$, $σ=0.48\pm0.03$ in ZTF $r$), and $g-r$ color evolution ($α_g - α_r$; $μ=0.20\pm0.02$, $σ=0.17\pm0.02$). These power-law fits are sensitive to the chosen truncation epoch if data beyond $\sim$$40\%$ of peak flux are included, but generally converge when restricted to earlier epochs. The relation between rise morphology and light-curve width ($\texttt{SALT2}$ $x_1$ stretch) bifurcates into two distinct regimes: high-stretch SNe Ia show clear trends where a higher $x_1$ correlates with shallower rises and more persistent blue colors, whereas low-stretch SNe Ia lack such trends. While rise times correlate positively with $x_1$ overall, this relation flattens significantly within the high-stretch population. Searching for anomalies, we identify several normal SNe Ia with unusually long rise times, which potentially exhibit short-duration ($\lesssim$2 days) flux excesses over a smooth rise. Long-duration ($\sim$5 days) flux excesses appear common within the high-stretch population and are tied to the shallow rises and early blue colors, pointing to widespread outward $^{56}$Ni mixing. Multi-dimensional explosion models with more realistic progenitor setups are needed to fully reproduce the observed dichotomy in rise morphology and stretch.

astro-ph.HE↗

Red-Teaming Coding Agents from a Tool-Invocation Perspective: An Empirical Security Assessment

Coding agents powered by large language models are becoming central modules of modern IDEs, helping users perform complex tasks by invoking tools. While powerful, tool invocation opens a substantial attack surface. Prior work has demonstrated attacks against general-purpose and domain-specific agents, but none have focused on the security risks of tool invocation in coding agents. To fill this gap, we conduct the first systematic red-teaming of six popular real-world coding agents: Cursor, Claude Code, Copilot, Windsurf, Cline, and Trae. Our red-teaming proceeds in two phases. In Phase 1, we perform prompt leakage reconnaissance to recover system prompts. We discover a general vulnerability, ToolLeak, which allows malicious prompt exfiltration through benign argument retrieval during tool invocation. In Phase 2, we hijack the agent's tool-invocation behavior using a novel two-channel prompt injection in the tool description and return values, achieving remote code execution (RCE). We adaptively construct payloads using security information leaked in Phase 1. In emulation across five backends, our method outperforms baselines on Claude-Sonnet-4, Claude-Sonnet-4.5, Grok-4, and GPT-5. On real agents, our approach succeeds on 19 of 25 agent-LLM pairs, achieving leakage on every agent using Claude and Grok backends. For tool-invocation hijacking, we obtain RCE on every tested agent-LLM pair, with our two-channel method delivering the highest success rate. We provide case studies on Cursor and Claude Code, analyze security guardrails of external and built-in tools, and conclude with practical defense recommendations.

cs.CR↗

Human and AI collaboration for pulmonary nodule segmentation

Medical expert annotators are scarce, and blind reliance on artificial intelligence (AI) can be misleading, motivating approaches in which humans, particularly junior medical trainees or even non-medical personnel, collaborate with AI to achieve robust medical segmentation. Although the Segment Anything Model (SAM) shows promise for general-purpose image segmentation, its performance in human-AI collaboration for specialized medical tasks has not been thoroughly evaluated. Here we present Hi-Seg, a human-in-the-loop segmentation framework for pulmonary nodules built on SAM. Humans iteratively refine prompts through trial-and-error learning and semantic reasoning, progressively guiding SAM toward higher-quality masks. Using chest CT scans from 1,179 patients across 12 centers, we conducted the first large-scale external validation of collaborative human-SAM segmentation. Across all annotator groups, Hi-Seg achieved a mean Dice score of almost 85%, outperforming five state-of-the-art deep learning models by 10-22% and 13 SAM variants by 1-29%. Hi-Seg improved segmentation accuracy while reducing annotation time for medical annotators, and briefly trained non-medical annotators achieved performance comparable to that of the junior medical student. These findings suggest that human-in-the-loop segmentation can reduce clinician workload, enable scalable crowdsourced annotation, and transform clinical workflows by facilitating the safe and efficient integration of foundation models into routine clinical practice.

cs.CV↗

Detection of persistent helium absorption in the 91bg-like type Ia Supernova 2022an

We present optical and near-infrared observations of the fast-declining Type Ia supernova (SN Ia) 2022an. The photometric and spectroscopic properties identify it as a standard 91bg-like event; however, our data reveal a relatively narrow absorption feature with a full width at half maximum (FWHM) of 75 angstroms near $1.037\,μ$m in the rest frame of the observed spectra that persists from around 30 days to nearly 90 days after maximum light. We attribute this feature to He I $1.083\,μ$m line with a blueshifted velocity of $1.3\times10^{4}$ km s$^{-1}$ and a FWHM of $2.1\times10^{3}$ km s$^{-1}$, supported by the detection of multiple optical He I transitions in earlier epochs at a higher velocity around $1.5\times10^{4}$ km s$^{-1}$. The high velocity of the helium could not be explained by helium external to the progenitor at the explosion, such as the stripped surface helium from a companion star. The properties of the helium absorption in SN 2022an spectra instead point to unburnt material in the outer ejecta, thus providing the most compelling evidence to date for helium-bearing ejecta in a 91bg-like SN Ia. Such helium has been predicted for sub-Chandrasekhar-mass double-detonation explosions involving a surface helium shell. No theoretical calculations of modern helium-shell double detonation have been performed at epochs similar to those observed for SN 2022an to study the effect of helium on their spectra, revealing a gap between observations and theoretical calculations in understanding the manifestation of helium in SNe Ia. Nevertheless, the discovery of persistent helium absorption in SN 2022an demonstrates the diagnostic power of NIR spectroscopy for understanding thermonuclear supernova explosions by probing the abundance and structure of their ejecta.

astro-ph.HE↗

A Disappearing Act: Constraints From "Missing" Flares of Repeating Partial TDE Candidates

Recurrent tidal disruption events (rTDEs) are sources that exhibit multiple TDE-like flares; many are likely powered by the recurring partial disruption of a bound star, in a repeating partial TDE (rpTDE). Two such sources, TDE 2022dbl (ASASSN-22ci) and TDE 2020vdq (ZTF20acaazkt), each exhibited two UV/optical flares and, under the assumption of periodicity, both were expected to exhibit a third flare in early 2026. Neither exhibited such a flare, to limits of $L_{\textrm{UV/optical}} \lesssim 10^{42}$ erg s$^{-1}$, $\sim$30$\times$ fainter than the previous flares. Here, we examine several possible explanations. Observing two independent TDEs from the same galaxy within $\sim$2 yr has a probability of $\lesssim$0.5% for measured average TDE rates and currently expected rate enhancements, unless there is extreme intrinsic dispersion in the rates. Theoretical predictions for a double TDE of both stars in a binary are inconsistent with the observed flares. We therefore conclude that TDE 2022dbl and TDE 2020vdq are rpTDEs. To produce only two observable flares with similar energetics, our semi-analytical modeling strongly favors a main-sequence star promptly placed on a bound orbit with a deep initial tidal encounter at pericenter. These results suggest that the majority of rTDEs with multiple flares over a few-year baseline are likely to be rpTDEs, and that a significant fraction of systems may produce only two observable flares. This has important implications for the use of r(p)TDEs as probes of TDE physics and dynamical processes in the nuclei of other galaxies, in addition to the expected yield from upcoming surveys.

astro-ph.HE↗

LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation

Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from diminishing transfer ratio -- the fraction of FM improvement captured by the VM -- as a single scalar cannot convey the rich intermediate knowledge that larger FMs learn. To address this bottleneck, we propose LoopFM (Learning frOm HistOrical RePresentations of FM), a framework that opens a high-bandwidth transfer channel by structuring FM intermediate embeddings as input features (e.g., user history sequence) for downstream VMs, without requiring real-time FM inference at serving and architectural coupling between FM and VM. We provide a theoretical framework for LoopFM with a gain decomposition and transfer-ratio analysis. On three public benchmarks, LoopFM demonstrates strong AUC improvements (e.g., 6%+ on TaobaoAd) and complementary knowledge transfer capability with KD. On industrial-scale systems (billions of examples, trillion-parameter FMs), LoopFM approximately doubles the knowledge transfer ratio on top of KD, delivering a +0.5% conversion improvement in the first half after its initial launch, and +1.03% and +1.22% conversion improvement from two individual launches in the subsequent half.

cs.LG↗

Capture and Stability of Resonant Planet Pairs in Turbulent Disk

We present a theoretical framework for the resonance capture and stability of two-planet systems in turbulent disks. By incorporating stochastic forcing (parameterized by $κ$) alongside laminar angular momentum and eccentricity damping timescales ($τ_{\rm m}, τ_{e}$), we derive an analytical criterion for the general $j:j-1$ mean motion resonances, and validate it through N-body simulations. The outcome is mapped in $κ$-$τ_{\rm m}/τ_{e}$ parameter space, revealing two distinct regimes: resonance trapping and turbulence-induced disruption -- which occurs either directly cross or via temporary capture followed by escape through turbulent diffusion. Crucially, our analysis identifies turbulence as a universal destabilizer. It amplifies the intrinsic overstability mechanism: In laminar disks, escape requires $τ_{\rm m}/τ_{e}$ to drop below a critical limit due to excessive eccentricity excitation. We demonstrate that turbulent diffusion lowers this limit, demanding stronger damping (larger $τ_{\rm m}/τ_{e}$) for stability. Thus, greater turbulence promotes escape, and sufficiently strong diffusion precludes resonance retention irrespective of eccentricity damping.

astro-ph.EP↗

Building a physics-aware AI ecosystem for solid-state hydrogen storage materials

Hydrogen storage remains a central bottleneck for scalable hydrogen energy systems due to the multiscale and coupled nature of the thermodynamics, kinetics, and microstructural evolution of hydrogen storage materials (HSMs). Although artificial intelligence (AI) has accelerated materials discovery, current approaches remain constrained by fragmented data, limited physical consistency, and weak integration with experimental validation. Here, we propose a unified framework that integrates coherent data infrastructure, physics-grounded modeling, and AI-driven inverse design within a closed-loop discovery paradigm. By embedding physical constraints and experimental feedback, this approach enables adaptive, physically consistent optimization, thereby establishing a pathway toward autonomous, digital-twin-enabled discovery of HSMs.

cond-mat.mtrl-sci↗

Low-Luminosity Type IIP Supernovae from the Zwicky Transient Facility Census of the Local Universe. I: Luminosity Function, Volumetric Rate

We present the luminosity function and volumetric rate of a sample of Type IIP supernovae (SNe) from the Zwicky Transient Facility Census of the Local Universe survey (CLU). This is the largest sample of Type IIP SNe from a systematic volume-limited survey to-date. The final sample includes 330 Type IIP SNe and 36 low-luminosity Type II (LLIIP) SNe with $M_{\textrm{r,peak}}>-16$ mag, which triples the literature sample of LLIIP SNe. The fraction of LLIIP SNe is $19^{+3}_{-4}\%$ of the total CLU Type IIP SNe population ($8^{+1}_{-2}\%$ of all core-collapse SNe). This implies that while LLIIP SNe likely represent the fate of core-collapse SNe of $8-12$ \Msun\ progenitors, they alone cannot account for the fate of all massive stars in this mass range. To derive an absolute rate, we estimate the ZTF pipeline efficiency as a function of the apparent magnitude and the local surface brightness. We derive a volumetric rate of $(3.9_{-0.4}^{+0.4}) \times 10^{4}\ \textrm{Gpc}^{-3}\ \textrm{yr}^{-1}$ for Type IIP SNe and $(7.3_{-0.6}^{+0.6}) \times 10^{3}\ \textrm{Gpc}^{-3}\ \textrm{yr}^{-1}$ for LLIIP SNe. Now that the rate of LLIIP SNe is robustly derived, the unresolved discrepancy between core-collapse SN rates and star-formation rates cannot be explained by LLIIP SNe alone.

astro-ph.HE↗

Fundamental Performance Limits of Non-Coherent ISAC: A Data-Aided Sensing Perspective

In this paper, we investigate a bistatic multiple-input multiple-output (MIMO) integrated sensing and communication (ISAC) system over block-fading channels, focusing on the scenario where the sensing and communication receivers (Rxs) are co-located. Under the assumption of unknown channel state information (CSI) at the Rx, two schemes are considered: pilot sensing (PS) and data-aided sensing (DAS). The communication rate-sensing distortion functions for both schemes are characterized. For the DAS scheme, a closed-form asymptotic expression for the sensing distortion is derived by using random matrix theory (RMT). The asymptotic performance analysis explicitly quantifies the significant gains of the DAS scheme, revealing a strict $3$ dB effective SNR improvement in the low-SNR regime and a strictly faster performance scaling rate in the high-SNR limit compared to the PS scheme.

cs.IT↗

HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models

Low-Rank Adaptation (LoRA) dominates parameter-efficient fine-tuning of large language models, yet most variants target dense architectures. Mixture-of-Experts (MoE) models scale parameters at near-constant per-token compute, and their sparse activation patterns create untapped opportunities for more efficient adaptation. We propose Hot-Experts Layer-level Low-Rank Adaptation (HELLoRA), which attaches LoRA modules only to the most frequently activated experts at each layer. This simple mechanism reduces trainable parameters and adapter-induced FLOPs while improving downstream performance, an effect we attribute to a form of structured regularization that preserves pretrained expert specialization. To stress-test HELLoRA under extreme parameter budgets, we further compose it with LoRI to form HELLoRI, which freezes the up-projection and sparsifies the down-projection. Across three MoE backbones, namely OlMoE-1B-7B, Mixtral-8x7B, and DeepSeekMoE, and three task families covering mathematical reasoning, code generation, and safety alignment, HELLoRA consistently outperforms strong PEFT baselines. Relative to vanilla LoRA on OlMoE, HELLoRA uses 15.7% of the trainable parameters, reduces adapter FLOPs by 38.7%, achieves 1.9x the training throughput, and improves accuracy by 9.2%. On DeepSeekMoE, HELLoRA outperforms LoRA while using only 23.2% of its trainable parameters. These results demonstrate that activation-aware adapter placement is an effective and practical route to scaling PEFT for MoE language models.

cs.LG↗

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption

Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by incorporating external knowledge, effectively mitigating their inherent knowledge limitations. However, RAG remains vulnerable to poisoning attacks that manipulate retrieved texts to mislead model outputs. Existing defense mechanisms often lack theoretical robustness guarantees and perform unreliably when the LLM has limited knowledge of the retrieved content. In this work, we propose PRA-RAG, a provably robust retrieval aggregation algorithm designed to defend against poisoning attacks on retrieved texts. PRA-RAG samples multiple combinations of retrieved texts and utilizes geometric structures in the embedding space to identify a robust subset, from which a stable aggregated representation is derived. We provide theoretical bounds on the maximum impact of poisoned retrieved content and establish a quantitative measure of RAG's robustness. Experiments across multiple benchmarks and RAG architectures demonstrate that PRA-RAG reduces the attack success rate to as low as 1% while maintaining an accuracy of 71%, significantly outperforming representative state-of-the-art methods.

cs.IR↗

ESARBench: A Benchmark for Agentic UAV Embodied Search and Rescue

The rapid advancement of Multimodal Large Language Models (MLLMs) has empowered Unmanned Aerial Vehicle (UAV) with exceptional capabilities in spatial reasoning, semantic understanding, and complex decision-making, making them inherently suited for UAV Search and Rescue (SAR). However, existing UAV SAR research is dominated by traditional vision and path-planning methods and lacks a comprehensive and unified benchmark for embodied agents. To bridge this gap, we first propose the novel task of \textbf{Embodied Search and Rescue (ESAR)}, which requires aerial agents to autonomously explore complex environments, identify rescue clues, and reason about victim locations to execute informed decision-making. Additionally, we present \textbf{ESARBench}, the first comprehensive benchmark designed to evaluate MLLM-driven UAV agents in highly realistic SAR scenarios. Leveraging Unreal Engine 5 and AirSim, we construct four high-fidelity, large-scale open environments mapped directly from real-world Geographic Information System (GIS) data to ensure photorealistic landscapes. To rigorously simulate actual rescue operations, our benchmark incorporates dynamic variables including weather conditions, time of day, and stochastic clue placement. Furthermore, we create a dataset of 600 tasks modeled after real-world rescue cases and propose a robust set of evaluation metrics. We evaluate diverse baselines, ranging from traditional heuristics to advanced ground and aerial MLLM-based ObjectNav agents. Experimental results highlight the challenges in ESAR, revealing critical bottlenecks in spatial memory, aerial adaptation, and the trade-off between search efficiency and flight safety. We hope ESARBench serves as a valuable resource to advance research on Embodied Search and Rescue domain. Source code and project page: https://4amgodvzx.github.io/ESAR.github.io.

cs.RO↗

AT2024wpp: An Extremely Luminous Fast Ultraviolet Transient Powered by Accretion onto a Black Hole

We present the discovery of AT 2024wpp ("Whippet"), a fast and luminous 18cow-like transient. At a redshift of z=0.0868, revealed by Keck Cosmic Web Imager spectroscopy of its faint star-forming host, it is the fourth-nearest example of its class to date. Rapid identification of the source in the Zwicky Transient Facility data stream permitted ultraviolet-through-optical observations to be obtained prior to peak, allowing the first determination of the peak bolometric luminosity (2x10^45 erg/s), maximum photospheric radius (10^15 cm), and total radiated energy (10^51 erg) of an 18cow-like object. We present results from a comprehensive multiwavelength observing campaign, including a far-UV spectrum from the Cosmic Origins Spectrograph on the Hubble Space Telescope and deep imaging extending >100 days post-explosion from the Very Large Telescope, Hubble Space Telescope, Very Large Array, and Atacama Large Millimetre Array. We interpret the observations under a model in which a rapidly-accreting central engine blows a fast (~0.2c) wind into the surrounding medium and irradiates it with X-rays. The high Doppler velocities and intense ionization within this wind prevent identifiable spectroscopic features from appearing in the ejecta or in the surrounding circumstellar material. Weak H and He signatures do emerge in the spectra after 35 days in the form of double-peaked narrow lines. Each peak is individually narrow (full width ~3000 km/s) but the two components are separated by ~6600 km/s, indicating stable structures of denser material, possibly representing streams of tidal ejecta or an ablated companion star.

astro-ph.HE↗

Classical and spin polarizabilities of singly heavy baryons within heavy baryon chiral perturbation theory

We present a systematic study of the electromagnetic and spin polarizabilities of spin-1/2 singly charmed baryons at $\mathcal{O}(p^4)$ within the framework of heavy baryon chiral perturbation theory. Our results show that the higher-order corrections to the electric polarizability are small, while those to the magnetic polarizability are relatively larger due to the small mass splitting of singly charmed baryons and are closely related to transition magnetic moments. Furthermore, we find that the spin polarizabilities of singly charmed baryons, except for $γ_{M1M1}$, are much smaller than those of the nucleons. We have also calculated the polarizabilities for singly bottom baryons, with the results showing generally larger values than those of singly charmed baryons.

hep-ph↗