arXiv Science⌕ Search

arXiv subjects

Ping Chen

Publications and source records attributed to Ping Chen.

At least 55 records · Page 3Linked to original sources

Classical and spin polarizabilities of singly heavy baryons within heavy baryon chiral perturbation theory

We present a systematic study of the electromagnetic and spin polarizabilities of spin-1/2 singly charmed baryons at $\mathcal{O}(p^4)$ within the framework of heavy baryon chiral perturbation theory. Our results show that the higher-order corrections to the electric polarizability are small, while those to the magnetic polarizability are relatively larger due to the small mass splitting of singly charmed baryons and are closely related to transition magnetic moments. Furthermore, we find that the spin polarizabilities of singly charmed baryons, except for $γ_{M1M1}$, are much smaller than those of the nucleons. We have also calculated the polarizabilities for singly bottom baryons, with the results showing generally larger values than those of singly charmed baryons.

hep-ph↗

SafeCtrl: Region-Aware Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress

The widespread deployment of text-to-image diffusion models is significantly challenged by the generation of visually harmful content, such as sexually explicit content, violence, and horror imagery. Common safety interventions, ranging from input filtering to model concept erasure, often suffer from two critical limitations: (1) a severe trade-off between safety and context preservation, where removing unsafe concepts degrades the fidelity of the safe content, and (2) vulnerability to adversarial attacks, where safety mechanisms are easily bypassed. To address these challenges, we propose SafeCtrl, a Region-Aware safety control framework operating on a Detect-Then-Suppress paradigm. Unlike global safety interventions, SafeCtrl first employs an attention-guided Detect module to precisely localize specific risk regions. Subsequently, a localized Suppress module, optimized via image-level Direct Preference Optimization (DPO), neutralizes harmful semantics only within the detected areas, effectively transforming unsafe objects into safe alternatives while leaving the surrounding context intact. Extensive experiments across multiple risk categories demonstrate that SafeCtrl achieves a superior trade-off between safety and fidelity compared to state-of-the-art methods. Crucially, our approach exhibits improved resilience against adversarial prompt attacks, offering a precise and robust solution for responsible generation.

cs.CV↗

Agentic Flow Steering and Parallel Rollout Search for Spatially Grounded Text-to-Image Generation

Precise Text-to-Image (T2I) generation has achieved great success but is hindered by the limited relational reasoning of static text encoders and the error accumulation in open-loop sampling. Without real-time feedback, initial semantic ambiguities during the Ordinary Differential Equation trajectory inevitably escalate into stochastic deviations from spatial constraints. To bridge this gap, we introduce AFS-Search (Agentic Flow Steering and Parallel Rollout Search), a training-free closed-loop framework built upon FLUX.1-dev. AFS-Search incorporates a training-free closed-loop parallel rollout search and flow steering mechanism, which leverages a Vision-Language Model (VLM) as a semantic critic to diagnose intermediate latents and dynamically steer the velocity field via precise spatial grounding. Complementarily, we formulate T2I generation as a sequential decision-making process, exploring multiple trajectories through lookahead simulations and selecting the optimal path based on VLM-guided rewards. Further, we provide AFS-Search-Pro for higher performance and AFS-Search-Fast for quicker generation. Experimental results show that our AFS-Search-Pro greatly boosts the performance of the original FLUX.1-dev, achieving state-of-the-art results across three different benchmarks. Meanwhile, AFS-Search-Fast also significantly enhances performance while maintaining fast generation speed.

cs.AI↗

Chain-of-Trajectories: Unlocking the Intrinsic Generative Optimality of Diffusion Models via Graph-Theoretic Planning

Diffusion models operate in a reflexive System 1 mode, constrained by a fixed, content-agnostic sampling schedule. This rigidity arises from the curse of state dimensionality, where the combinatorial explosion of possible states in the high-dimensional noise manifold renders explicit trajectory planning intractable and leads to systematic computational misallocation. To address this, we introduce Chain-of-Trajectories (CoTj), a train-free framework enabling System 2 deliberative planning. Central to CoTj is Diffusion DNA, a low-dimensional signature that quantifies per-stage denoising difficulty and serves as a proxy for the high-dimensional state space, allowing us to reformulate sampling as graph planning on a directed acyclic graph. Through a Predict-Plan-Execute paradigm, CoTj dynamically allocates computational effort to the most challenging generative phases. Experiments across multiple generative models demonstrate that CoTj discovers context-aware trajectories, improving output quality and stability while reducing redundant computation. This work establishes a new foundation for resource-aware, planning-based diffusion modeling. The code is available at https://github.com/UnicomAI/CoTj.

cs.LG↗

IdentityGuard: Context-Aware Restriction and Provenance for Personalized Synthesis

The nature of personalized text-to-image models poses a unique safety challenge that generic context-blind methods are ill-equipped to handle. Such global filters create a dilemma: to prevent misuse, they are forced to damage the model's broader utility by erasing concepts entirely, causing unacceptable collateral damage.Our work presents a more precisely targeted approach, built on the principle that security should be as context-aware as the threat itself, intrinsically bound to the personalized concept. We present IDENTITYGUARD, which realizes this principle through a conditional restriction that blocks harmful content only when combined with the personalized identity, and a concept-specific watermark for precise traceability. Experiments show our approach prevents misuse while preserving the model's utility and enabling robust traceability. By moving beyond blunt, global filters, our work demonstrates a more effective and responsible path toward AI safety.

cs.CR↗

Mass Production of 2023 KMTNet Microlensing Planets I: Low Mass Ratio

We initiate the systematic search for planets in the 2023 data of the Korea Microlensing Telescope Network (KMTNet), focusing on those planets found by the KMTNet AnomalyFinder with low preliminary estimates of the mass-ratio, $q<2\times 10^{-4}$. The 2023 season is the first for which the photometry of all events was re-reduced prior to the AnomalyFinder search, potentially increasing its sensitivity to planets. We find three strong low-$q$ planet candidates, KMT-2023-BLG-0164 ($q\sim 1.3\times 10^{-4}$), KMT-2023-BLG-1286 ($q\sim 1.9\times 10^{-4}$), and KMT-2023-BLG-1746 ($q\sim 8\times 10^{-5}$). KMT-2023-BLG-0164 is notable in that the source is projected on a very bright ($I=16.0$) foreground star, which is either the planet's host or (more likely) a companion to the host. We obtain a spectrum, finding that its mass and distance are $M\sim 1.0\,M_\odot$ and $D\sim 1.5$ kpc, the latter being the distance of the lens ($D_L$) regardless of whether the spectroscopic target is the host or its companion. We also analyze two other candidates, KMT-2023-BLG-0614 and KMT-2023-BLG-1593, which are unlikely to enter the statistical sample due to their ambiguous interpretations as possible non-planetary events.

astro-ph.EP↗

MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference

We present MeanCache, a training-free caching framework for efficient Flow Matching inference. Existing caching methods reduce redundant computation but typically rely on instantaneous velocity information (e.g., feature caching), which often leads to severe trajectory deviations and error accumulation under high acceleration ratios. MeanCache introduces an average-velocity perspective: by leveraging cached Jacobian--vector products (JVP) to construct interval average velocities from instantaneous velocities, it effectively mitigates local error accumulation. To further improve cache timing and JVP reuse stability, we develop a trajectory-stability scheduling strategy as a practical tool, employing a Peak-Suppressed Shortest Path under budget constraints to determine the schedule. Experiments on FLUX.1, Qwen-Image, and HunyuanVideo demonstrate that MeanCache achieves 4.12X and 4.56X and 3.59X acceleration, respectively, while consistently outperforming state-of-the-art caching baselines in generation quality. We believe this simple yet effective approach provides a new perspective for Flow Matching inference and will inspire further exploration of stability-driven acceleration in commercial-scale generative models.

cs.LG↗

Beyond Geometry: Artistic Disparity Synthesis for Immersive 2D-to-3D

Current 2D-to-3D conversion methods achieve geometric accuracy but are artistically deficient, failing to replicate the immersive and emotionally resonant experience of professional 3D cinema. This is because geometric reconstruction paradigms mistake deliberate artistic intent, such as strategic zero-plane shifts for pop-out effects and local depth sculpting, for data noise or ambiguity. This paper argues for a new paradigm: Artistic Disparity Synthesis, shifting the goal from physically accurate disparity estimation to artistically coherent disparity synthesis. We propose Art3D, a preliminary framework exploring this paradigm. Art3D uses a dual-path architecture to decouple global depth parameters (macro-intent) from local artistic effects (visual brushstrokes) and learns from professional 3D film data via indirect supervision. We also introduce a preliminary evaluation method to quantify cinematic alignment. Experiments show our approach demonstrates potential in replicating key local out-of-screen effects and aligning with the global depth styles of cinematic 3D content, laying the groundwork for a new class of artistically-driven conversion tools.

cs.CV↗

RandSet: Randomized Corpus Reduction for Fuzzing Seed Scheduling

Seed explosion is a fundamental problem in fuzzing seed scheduling, where a fuzzer maintains a huge corpus and fails to choose promising seeds. Existing works focus on seed prioritization but still suffer from seed explosion since corpus size remains huge. We tackle this from a new perspective: corpus reduction, i.e., computing a seed corpus subset. However, corpus reduction could lead to poor seed diversity and large runtime overhead. Prior techniques like cull_queue, AFL-Cmin, and MinSet suffer from poor diversity or prohibitive overhead, making them unsuitable for high-frequency seed scheduling. We propose RandSet, a novel randomized corpus reduction technique that reduces corpus size and yields diverse seed selection simultaneously with minimal overhead. Our key insight is introducing randomness into corpus reduction to enjoy two benefits of a randomized algorithm: randomized output (diverse seed selection) and low runtime cost. Specifically, we formulate corpus reduction as a set cover problem and compute a randomized subset covering all features of the entire corpus. We then schedule seeds from this small, randomized subset rather than the entire corpus, effectively mitigating seed explosion. We implement RandSet on three popular fuzzers: AFL++, LibAFL, and Centipede, and evaluate it on standalone programs, FuzzBench, and Magma. Results show RandSet achieves significantly more diverse seed selection than other reduction techniques, with average subset ratios of 4.03% and 5.99% on standalone and FuzzBench programs. RandSet achieves a 16.58% coverage gain on standalone programs and up to 3.57% on FuzzBench in AFL++, triggers up to 7 more ground-truth bugs than the state-of-the-art on Magma, while introducing only 1.17%-3.93% overhead.

cs.SE↗

A Faint Progenitor System for the Faint Supernova 2024vjm

Type Ia Supernovae (SNe Ia) are well known for their role as standardizable cosmological candles. Their uniformity is credited to their single origin as thermonuclear explosions of White dwarf (WD) stars. Nevertheless, some SNe Ia break this regularity. Prominently, the Iax subclass are less energetic and remarkably diverse, raising questions about their progenitor systems. While no progenitor system of a normal SN Ia has ever been detected, a luminous blue star was identified in pre-explosion images of the site of the bright SN Iax SN 2012Z, suggested to be a helium giant companion star acting as a mass donor to a WD SN progenitor. This is in line with models of weak mass accretion of a WD from a binary companion, producing an explosion that does not fully disrupt the star. However, these models fail to explain the properties of the faintest Type Iax explosions, suggesting either they originate from other WD binary systems, or even from massive progenitor stars. Here, we present the faint SN Iax SN 2024vjm - possibly the faintest supernova observed to date. Using a deep pre-explosion image taken by the recently launched Euclid space mission, we show that its progenitor system must be fainter than the helium giant SN Iax progenitor candidate of SN 2012Z, as well as that of the luminous red companion or remnant of the faint SN 2008ha, and may require a subdwarf helium star as a mass donor. The deep image also provides strong arguments against a massive star origin for this faint supernova. Our observations argue that SN 2024vjm is a WD explosion, but we find that remarkably faint SNe Iax fade more slowly than bright ones, i.e., they evolve in an opposite manner from the famous Phillips relation that makes regular SNe Ia cosmological candles.

astro-ph.HE↗

Photometry and Spectroscopy of SN 2024pxl: A Luminosity Link Among Type Iax Supernovae

We present extensive ultraviolet to optical photometric and optical to near-infrared (NIR) spectroscopic follow-up observations of the nearby intermediate-luminosity ($M_V = -16.81\pm0.19$~mag) Type Iax supernova (SN) 2024pxl in NGC 6384. SN~2024pxl exhibits a faster light curve than the high-luminosity members of this class, and slower than low-luminosity events. The observationally well-constrained rise time of $\sim$11 days and an estimated synthesized $^{56}$Ni mass of 0.03\, M$_\odot$, based on analytical modeling of the integrated spectral energy distribution light curve, are consistent with models of the weak deflagration of a carbon-oxygen white dwarf. Our optical spectral sequence of SN~2024pxl shows weak \ion{Si}{2} lines and spectral evolution similar to other high-luminosity Type Iax SNe, but also a prominent early-time \ion{C}{2} line, like lower-luminosity Type Iax SNe. The late-time optical spectrum of SN~2024pxl closely matches that of SN~2014dt, and its NIR spectral evolution aligns with that of other well-studied, high-luminosity Type Iax SNe. The spectral-line expansion velocities of SN~2024pxl are at the lower end of the Type Iax SN velocity distribution, and the velocity distribution of iron-group elements compared to intermediate-mass elements suggests that the ejecta are mixed on large scales, as expected in pure deflagration models. SN~2024pxl exhibits characteristics intermediate between those of high-luminosity and low-luminosity Type~Iax SNe, further establishing a link across this diverse class.

astro-ph.HE↗

APEX: A Decoupled Memory-based Explorer for Asynchronous Aerial Object Goal Navigation

Aerial Object Goal Navigation, a challenging frontier in Embodied AI, requires an Unmanned Aerial Vehicle (UAV) agent to autonomously explore, reason, and identify a specific target using only visual perception and language description. However, existing methods struggle with the memorization of complex spatial representations in aerial environments, reliable and interpretable action decision-making, and inefficient exploration and information gathering. To address these challenges, we introduce \textbf{APEX} (Aerial Parallel Explorer), a novel hierarchical agent designed for efficient exploration and target acquisition in complex aerial settings. APEX is built upon a modular, three-part architecture: 1) Dynamic Spatio-Semantic Mapping Memory, which leverages the zero-shot capability of a Vision-Language Model (VLM) to dynamically construct high-resolution 3D Attraction, Exploration, and Obstacle maps, serving as an interpretable memory mechanism. 2) Action Decision Module, trained with reinforcement learning, which translates this rich spatial understanding into a fine-grained and robust control policy. 3) Target Grounding Module, which employs an open-vocabulary detector to achieve definitive and generalizable target identification. All these components are integrated into a hierarchical, asynchronous, and parallel framework, effectively bypassing the VLM's inference latency and boosting the agent's proactivity in exploration. Extensive experiments show that APEX outperforms the previous state of the art by +4.2\% SR and +2.8\% SPL on challenging UAV-ON benchmarks, demonstrating its superior efficiency and the effectiveness of its hierarchical asynchronous design. Our source code is provided in \href{https://github.com/4amGodvzx/apex}{GitHub}

cs.RO↗

Vision-Language Controlled Deep Unfolding for Joint Medical Image Restoration and Segmentation

We propose VL-DUN, a principled framework for joint All-in-One Medical Image Restoration and Segmentation (AiOMIRS) that bridges the gap between low-level signal recovery and high-level semantic understanding. While standard pipelines treat these tasks in isolation, our core insight is that they are fundamentally synergistic: restoration provides clean anatomical structures to improve segmentation, while semantic priors regularize the restoration process. VL-DUN resolves the sub-optimality of sequential processing through two primary innovations. (1) We formulate AiOMIRS as a unified optimization problem, deriving an interpretable joint unfolding mechanism where restoration and segmentation are mathematically coupled for mutual refinement. (2) We introduce a frequency-aware Mamba mechanism to capture long-range dependencies for global segmentation while preserving the high-frequency textures necessary for restoration. This allows for efficient global context modeling with linear complexity, effectively mitigating the spectral bias of standard architectures. As a pioneering work in the AiOMIRS task, VL-DUN establishes a new state-of-the-art across multi-modal benchmarks, improving PSNR by 0.92 dB and the Dice coefficient by 9.76\%. Our results demonstrate that joint collaborative learning offers a superior, more robust solution for complex clinical workflows compared to isolated task processing. The codes are provided in https://github.com/cipi666/VLDUN.

eess.IV↗

Effects of Stellar X-ray Photoevaporation on Planetesimal Formation via the Streaming Instability

The formation of planetesimals via the streaming instability (SI) is a crucial step in planet formation, yet its triggering conditions and efficiency are highly sensitive to both disk properties and specific evolutionary processes. We aim to study the planetesimal formation via the SI, driven by the stellar X-ray photoevaporation during the late stages of disk dispersal, and quantify its dependence on key disk and stellar parameters. We use the DustPy code to simulate the dust dynamics including coagulation, fragmentation, and radial drift in a viscously accreting disk undergoing stellar X-ray photoevaporation. Stellar X-rays drive the disk dispersal, opening a cavity at a few au orbital distance and inducing the formation of an associated local pressure maximum. This pressure maximum acts as a trap for radially drifting dust, therefore enhancing the dust density to the critical level required to initiate the streaming instability and the subsequent collapse into planetesimals. The fiducial model produces 31.4 M_\oplus of planetesimals with an initial dust to final planetesimal conversion efficiency of 20.4%. This pathway is most efficient in larger disks with higher metallicities, lower viscosities, higher dust fragmentation threshold velocities, and/or around stars with higher X-ray luminosities. This work demonstrates that stellar X-ray photoevaporation is a robust and feasible mechanism for triggering planetesimal formation via the SI during the final clearing phase of protoplanetary disk evolution.

astro-ph.EP↗

MARO: Learning Stronger Reasoning from Social Interaction

Humans face countless scenarios that require reasoning and judgment in daily life. However, existing large language model training methods primarily allow models to learn from existing textual content or solve predetermined problems, lacking experience in real scenarios involving interaction, negotiation, and competition with others. To address this, this paper proposes Multi-Agent Reward Optimization (MARO), a method that enables large language models (LLMs) to acquire stronger reasoning abilities by learning and practicing in multi-agent social environments. Specifically, MARO first addresses the sparse learning signal problem by decomposing final success or failure outcomes into each specific behavior during the interaction process; second, it handles the uneven role distribution problem by balancing the training sample weights of different roles; finally, it addresses environmental instability issues by directly evaluating the utility of each behavior. Experimental results demonstrate that MARO not only achieves significant improvements in social reasoning capabilities, but also that the abilities acquired through social simulation learning can effectively transfer to other tasks such as mathematical reasoning and instruction following. This reveals the tremendous potential of multi-agent social learning in enhancing the general reasoning capabilities of LLMs.

cs.AI↗

MIRAGE: Exploring How Large Language Models Perform in Complex Social Interactive Environments

Large Language Models (LLMs) have shown remarkable capabilities in environmental perception, reasoning-based decision-making, and simulating complex human behaviors, particularly in interactive role-playing contexts. This paper introduces the Multiverse Interactive Role-play Ability General Evaluation (MIRAGE), a comprehensive framework designed to assess LLMs' proficiency in portraying advanced human behaviors through murder mystery games. MIRAGE features eight intricately crafted scripts encompassing diverse themes and styles, providing a rich simulation. To evaluate LLMs' performance, MIRAGE employs four distinct methods: the Trust Inclination Index (TII) to measure dynamics of trust and suspicion, the Clue Investigation Capability (CIC) to measure LLMs' capability of conducting information, the Interactivity Capability Index (ICI) to assess role-playing capabilities and the Script Compliance Index (SCI) to assess LLMs' capability of understanding and following instructions. Our experiments indicate that even popular models like GPT-4 face significant challenges in navigating the complexities presented by the MIRAGE. The datasets and simulation codes are available in \href{https://github.com/lime728/MIRAGE}{github}.

cs.CL↗

QueryIPI: Query-agnostic Indirect Prompt Injection on Coding Agents

Modern coding agents integrated into IDEs orchestrate powerful tools and high-privilege system access, creating a high-stakes attack surface. Prior work on Indirect Prompt Injection (IPI) is mainly query-specific, requiring particular user queries as triggers and leading to poor generalizability. We propose query-agnostic IPI, a new attack paradigm that reliably executes malicious payloads under arbitrary user queries. Our key insight is that malicious payloads should leverage the invariant prompt context (i.e., system prompt and tool descriptions) rather than variant user queries. We present QueryIPI, an automated framework that uses tool descriptions as optimizable payloads and refines them via iterative, prompt-based blackbox optimization. QueryIPI leverages system invariants for initial seed generation aligned with agent conventions, and iterative reflection to resolve instruction-following failures and safety refusals. Experiments on five simulated agents show that QueryIPI achieves up to 87% success rate, outperforming the best baseline (50%). Crucially, generated malicious descriptions transfer to real-world coding agents, highlighting a practical security risk.

cs.CR↗

A free-floating-planet microlensing event caused by a Saturn-mass object

A population of free-floating planets is known from gravitational microlensing surveys. None have a directly measured mass, owing to a degeneracy with the distance, but the population statistics indicate that many are less massive than Jupiter. We report a microlensing event -- KMT-2024-BLG-0792/OGLE-2024-BLG-0516, which was observed from both ground- and space-based telescopes -- that breaks the mass-distance degeneracy. The event was caused by an object with 0.219^{+0.075}_{-0.046} Jupiter masses that is either gravitationally unbound or on a very wide orbit. Through comparison with the statistical properties of other observed microlensing events and predictions from simulations, we infer that this object likely formed in a protoplanetary disk (like a planet), not in isolation (like a brown dwarf), and dynamical processes then ejected it from its birth place, producing a free-floating object.

astro-ph.EP↗