arXiv ScienceSearch

arXiv subjects

Yuanyuan Zhang

Publications and source records attributed to Yuanyuan Zhang.

At least 19 recordsLinked to original sources

Search for neutrinoless quadruple beta decay of $^{136}$Xe in PandaX-4T detector

The observation of neutrinoless quadruple beta decay (0$ν$4$β$) in the absence of neutrinoless double beta decay (0$ν$2$β$) has been argued to provide a strong indication that neutrinos are Dirac particles. We report a search for 0$ν$4$β$ decay of $^{136}\text{Xe}$ using a total $^{136}\text{Xe}$ exposure of 148.4 kg$\cdot$yr, collected during the commissioning and the first science runs of the PandaX-4T experiment. No significant excess of events over the background is observed. A lower limit on the 0$ν$4$β$ decay half-life of $^{136}\text{Xe}$ is set at 6.01 x $10^{24}$ yr at the 90% confidence level. This result establishes the most stringent constraint on this process in xenon, demonstrating the unique capability of the PandaX-4T detector in probing lepton number violation and shedding light on the fundamental nature of neutrinos.

nucl-ex

Compressing Streaming Neural Audio Encoders via Latent-Space Distillation

System-wide Dictation on Apple devices runs entirely on-device, and the speech it transcribes reaches the foundation model through a tokenizer: an encoder that maps short windows of waveform onto the representation the language model reads. Because that model is sparsely activated under Instruction-Following Pruning, only a small subset of its experts occupies DRAM at any time, so the always-on tokenizer competes for the same memory, and its parameter count bears directly on power and latency. In this work we study how to compress such a tokenizer by distillation, taking as the supervision target neither the discrete token nor the output distribution but the pre-quantizer latent the model actually consumes - the last representation the two token interfaces share. We train only the student encoder to regress the teacher's per-frame latent under a squared-error objective, with a single affine layer absorbing the teacher-student width mismatch. Because the target precedes both the quantizer and the language-model bridge, one recipe covers both token interfaces we support, and applies both to a tokenizer pretrained alone and to one jointly trained with a language model. At 2.8x compression the distilled student stays within 1.9% relative WER of its teacher on five of six teacher-student pairs without any fine-tuning, and improves on an independently trained tokenizer of identical capacity by 3.9% relative.

cs.SD

A Splashback-like Feature of Central Galaxies in Galaxy Clusters

We investigate a splashback-like feature in the outer region of central galaxies (CGs) in clusters. This feature is detected as a "dip" in the radial slope of the CG surface brightness, derived through the stacking of Dark Energy Survey data of over 4000 galaxy clusters in the redshift range of 0.2 to 0.5 with richness 20 and above. The local minimum of the dip occurs between 40 to 60 kpc from the CG center, with a mild dependence on cluster richness. This feature resembles the density transition caused by the splashback effect at the outskirts of galaxy clusters, when accreted matter reaches the apocenter for the first time. We turn to the IllustrisTNG hydro-dynamic simulation to gain theoretical insights. Density bumps, shells and accretion streaks are identified in the diffuse stellar content of the CGs and intra-cluster light which relate to the recent history of disruption and accretion. These features occur at the outskirts of the CGs, up to several hundred kiloparsecs from the cluster center. Thus, the location of the splashback-like dip in the data potentially marks the edge of the CG and the beginning of a region with the cluster diffuse light undergoing active or recent accretion.

astro-ph.GA

Physics-Constrained Deep Learning Model for Contactless Blood Pressure Monitoring from Triaxial Bodyseismography

Ballistocardiography (BCG) is promising for unobtrusive long-term blood pressure (BP) monitoring in laboratory settings, but traditional BCG signals are vulnerable to the variations in body-bed interaction with shifted fiducial points in temporal or amplitude axis, and BP varies with personal hemodynamic changes, causing misaligned representations that affect model generalizability and robustness. In this work, we propose a non-invasive BP estimation framework, Phy-BP, based on triaxial bodyseismography (BSG) as an extension of BCG. Firstly, an adaptive quality-control algorithm is designed to select BSG segments enriched with cardiogenic components by jointly considering neighboring beat patterns and universal cardiogenic templates. Furthermore, a physical model is established to describe 3D wave propagation in the body-bed system and is subsequently embedded into the deep learning model to characterize the intrinsic coupling among triaxial BSG signals driven by a single cardiogenic excitation. Thus, multi-axis features are aligned during model training, improving robustness against distortions in real scenarios. Experiments on a 162-hour hospital dataset collected from 21 subjects reveal that the proposed Phy-BP can dynamically filter out low-quality measurements, and the deep learning model training is constrained by physical consistency across different axes to provide faithful BP monitoring, especially when training samples are limited.

eess.SP

Disuccessors, Gröbner-Shirshov bases and free L-dendriform algebras

In this paper, we prove respectively that the disuccessor operations on the associative operad $\as$, the Lie operad $\lie$, and the pre-Lie operad $\prelie$ preserve Gröbner-Shirshov bases. This structural preservation enables the transfer of known bases to more complex operads. As a consequence, we introduce new methods for constructing Gröbner-Shirshov bases for the $\dend$ and $\prelie$ operads, and explicitly construct a Gröbner-Shirshov basis for the free L-dendriform algebra. This provides a conceptual and computationally efficient resolution of Madariaga's problem and offers new insights into the combinatorial structure of L-dendriform algebras.

math.RA

The Radioactive Background of the JUNO Calibration System

The Jiangmen Underground Neutrino Observatory (JUNO) experiment is a reactor antineutrino detector employing 20 kton of ultra-pure liquid scintillator to determine the neutrino mass ordering and to precisely measure oscillation parameters. The total singles background rate from radioactivity is required to be below 10 Hz in the energy range of 0.7-12 MeV within the fiducial volume for reactor neutrino analysis. The calibration system is designed to characterize the detector energy and position responses, while several of its components are located close to the target and may contribute to the background budget. Therefore, extensive material screening and selection are required to construct a low-background calibration system and to ensure that its contribution remains within the design requirements. In this work, a comprehensive study of the radioactive background induced by the calibration system is presented, including material radioactivity measurements using high-purity germanium detectors and neutron activation analysis techniques, detailed Monte Carlo simulations to evaluate the background, and comparisons with in-situ detector data to validate the predictions. In this data analysis, dedicated spatial selection methods are developed to isolate calibration-related contributions and to suppress the liquid scintillator background. The total radioactivity contribution from the calibration system is estimated to be less than 76 mHz, which satisfies the requirement of 200 mHz (2% of the total background budget). The results from in-situ data are found to be consistent with the expectations based on material assay and simulation within uncertainties. These results demonstrate that the calibration-induced background is well understood, in agreement between data and simulation, and negligible for reactor antineutrino measurements in JUNO.

physics.ins-det

The Cross-Survey Decade: A Call to Action

By 2027, three flagship wide-field surveys will be operating simultaneously from ground and space, observing overlapping sky and representing more than $6 billion in US and European public investment. Together they will produce overlapping petabyte-scale datasets across thousands of square degrees. This is a different class of challenge: the observations are no longer the bottleneck; realizing their joint scientific return now depends on shared computational infrastructure and coordination. Decades of community studies show that combining these datasets does more than improve precision. For science ranging from weak lensing to transient discovery and Galactic-plane astronomy, joint processing and analysis can unlock capabilities no single survey provides alone. Yet the required infrastructure -- joint pixel-level processing, cross-calibration and validation, interoperable data access, and the people to build and sustain it -- falls outside any single mission or institution's mandate. We issue a call to action for cross-survey science infrastructure, built around four pillars: (1) joint pixel-level processing and validation; (2) an AI-ready data substrate for scientific foundation models; (3) standardized, interoperable data access across surveys, democratizing participation in astrophysical discovery; and (4) dedicated personnel and career pathways. We outline concrete steps for policymakers, agencies, observatories, universities, the research community, and philanthropy, and argue that the moment to act is now, while foundational technical choices can still be aligned at a fraction of the cost of reconciling them later.

astro-ph.IM

NormAct: Benchmarking Embodied Agents' Proactive Compliance with Unspoken Social Norms

Embodied agents driven by multimodal large language models (MLLMs) can often complete everyday tasks from visual observations, but goal achievement does not establish whether they proactively respect unstated social norms. Existing benchmarks assess explicit norm judgments or constrained behavior, but rarely test whether agents infer and apply scene-relevant norms during ordinary tasks. We introduce NormAct, a benchmark of 550 TongSim scenarios in which the same goal permits norm-compliant or norm-violating action sequences. Norm-relevant evidence is embedded in each scenario while the applicable rule is omitted from the goal instruction. By progressively increasing normative guidance while holding the goal and scene fixed, NormAct tests whether compliant behavior emerges autonomously or only after prompting. Across three MLLM planners, goal achievement substantially exceeds norm compliance without guidance (67.4% versus 24.7%), while both broad and rule-specific guidance improve compliance, indicating that planners can often comply when prompted but not reliably on their own. With a fixed planner, general-norm retrieval is less effective than norm-relevant scene descriptions or generated norm cues, suggesting that identifying relevant visual evidence is a greater challenge than accessing general norm knowledge. NormAct therefore supports the development of embodied agents that pursue everyday goals while proactively respecting unstated social norms.

cs.AI

Benchmarking Human and Automatic Speech Recognition of Diverse Speech: Initial Results

Humans are often considered to be the best listeners and seen as the upper-bound performance of automatic speech recognition (ASR) systems. We present a preliminary comparison of the performances of state-of-the-art ASR systems and Dutch native listeners on the recognition of "diverse" speech, specifically Dutch child and older adults' speech and Flemish. Google Telephony outperformed the other ASR systems. Importantly, the ASR systems showed similar performance to the listeners, and in specific cases even outperformed them. Slight performance differences between the listeners and ASR systems were found related to speaker's age and regional accents and utterance length. Future research should focus on making ASR systems more robust to acoustic variability related to aging and regional accents. A comparison of ASR recognition performances on the test stimuli and the full Jasmin-CGN test sets showed the influence of the specific test sets on the conclusions regarding benchmarking human and ASR performance.

cs.CL

A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition

We present DRES: a 1.5-hour Dutch realistic elicited (semi-spontaneous) speech dataset from 80 speakers recorded in noisy, public indoor environments. DRES was designed as a test set for the evaluation of state-of-the-art (SotA) automatic speech recognition (ASR) and speech enhancement (SE) models in a real-world scenario: a person speaking in a public indoor space with background talkers and noise. The speech was recorded with a four-channel linear microphone array. In this work we evaluate the speech quality of five well-known single-channel SE algorithms and the recognition performance of eight SotA off-the-shelf ASR models before and after applying SE on the speech of DRES. We found that five out of the eight ASR models have WERs lower than 22\% on DRES, despite the challenging conditions. In contrast to recent work, we did not find a positive effect of modern single-channel SE on ASR performance, emphasizing the importance of evaluating in realistic conditions.

eess.AS

Measurement of solar $pp$ neutrino flux with the new PandaX-4T data

We report a new measurement of the solar proton--proton ($pp$) neutrino flux via neutrino--electron elastic scattering using the PandaX-4T Run 2 data set collected between 2024 and 2026, corresponding to an exposure of 1.9 tonne$\cdot$yr. Before Run 2 data taking, the detector underwent a series of upgrades to improve its response and background conditions. Time variations of radioactive noble-gas impurities are constrained using the physics data themselves, complemented by measurements from the gas-assay system. The analysis introduced improvements in the data processing chain, detector response characterization, and background models. A blind spectral analysis was then performed on the electronic-recoil data across a wide energy range from 20 to 1000 keV. In combination with the Run 0 data published earlier, the fitted $pp$ flux is $(8.5 \pm 3.5)\times 10^{10}$ $\mathrm{cm^{-2}s^{-1}}$, consistent with the prediction of the Standard Solar Model. With a statistical significance of $2.2σ$ above background, this marks the first positive indication of solar $pp$ neutrino--electron scattering below an electronic-recoil energy of 165 keV.

hep-ex

Comparing Human and Automatic Recognition of Dutch Dysarthric Continuous Speech: A Case Study

In our goal to develop personalised dysarthric speech recognition (DSR) models, this study compared the recognition performances of human listeners and those of three state-of-the-art, off-the-shelf ASR systems (Whisper-large-V3, Google Chirp 3, and Omnilingual) on the recognition of Dutch continuous read and spontaneous speech from a single speaker with severe dysarthria. Results showed that both humans listeners and the three off-the-shelf ASR systems exhibit word error rates (WER) exceeding 70% on average, indicating that DSR is highly challenging for both humans and ASR systems. Fine-tuning on the dysarthric speech significantly reduced WER. Although overall WERs are still quite high (>23%), the personalised DSR models outperformed the human listeners, and performance is getting closer to being useful for supporting day-to-day communication of dysarthric speakers. Future research should focus on improving personalized DSR on spontaneous speech and longer utterances in the case of read speech, with a specific focus on particular phonemes.

cs.CL

Nijenhuis BiHom-Lie bialgebras and differential Lie bialgebras

In this paper, we first introduce the concept of Nijenhuis BiHom-Lie algebras. We then establish the equivalence relations between the Manin triples of Nijenhuis BiHom-Lie algebras, Nijenhuis BiHom-Lie bialgebras, and matched pairs of Nijenhuis BiHom-Lie algebras. Furthermore, we show that such an equivalence also holds for differential Lie bialgebras, together with their associated Manin triples and corresponding matched pairs.

math.RA

Contract Structure and Risk Aversion in Longevity Risk Transfers

This paper introduces an economic framework to assess optimal longevity risk transfers between institutions, focusing on the interactions between a buyer exposed to long-term longevity risk and a seller offering longevity protection. While most longevity risk transfers have occurred in the reinsurance sector, where global reinsurers provide long-term protections, the capital market for longevity risk transfer has struggled to gain traction, resulting in only a few short-term instruments. We investigate how differences in risk aversion between the two parties affect the equilibrium structure of longevity risk transfer contracts, contrasting `static' contracts that offer long-term protection with `dynamic' contracts that provide short-term, variable coverage. Our analysis shows that static contracts are preferred by more risk-averse buyers, while dynamic contracts are favored by more risk-averse sellers who are reluctant to commit to long-term agreements. When incorporating information asymmetry through ambiguity, we find that ambiguity can cause more risk-averse sellers to stop offering long-term contracts. With the assumption that global reinsurers, acting as sellers in the reinsurance sector and buyers in the capital market, are generally less risk-averse than other participants, our findings provide theoretical explanations for current market dynamics and suggest that short-term instruments offer valuable initial steps toward developing an efficient and active capital market for longevity risk transfer.

econ.GN

Constraining Galaxy Cluster Triaxiality via Weak Lensing -- I. Preparation for the Rubin Data Beyond Leading Order

The 3D mass distributions of galaxy clusters are generally triaxial, a geometry that is difficult to constrain from projected observations. In this work, we measure the projected halo shapes of clusters from their weak lensing signatures using the triaxiality functionality in the Cluster Lensing Mass Modeling software, a tool developed by the Dark Energy Science Collaboration to analyze data from NSF-DOE Rubin Observatory's Legacy Survey of Space and Time (LSST). We measure ensemble halo ellipticity on the plane of the sky via axis-aligned stacking and multipole expansion of the weak lensing data. We study a precursor dataset -- the redMaPPer cluster catalog, the metacalibration shape catalog, and the Directional Neighborhood Fitting photometric redshift catalog from the Dark Energy Survey Year 3 public data release. We select clusters that have a high centering probability (>90%) of the identified central galaxy, and use the satellite galaxy distribution to determine the major-axis orientation for stacking. We extend the analysis to the second order of ellipticity in the monopole and quadrupole measurement. The projected ellipticity of the cluster sample is found to be $0.310^{+0.017}_{-0.016}$ (axis ratio $0.527^{+0.018}_{-0.019}$). The projected cluster ellipticity shows no statistically significant dependence on mass and redshift. We further verify the accuracy of the cluster shape measurement using mock catalogs. This analysis is applicable to datasets from upcoming wide-area cosmic surveys such as LSST, Euclid, and the Roman Space Telescope, where larger sample sizes will lead to tighter constraints on the cluster ellipticities.

astro-ph.CO

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some studies involve heavy frameworks requiring extensive training, posing challenges for convenient and efficient assessment of docking methods. To fill these gaps, we design PoseX, an open-source benchmark to evaluate both self-docking and cross-docking, enabling a practical and comprehensive assessment of algorithmic advances. Specifically, we curated a novel dataset comprising 718 entries for self-docking and 1,312 entries for cross-docking; second, we incorporated 23 docking methods in three methodological categories, including physics-based methods (e.g., Schrödinger Glide), AI docking methods (e.g., DiffDock) and AI co-folding methods (e.g., AlphaFold3); third, we developed a relaxation method for post-processing to minimize conformational energy and refine binding poses; fourth, we built a leaderboard to rank submitted models in real-time. We derived some key insights and conclusions from extensive experiments: (1) AI approaches have consistently outperformed physics-based methods in overall docking success rate. (2) Most intra- and intermolecular clashes of AI approaches can be greatly alleviated with relaxation, which means combining AI modeling with physics-based post-processing could achieve excellent performance. (3) AI co-folding methods exhibit ligand chirality issues, except for Boltz-1x, which introduced physics-inspired potentials to fix hallucinations, suggesting modeling on stereochemistry improves the structural plausibility markedly. (4) Specifying binding pockets significantly promotes docking performance, indicating that pocket information can be leveraged adequately, particularly for AI co-folding methods, in future modeling efforts. The code, dataset, and leaderboard are released at https://github.com/CataAI/PoseX.

cs.LG

Averaging pre-Lie bialgebras

In this paper, we first introduce representations of averaging pre-Lie algebras and study their matched pairs, Manin triples, and bialgebra theories. We prove that these three notions are equivalent under certain conditions. Moreover, by introducing averaging operators on quadratic Rota-Baxter pre-Lie algebras, we show that such operators give rise to averaging pre-Lie bialgebras. Then we introduce the notion of admissible classical Yang-Baxter equations in averaging pre-Lie algebras, as well as the relative Rota-Baxter operators on averaging pre-Lie algebras, and show that the relative Rota-Baxter operators on averaging pre-Lie algebras yield symmetric solutions of admissible classical Yang-Baxter equations in averaging pre-Lie algebras. Finally, we show that every averaging pre-Lie bialgebra induces an averaging Lie bialgebra.

math.RA

Retrieval-Reasoning Large Language Model-based Synthetic Clinical Trial Generation

Machine learning (ML) holds great promise for clinical applications but is often hindered by limited access to high-quality data due to privacy concerns, high costs, and long timelines associated with clinical trials. While large language models (LLMs) have demonstrated strong performance in general-purpose generation tasks, their application to synthesizing realistic clinical trials remains underexplored. In this work, we propose a novel Retrieval-Reasoning framework that leverages few-shot prompting with LLMs to generate synthetic clinical trial reports annotated with binary success/failure outcomes. Our approach integrates a retrieval module to ground the generation on relevant trial data and a reasoning module to ensure domain-consistent justifications. Experiments conducted on real clinical trials from the ClinicalTrials.gov database demonstrate that the generated synthetic trials effectively augment real datasets. Fine-tuning a BioBERT classifier on synthetic data, real data, or their combination shows that hybrid fine-tuning leads to improved performance on clinical trial outcome prediction tasks. Our results suggest that LLM-based synthetic data can serve as a powerful tool for privacy-preserving data augmentation in clinical research. The code is available at https://github.com/XuZR3x/Retrieval_Reasoning_Clinical_Trial_Generation.

cs.CL