arXiv Science⌕ Search

arXiv subjects

:

Publications and source records attributed to :.

At least 127 records · Page 7Linked to original sources

Measurement of reactor antineutrino oscillation at SNO+

The SNO+ collaboration reports its second spectral analysis of reactor antineutrino oscillation using 286 tonne-years of new data. The measured energies of reactor antineutrino candidates were fitted to obtain the second-most precise determination of the neutrino mass-squared difference $Δm^2_{21}$ = ($7.96^{+0.48}_{-0.42}$) $\times$ 10$^{-5}$ eV$^2$. Constraining $Δm^2_{21}$ and $\sin^2θ_{12}$ with measurements from long-baseline reactor antineutrino and solar neutrino experiments yields $Δm^2_{21}$ = ($7.58^{+0.18}_{-0.17}$) $\times$ 10$^{-5}$ eV$^2$ and $\sin^2θ_{12} = 0.308 \pm 0.013$. This fit also yields a first measurement of the flux of geoneutrinos in the Western Hemisphere, with $73^{+47}_{-43}$ TNU at SNO+.

hep-ex↗

On Soft Clustering For Correlation Estimators

Properly estimating correlations between objects at different spatial scales necessitates $\mathcal{O}(n^2)$ distance calculations. For this reason, most widely adopted packages for estimating correlations use clustering algorithms to approximate local trends. However, methods for quantifying the error introduced by this clustering have been understudied. In response, we present an algorithm for estimating correlations that is probabilistic in the way that it clusters objects, enabling us to quantify the uncertainty caused by clustering simply through model inference. These soft clustering assignments enable correlation estimators that are theoretically differentiable with respect to their input catalogs. Thus, we also build a theoretical framework for differentiable correlation functions and describe their utility in comparison to existing surrogate models. Notably, we find that repeated normalization and distance function calls slow gradient calculations and that sparse Jacobians destabilize precision, pointing towards either approximate or surrogate methods as a necessary solution to exact gradients from correlation functions. To that end, we close with a discussion of surrogate models as proxies for correlation functions. We provide an example that demonstrates the efficacy of surrogate models to enable gradient-based optimization of astrophysical model parameters, successfully minimizing a correlation function output. Our numerical experiments cover science cases across cosmology, from point spread function (PSF) modeling efforts to gravitational simulations to galaxy intrinsic alignment (IA).

astro-ph.IM↗

Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models

As inference-time scaling becomes critical for enhanced reasoning capabilities, it is increasingly becoming important to build models that are efficient to infer. We introduce Nemotron-H, a family of 8B and 56B/47B hybrid Mamba-Transformer models designed to reduce inference cost for a given accuracy level. To achieve this goal, we replace the majority of self-attention layers in the common Transformer model architecture with Mamba layers that perform constant computation and require constant memory per generated token. We show that Nemotron-H models offer either better or on-par accuracy compared to other similarly-sized state-of-the-art open-sourced Transformer models (e.g., Qwen-2.5-7B/72B and Llama-3.1-8B/70B), while being up to 3$\times$ faster at inference. To further increase inference speed and reduce the memory required at inference time, we created Nemotron-H-47B-Base from the 56B model using a new compression via pruning and distillation technique called MiniPuzzle. Nemotron-H-47B-Base achieves similar accuracy to the 56B model, but is 20% faster to infer. In addition, we introduce an FP8-based training recipe and show that it can achieve on par results with BF16-based training. This recipe is used to train the 56B model. We are releasing Nemotron-H base model checkpoints with support in Hugging Face and NeMo.

cs.CL↗

NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

We introduce Nemotron-Nano-9B-v2, a hybrid Mamba-Transformer language model designed to increase throughput for reasoning workloads while achieving state-of-the-art accuracy compared to similarly-sized models. Nemotron-Nano-9B-v2 builds on the Nemotron-H architecture, in which the majority of the self-attention layers in the common Transformer architecture are replaced with Mamba-2 layers, to achieve improved inference speed when generating the long thinking traces needed for reasoning. We create Nemotron-Nano-9B-v2 by first pre-training a 12-billion-parameter model (Nemotron-Nano-12B-v2-Base) on 20 trillion tokens using an FP8 training recipe. After aligning Nemotron-Nano-12B-v2-Base, we employ the Minitron strategy to compress and distill the model with the goal of enabling inference on up to 128k tokens on a single NVIDIA A10G GPU (22GiB of memory, bfloat16 precision). Compared to existing similarly-sized models (e.g., Qwen3-8B), we show that Nemotron-Nano-9B-v2 achieves on-par or better accuracy on reasoning benchmarks while achieving up to 6x higher inference throughput in reasoning settings like 8k input and 16k output tokens. We are releasing Nemotron-Nano-9B-v2, Nemotron-Nano12B-v2-Base, and Nemotron-Nano-9B-v2-Base checkpoints along with the majority of our pre- and post-training datasets on Hugging Face.

cs.CL↗

Constraining the TeV gamma-ray emission of SN 2024bch, a possible type IIn-L from a red supergiant progenitor. Multiwavelength observations and analysis of the progenitor

We present very high-energy optical photometry and spectroscopic observations of SN 2024bch in the nearby galaxy NGC 3206 (\sim 20 Mpc). We used gamma-ray observations performed with the first Large-Sized Telescope (LST-1) of the Cherenkov Telescope Array Observatory (CTAO) and optical observations with the Liverpool Telescope (LT) combined with data from public repositories to evaluate the general properties of the event and the progenitor star. No significant emission above the LST-1 energy threshold for this observation (\sim 100 GeV) was detected in the direction of SN 2024bch, and we computed an integral upper limit on the photon flux of F_γ(>100 GeV) \le 3.61 \times 10^{-12} cm^{-2} s^{-1} based on six nonconsecutive nights of observations with the LST-1, between 16 and 38 days after the explosion. Employing a general model for the gamma-ray flux emission, we found an upper limit on the mass-loss-rate to wind-velocity ratio of \dot M/u_{w} \le 10^{-4} \frac{M_\odot}{yr}\frac{s}{km}, although gamma-gamma absorption could potentially have skewed this estimation, effectively weakening our constraint. From spectro-photometric observations we found progenitor parameters of M_{pr} = 11 - 20 M_\odot and R_{pr} = 531 \pm 125 R_\odot. Finally, using archival images from the Hubble Space Telescope, we constrained the luminosity of the progenitor star to log(L_{pr}/L_\odot) \le 4.82 and its effective temperature to T_{pr} \le 4000 K. Our results suggest that SN 2024bch is a type IIn-L supernova that originated from a progenitor star consistent with a red supergiant. We show how the correct estimation of the mass-loss history of a supernova will play a major role in future multiwavelength observations.

astro-ph.HE↗

Combined dark matter search towards dwarf spheroidal galaxies with Fermi-LAT, HAWC, H.E.S.S., MAGIC, and VERITAS

Dwarf spheroidal galaxies (dSphs) are excellent targets for indirect dark matter (DM) searches using gamma-ray telescopes because they are thought to have high DM content and a low astrophysical background. The sensitivity of these searches is improved by combining the observations of dSphs made by different gamma-ray telescopes. We present the results of a combined search by the most sensitive currently operating gamma-ray telescopes, namely: the satellite-borne Fermi-LAT telescope; the ground-based imaging atmospheric Cherenkov telescope arrays H.E.S.S., MAGIC, and VERITAS; and the HAWC water Cherenkov detector. Individual datasets were analyzed using a common statistical approach. Results were subsequently combined via a global joint likelihood analysis. We obtain constraints on the velocity-weighted cross section $\langle σ\mathit{v} \rangle$ for DM self-annihilation as a function of the DM particle mass. This five-instrument combination allows the derivation of up to 2-3 times more constraining upper limits on $\langle σ\mathit{v} \rangle$ than the individual results over a wide mass range spanning from 5 GeV to 100 TeV. Depending on the DM content modeling, the 95% confidence level observed limits reach $1.5\times$10$^{-24}$ cm$^3$s$^{-1}$ and $3.2\times$10$^{-25}$ cm$^3$s$^{-1}$, respectively, in the $τ^+τ^-$ annihilation channel for a DM mass of 2 TeV.

astro-ph.HE↗

HAWC, VERITAS, Fermi-LAT and XMM-Newton follow-up observations of the unidentified ultra-high-energy gamma-ray source LHAASO J2108+5157

We report observations of the ultra-high-energy gamma-ray source LHAASO J2108$+$5157, utilizing VERITAS, HAWC, Fermi-LAT, and XMM-Newton. VERITAS has collected $\sim$ 40 hours of data that we used to set ULs to the emission above 200 GeV. The HAWC data, collected over $\sim 2400$ days, reveal emission between 3 and 146 TeV, with a significance of $7.5~σ$, favoring an extended source model. The best-fit spectrum measured by HAWC is characterized by a simple power-law with a spectral index of $2.45\pm0.11_{stat}$. Fermi-LAT analysis finds a point source with a very soft spectrum in the LHAASO J2108+5157 region, consistent with the 4FGL-DR3 catalog results. The XMM-Newton analysis yields a null detection of the source in the 2 - 7 keV band. The broadband spectrum can be interpreted as a pulsar and a pulsar wind nebula system, where the GeV gamma-ray emission originates from an unidentified pulsar, and the X-ray and TeV emission is attributed to synchrotron radiation and inverse Compton scattering of electrons accelerated within a pulsar wind nebula. In this leptonic scenario, our X-ray upper limit provides a stringent constraint on the magnetic field, which is $\lesssim 1.5\ μ$G.

astro-ph.HE↗

Audio2Face-3D: Audio-driven Realistic Facial Animation For Digital Avatars

Audio-driven facial animation presents an effective solution for animating digital avatars. In this paper, we detail the technical aspects of NVIDIA Audio2Face-3D, including data acquisition, network architecture, retargeting methodology, evaluation metrics, and use cases. Audio2Face-3D system enables real-time interaction between human users and interactive avatars, facilitating facial animation authoring for game characters. To assist digital avatar creators and game developers in generating realistic facial animations, we have open-sourced Audio2Face-3D networks, SDK, training framework, and example dataset.

cs.GR↗

BeyondWeb: Lessons from Scaling Synthetic Data for Trillion-scale Pretraining

Recent advances in large language model (LLM) pretraining have shown that simply scaling data quantity eventually leads to diminishing returns, hitting a data wall. In response, the use of synthetic data for pretraining has emerged as a promising paradigm for pushing the frontier of performance. Despite this, the factors affecting synthetic data quality remain poorly understood. In this work, we introduce BeyondWeb, a synthetic data generation framework that produces high-quality synthetic data for pretraining. BeyondWeb significantly extends the capabilities of traditional web-scale datasets, outperforming state-of-the-art synthetic pretraining datasets such as Cosmopedia and Nemotron-CC's high-quality synthetic subset (Nemotron-Synth) by up to 5.1 percentage points (pp) and 2.6pp, respectively, when averaged across a suite of 14 benchmark evaluations. It delivers up to 7.7x faster training than open web data and 2.7x faster than Nemotron-Synth. Remarkably, a 3B model trained for 180B tokens on BeyondWeb outperforms an 8B model trained for the same token budget on Cosmopedia. We also present several insights from BeyondWeb on synthetic data for pretraining: what drives its benefits, which data to rephrase and how, and the impact of model size and family on data quality. Overall, our work shows that there's no silver bullet for generating high-quality synthetic pretraining data. The best outcomes require jointly optimizing many factors, a challenging task that requires rigorous science and practical expertise. Naive approaches can yield modest improvements, potentially at great cost, while well-executed methods can yield transformative improvements, as exemplified by BeyondWeb.

cs.LG↗

gpt-oss-120b & gpt-oss-20b Model Card

We present gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models that push the frontier of accuracy and inference cost. The models use an efficient mixture-of-expert transformer architecture and are trained using large-scale distillation and reinforcement learning. We optimize the models to have strong agentic capabilities (deep research browsing, python tool use, and support for developer-provided functions), all while using a rendered chat format that enables clear instruction following and role delineation. Both models achieve strong results on benchmarks ranging from mathematics, coding, and safety. We release the model weights, inference implementations, tool environments, and tokenizers under an Apache 2.0 license to enable broad use and further research.

cs.CL↗

SafeWork-R1: Coevolving Safety and Intelligence under the AI-45$^{\circ}$ Law

We introduce SafeWork-R1, a cutting-edge multimodal reasoning model that demonstrates the coevolution of capabilities and safety. It is developed by our proposed SafeLadder framework, which incorporates large-scale, progressive, safety-oriented reinforcement learning post-training, supported by a suite of multi-principled verifiers. Unlike previous alignment methods such as RLHF that simply learn human preferences, SafeLadder enables SafeWork-R1 to develop intrinsic safety reasoning and self-reflection abilities, giving rise to safety `aha' moments. Notably, SafeWork-R1 achieves an average improvement of $46.54\%$ over its base model Qwen2.5-VL-72B on safety-related benchmarks without compromising general capabilities, and delivers state-of-the-art safety performance compared to leading proprietary models such as GPT-4.1 and Claude Opus 4. To further bolster its reliability, we implement two distinct inference-time intervention methods and a deliberative search mechanism, enforcing step-level verification. Finally, we further develop SafeWork-R1-InternVL3-78B, SafeWork-R1-DeepSeek-70B, and SafeWork-R1-Qwen2.5VL-7B. All resulting models demonstrate that safety and capability can co-evolve synergistically, highlighting the generalizability of our framework in building robust, reliable, and trustworthy general-purpose AI.

cs.AI↗

Multi-wavelength Study of HESS J0632+057: New Insights into Pulsar-Disk Interaction

We present an analysis of new multi-wavelength observations of the TeV gamma-ray binary HESS J0632+057, conducted using SALT, Swift, NuSTAR, and VERITAS in 2023--2024. By combining these new data with archival observations, we confirm previous suggestions of orbital variability in the source's X-ray spectrum, including increased X-ray absorption at the orbital phase interval of $ϕ\approx0.3\textrm{--}0.4$. The source's X-ray flux within this phase interval seems to have exhibited a significant change on an orbital timescale. Additionally, occasional short-term variations in the X-ray band on a timescale of less than 3 days have been observed. The measured duration of the increased absorbing column density and the flux variability timescales can provide clues about the interaction between the putative pulsar and the Be companion's disk if, as previously suggested, the pulsar crosses the disk at this phase interval. Moreover, the new contemporaneous X-ray and TeV observations around the pulsar-crossing phases revealed independent variability in the X-ray and TeV fluxes, contrary to a previous observation of concurrent flux increases. While these observations alone cannot provide definitive conclusions, we discuss our results in the context of pulsar-disk interaction and intrabinary shock emission scenarios.

astro-ph.HE↗

Measurements of the branching fractions of $Ξ_{c}^{+}\to Σ^{+}K_{S}^{0}$, $Ξ_{c}^{+}\to Ξ^{0}π^{+}$, and $Ξ_{c}^{+}\to Ξ^{0}K^{+}$ at Belle and Belle II

Using 983.0 $\rm{fb}^{-1}$ and 427.9 $\rm{fb}^{-1}$ data samples collected with the Belle and Belle II detectors at the KEKB and SuperKEKB asymmetric energy $e^+e^-$ colliders, respectively, we present studies of the Cabibbo-favored $Ξ_c^+$ decays ${Ξ_{c}^{+}\to Σ^{+}K_{S}^{0}}$ and $Ξ_{c}^{+}\to Ξ^{0}π^{+}$, and the singly Cabibbo-suppressed decay $Ξ_{c}^{+}\to Ξ^{0}K^{+}$. The ratios of branching fractions of ${Ξ_{c}^{+}\to Σ^{+}K_{S}^{0}}$ and $Ξ_{c}^{+}\to Ξ^{0}K^{+}$ relative to that of $Ξ_{c}^{+}\toΞ^{-}π^{+}π^{+}$ are measured for the first time, while the ratio ${\cal B}(Ξ_{c}^{+}\toΞ^{0}π^{+})/{\cal B}(Ξ_{c}^{+}\toΞ^{-}π^{+}π^{+}) $ is also determined and improved by an order of magnitude in precision. The measured branching fraction ratios are $\frac{\cal{B}(Ξ_{c}^{+} \to Σ^{+}K_{S}^{0})}{\cal{B}(Ξ_{c}^{+}\to Ξ^{-}π^{+}π^+)}= 0.067 \pm 0.007 \pm 0.003$, $\frac{\cal{B}(Ξ_c^{+} \to Ξ^{0}π^{+})}{\cal{B}(Ξ_{c}^{+}\to Ξ^{-}π^{+}π^+)} = 0.251 \pm 0.005 \pm 0.010$, $\frac{\cal{B}(Ξ_c^{+} \to Ξ^{0}K^{+})}{\cal{B}(Ξ_{c}^{+}\to Ξ^{-}π^{+}π^+)} = 0.017 \pm 0.003 \pm 0.001$. Additionally, the ratio ${\cal B}(Ξ_{c}^{+}\toΞ^{0}K^{+})/{\cal B}(Ξ_{c}^{+}\toΞ^{0}π^{+})$ is measured to be $ 0.068 \pm 0.010 \pm 0.004$. Here, the first and second uncertainties are statistical and systematic, respectively. Multiplying the ratios by the branching fraction of the normalization mode, ${\mathcal B}(Ξ_{c}^{+}\toΞ^{-}π^{+}π^+)= (2.9\pm 1.3)\%$, we obtain the following absolute branching fractions ${\cal B}(Ξ_{c}^{+}\toΣ^{+}K^{0}_{S}) = (0.194 \pm 0.021 \pm 0.009 \pm 0.087 )%$, ${\cal B}(Ξ_{c}^{+}\toΞ^{0}π^{+}) = (0.728 \pm 0.014 \pm 0.027 \pm 0.326 )%$, ${\cal B}(Ξ_{c}^{+}\toΞ^{0}K^{+}) = (0.049 \pm 0.007 \pm 0.003 \pm 0.022 )%$.

hep-ex↗

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report

To understand and identify the unprecedented risks posed by rapidly advancing artificial intelligence (AI) models, this report presents a comprehensive assessment of their frontier risks. Drawing on the E-T-C analysis (deployment environment, threat source, enabling capability) from the Frontier AI Risk Management Framework (v1.0) (SafeWork-F1-Framework), we identify critical risks in seven areas: cyber offense, biological and chemical risks, persuasion and manipulation, uncontrolled autonomous AI R\&D, strategic deception and scheming, self-replication, and collusion. Guided by the "AI-$45^\circ$ Law," we evaluate these risks using "red lines" (intolerable thresholds) and "yellow lines" (early warning indicators) to define risk zones: green (manageable risk for routine deployment and continuous monitoring), yellow (requiring strengthened mitigations and controlled deployment), and red (necessitating suspension of development and/or deployment). Experimental results show that all recent frontier AI models reside in green and yellow zones, without crossing red lines. Specifically, no evaluated models cross the yellow line for cyber offense or uncontrolled AI R\&D risks. For self-replication, and strategic deception and scheming, most models remain in the green zone, except for certain reasoning models in the yellow zone. In persuasion and manipulation, most models are in the yellow zone due to their effective influence on humans. For biological and chemical risks, we are unable to rule out the possibility of most models residing in the yellow zone, although detailed threat modeling and in-depth assessment are required to make further claims. This work reflects our current understanding of AI frontier risks and urges collective action to mitigate these challenges.

cs.AI↗

Step-3 is Large yet Affordable: Model-system Co-design for Cost-effective Decoding

Large language models (LLMs) face low hardware efficiency during decoding, especially for long-context reasoning tasks. This paper introduces Step-3, a 321B-parameter VLM with hardware-aware model-system co-design optimized for minimizing decoding costs. Step-3 innovates in two key dimensions: (1) A novel Multi-Matrix Factorization Attention (MFA) mechanism that significantly reduces both KV cache size and computation while maintaining high attention expressiveness, and (2) Attention-FFN Disaggregation (AFD), a distributed inference system that decouples attention and Feed-Forward Network (FFN) layers into specialized subsystems. This co-design achieves unprecedented cost efficiency: Step-3 significantly reduces theoretical decoding costs compared with models like DeepSeek-V3 and Qwen3 MoE 235B, with the gains widening at longer context. Step-3 achieves low cost while activating 38B parameters per token (more than DeepSeek-V3 and Qwen3 MoE 235B), demonstrating that hardware-aligned attention arithmetic intensity, MoE sparsity, and AFD are critical to cost-effectiveness. We perform a head-to-head comparison with DeepSeek-V3 in its favorable scenarios. Our implementation on Hopper GPUs achieves a decoding throughput of up to 4,039 tokens per second per GPU under 50ms TPOT SLA (4K context, FP8, no MTP). It is higher than DeepSeek-V3's 2,324 in the same setup and sets a new Pareto frontier for LLM decoding.

cs.LG↗

Probing Long-Range Forces in Neutrino Oscillations at the ESSnuSB Experiment

Neutrino oscillations constitute an excellent tool to probe physics beyond the Standard Model. In this paper, we investigate the potential of the ESSnuSB experiment to constrain the effects of flavour-dependent long-range forces (LRFs) in neutrino oscillations, which may arise due to the extension of the Standard Model gauge group by introducing new $U(1)$ symmetries. Focusing on three specific $U(1)$ symmetries -- $L_e - L_μ$, $L_e - L_τ$, and $L_μ- L_τ$, we demonstrate that ESSnuSB offers a favourable environment to search for LRF effects. Our analyses reveal that ESSnuSB can set $90\%$ confidence level bounds of $V_{eμ} < 2.99 \times 10^{-14} \, \text{eV}$, $V_{eτ} < 2.05 \times 10^{-14} \, \text{eV}$, and $V_{μτ} < 1.81 \times 10^{-14} \, \text{eV}$, which are competitive to the upcoming Deep Underground Neutrino Experiment (DUNE). It is also observed that reducing the systematic uncertainties from $5\%$ to $2\%$ improves the ESSnuSB limits on $V_{αβ}$. Interestingly, we find limited correlations between LRF parameters and the less constrained lepton mixing parameters $θ_{23}$ and $δ_{\text{CP}}$, preserving the robustness of ESSnuSB's sensitivity to CP violation. Even under extreme LRF potentials ($V_{αβ} \gg 10^{-13} \, \text{eV}$), the CP-violation sensitivity and $δ_{\text{CP}}$ precision remain largely unaffected. These results establish ESSnuSB as a competitive experimental setup for probing LRF effects, complementing constraints from other neutrino sources and offering critical insights into the physics of long-range forces.

hep-ph↗

M2-Reasoning: Empowering MLLMs with Unified General and Spatial Reasoning

Recent advancements in Multimodal Large Language Models (MLLMs), particularly through Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced their reasoning abilities. However, a critical gap persists: these models struggle with dynamic spatial interactions, a capability essential for real-world applications. To bridge this gap, we introduce M2-Reasoning-7B, a model designed to excel in both general and spatial reasoning. Our approach integrates two key innovations: (1) a novel data pipeline that generates 294.2K high-quality data samples (168K for cold-start fine-tuning and 126.2K for RLVR), which feature logically coherent reasoning trajectories and have undergone comprehensive assessment; and (2) a dynamic multi-task training strategy with step-wise optimization to mitigate conflicts between data, and task-specific rewards for delivering tailored incentive signals. This combination of curated data and advanced training allows M2-Reasoning-7B to set a new state-of-the-art (SOTA) across 8 benchmarks, showcasing superior performance in both general and spatial reasoning domains.

cs.AI↗

Combining Hybrid and Opaque Scintillator Techniques in the Search for Double Beta Plus Decays

Double beta plus decay is a rare nuclear disintegration process. Difficulties in its measurement arise from suppressed decay probabilities, experimentally challenging decay signatures and low natural abundances of suitable candidate nuclei. In this article, we propose a new detector concept to overcome these challenges. It is based on the first-time combination of hybrid and opaque scintillation detector technology paired with novel light read-out techniques. This approach is particularly suitable detecting positron (beta plus) signatures. We expect to discover two-neutrino double beta plus decay modes within 1 tonne-week exposure and are able to probe neutrinoless double beta plus decays at several orders of magnitude improved significance compared to current experimental limits.

physics.ins-det↗