arXiv Science⌕ Search

arXiv subjects

:

Publications and source records attributed to :.

At least 181 records · Page 10Linked to original sources

Qwen2.5 Technical Report

In this report, we introduce Qwen2.5, a comprehensive series of large language models (LLMs) designed to meet diverse needs. Compared to previous iterations, Qwen 2.5 has been significantly improved during both the pre-training and post-training stages. In terms of pre-training, we have scaled the high-quality pre-training datasets from the previous 7 trillion tokens to 18 trillion tokens. This provides a strong foundation for common sense, expert knowledge, and reasoning capabilities. In terms of post-training, we implement intricate supervised finetuning with over 1 million samples, as well as multistage reinforcement learning. Post-training techniques enhance human preference, and notably improve long text generation, structural data analysis, and instruction following. To handle diverse and varied use cases effectively, we present Qwen2.5 LLM series in rich sizes. Open-weight offerings include base and instruction-tuned models, with quantized versions available. In addition, for hosted solutions, the proprietary models currently include two mixture-of-experts (MoE) variants: Qwen2.5-Turbo and Qwen2.5-Plus, both available from Alibaba Cloud Model Studio. Qwen2.5 has demonstrated top-tier performance on a wide range of benchmarks evaluating language understanding, reasoning, mathematics, coding, human preference alignment, etc. Specifically, the open-weight flagship Qwen2.5-72B-Instruct outperforms a number of open and proprietary models and demonstrates competitive performance to the state-of-the-art open-weight model, Llama-3-405B-Instruct, which is around 5 times larger. Qwen2.5-Turbo and Qwen2.5-Plus offer superior cost-effectiveness while performing competitively against GPT-4o-mini and GPT-4o respectively. Additionally, as the foundation, Qwen2.5 models have been instrumental in training specialized models such as Qwen2.5-Math, Qwen2.5-Coder, QwQ, and multimodal models.

cs.CL↗

LLM-jp: A Cross-organizational Project for the Research and Development of Fully Open Japanese LLMs

This paper introduces LLM-jp, a cross-organizational project for the research and development of Japanese large language models (LLMs). LLM-jp aims to develop open-source and strong Japanese LLMs, and as of this writing, more than 1,500 participants from academia and industry are working together for this purpose. This paper presents the background of the establishment of LLM-jp, summaries of its activities, and technical reports on the LLMs developed by LLM-jp. For the latest activities, visit https://llm-jp.nii.ac.jp/en/.

cs.CL↗

Imagen 3

We introduce Imagen 3, a latent diffusion model that generates high quality images from text prompts. We describe our quality and responsibility evaluations. Imagen 3 is preferred over other state-of-the-art (SOTA) models at the time of evaluation. In addition, we discuss issues around safety and representation, as well as methods we used to minimize the potential harm of our models.

cs.CV↗

Search for lepton flavor-violating decay modes $B^0\to K_S^0τ^\pm\ell^\mp~(\ell=μ, e)$ with hadronic $B$-tagging at Belle and Belle II

We present the first search for the lepton flavor-violating decay modes $B^0 \rightarrow K_S^0 τ^\pm \ell^\mp~(\ell=μ, e)$ using the 711 fb$^{-1}$ and 365 fb$^{-1}$ data samples recorded by the Belle and Belle II detectors, respectively. We use a hadronic $B$-tagging technique, and search for the signal decay in the system recoiling against the fully reconstructed $B$ meson. We find no evidence for $B^0 \rightarrow K_S^0 τ^\pm \ell^\mp$ decays and set 90\% confidence level upper limits on the branching fractions in the range of $[0.8,\,3.6]\times10^{-5}$.

hep-ex↗

Measurement of the $^8$B Solar Neutrino Flux Using the Full SNO+ Water Phase Dataset

The SNO+ detector operated initially as a water Cherenkov detector. The implementation of a sealed covergas system midway through water data taking resulted in a significant reduction in the activity of $^{222}$Rn daughters in the detector and allowed the lowest background to the solar electron scattering signal above 5 MeV achieved to date. This paper reports an updated SNO+ water phase $^8$B solar neutrino analysis with a total livetime of 282.4 days and an analysis threshold of 3.5 MeV. The $^8$B solar neutrino flux is found to be $\left(2.32^{+0.18}_{-0.17}\text{(stat.)}^{+0.07}_{-0.05}\text{(syst.)}\right)\times10^{6}$ cm$^{-2}$s$^{-1}$ assuming no neutrino oscillations, or $\left(5.36^{+0.41}_{-0.39}\text{(stat.)}^{+0.17}_{-0.16}\text{(syst.)} \right)\times10^{6}$ cm$^{-2}$s$^{-1}$ assuming standard neutrino oscillation parameters, in good agreement with both previous measurements and Standard Solar Model Calculations. The electron recoil spectrum is presented above 3.5 MeV.

hep-ex↗

Quantum Decoherence at ESSnuSB Experiment

In this proceedings we study the sensitivity of the ESSnuSB experiment to probe quantum decoherence. ESSnuSB is a future long-baseline neutrino oscillation experiment which aims to measure $δ_{\rm CP}$ by probing the second oscillation maximum. Using the open quantum system formalism for decoherence, we have shown that the sensitivity of ESSnuSB to constrain the decoherence parameters is better than MINOS but comparable to DUNE. We have also shown that the CP measurement capability of ESSnuSB is robust in presence of decoherence.

hep-ph↗

Observations of the singly Cabibbo-suppressed decays $Ξ_c^{+} \to pK_{S}^{0}$, $Ξ_c^+ \to Λπ^+$, and $Ξ_c^+ \to Σ^{0} π^+$ at Belle and Belle II

Using data samples of 983.0~$\rm fb^{-1}$ and 427.9~$\rm fb^{-1}$ accumulated with the Belle and Belle~II detectors operating at the KEKB and SuperKEKB asymmetric-energy $e^+e^-$ colliders, singly Cabibbo-suppressed decays $Ξ_c^{+} \to pK_{S}^{0}$, $Ξ_c^+ \to Λπ^+$, and $Ξ_c^+ \to Σ^{0} π^+$ are observed for the first time. The ratios of branching fractions of $Ξ_{c}^{+}\to p K_{S}^{0}$, $Ξ_{c}^{+}\to Λπ^{+}$, and $Ξ_{c}^{+}\to Σ^{0} π^{+}$ relative to that of $Ξ_c^+ \to Ξ^- π^{+} π^{+}$ are measured to be \begin{equation} \frac{{\cal B}(Ξ_c^+ \to pK_S^0)}{{\cal B}(Ξ_c^{+} \to Ξ^{-} π^+ π^+)} = (2.47 \pm 0.16 \pm 0.07)\% \notag, \end{equation} \begin{equation} \frac{{\cal B}(Ξ_c^+ \to Λπ^+)}{{\cal B}(Ξ_c^{+} \to Ξ^{-} π^+ π^+)} = (1.56 \pm 0.14 \pm 0.09)\% \notag, \end{equation} \begin{equation} \frac{{\cal B}(Ξ_c^+ \to Σ^0 π^+)}{{\cal B}(Ξ_c^{+} \to Ξ^{-} π^+ π^+)} = (4.13 \pm 0.26 \pm 0.22)\% \notag. \end{equation} Multiplying these values by the branching fraction of the normalization channel, ${\cal B}(Ξ_c^{+} \to Ξ^{-} π^+π^+) = (2.9 \pm 1.3)\%$, the absolute branching fractions are determined to be \begin{equation} {\cal B}(Ξ_c^{+} \to p K_{S}^{0}) = (7.16 \pm 0.46 \pm 0.20 \pm 3.21) \times 10^{-4} \notag, \end{equation} \begin{equation} {\cal B}(Ξ_c^{+} \to Λπ^+) = (4.52 \pm 0.41 \pm 0.26 \pm 2.03) \times 10^{-4} \notag, \end{equation} \begin{equation} {\cal B}(Ξ_c^{+} \to Σ^0 π^+) = (1.20 \pm 0.08 \pm 0.07 \pm 0.54) \times 10^{-3} \notag. \end{equation} The first and second uncertainties above are statistical and systematic, respectively, while the third ones arise from the uncertainty in ${\cal B}(Ξ_c^{+} \to Ξ^{-} π^{+} π^{+})$.

hep-ex↗

Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models

We introduce Edify Image, a family of diffusion models capable of generating photorealistic image content with pixel-perfect accuracy. Edify Image utilizes cascaded pixel-space diffusion models trained using a novel Laplacian diffusion process, in which image signals at different frequency bands are attenuated at varying rates. Edify Image supports a wide range of applications, including text-to-image synthesis, 4K upsampling, ControlNets, 360 HDR panorama generation, and finetuning for image customization.

cs.CV↗

Edify 3D: Scalable High-Quality 3D Asset Generation

We introduce Edify 3D, an advanced solution designed for high-quality 3D asset generation. Our method first synthesizes RGB and surface normal images of the described object at multiple viewpoints using a diffusion model. The multi-view observations are then used to reconstruct the shape, texture, and PBR materials of the object. Our method can generate high-quality 3D assets with detailed geometry, clean shape topologies, high-resolution textures, and materials within 2 minutes of runtime.

cs.CV↗

New Cold Subdwarf Discoveries from Backyard Worlds and a Metallicity Classification System for T Subdwarfs

We report the results of a spectroscopic survey of candidate T subdwarfs identified by the Backyard Worlds: Planet 9 program. Near-infrared spectra of 31 sources with red $J-W2$ colors and large $J$-band reduced proper motions show varying signatures of subsolar metallicity, including strong collision-induced H$_2$ absorption, obscured methane and water features, and weak K I absorption. These metallicity signatures are supported by spectral model fits and 3D velocities, indicating thick disk and halo population membership for several sources. We identify three new metal-poor T subdwarfs ([M/H] $\lesssim$ $-$0.5), CWISE J062316.19+071505.6, WISEA J152443.14$-$262001.8, and CWISE J211250.11-052925.2; and 19 new "mild" subdwarfs with modest metal deficiency ([M/H] $\lesssim$ $-$0.25). We also identify three metal-rich brown dwarfs with thick disk kinematics. We provide kinematic evidence that the extreme L subdwarf 2MASS J053253.46+824646.5 and the mild T subdwarf CWISE J113010.07+313944.7 may be part of the Thamnos population, while the T subdwarf CWISE J155349.96+693355.2 may be part of the Helmi stream. We define a metallicity classification system for T dwarfs that adds mild subdwarfs (d/sdT), subdwarfs (sdT), and extreme subdwarfs (esdT) to the existing dwarf sequence. We also define a metallicity spectral index that correlates with metallicities inferred from spectral model fits and iron abundances from stellar primaries of benchmark T dwarf companions. This expansion of the T dwarf classification system supports investigations of ancient, metal-poor brown dwarfs now being uncovered in deep imaging and spectroscopic surveys.

astro-ph.SR↗

International comparison of optical frequencies with transportable optical lattice clocks

Optical clocks have improved their frequency stability and estimated accuracy by more than two orders of magnitude over the best caesium microwave clocks that realise the SI second. Accordingly, an optical redefinition of the second has been widely discussed, prompting a need for the consistency of optical clocks to be verified worldwide. While satellite frequency links are sufficient to compare microwave clocks, a suitable method for comparing high-performance optical clocks over intercontinental distances is missing. Furthermore, remote comparisons over frequency links face fractional uncertainties of a few $10^{-18}$ due to imprecise knowledge of each clock's relativistic redshift, which stems from uncertainty in the geopotential determined at each distant location. Here, we report a landmark campaign towards the era of optical clocks, where, for the first time, state-of-the-art transportable optical clocks from Japan and Europe are brought together to demonstrate international comparisons that require neither a high-performance frequency link nor information on the geopotential difference between remote sites. Conversely, the reproducibility of the clocks after being transported between countries was sufficient to determine geopotential height offsets at the level of 4 cm. Our campaign paves the way for redefining the SI second and has a significant impact on various applications, including tests of general relativity, geodetic sensing for geosciences, precise navigation, and future timing networks.

physics.atom-ph↗

GPT-4o System Card

GPT-4o is an autoregressive omni model that accepts as input any combination of text, audio, image, and video, and generates any combination of text, audio, and image outputs. It's trained end-to-end across text, vision, and audio, meaning all inputs and outputs are processed by the same neural network. GPT-4o can respond to audio inputs in as little as 232 milliseconds, with an average of 320 milliseconds, which is similar to human response time in conversation. It matches GPT-4 Turbo performance on text in English and code, with significant improvement on text in non-English languages, while also being much faster and 50\% cheaper in the API. GPT-4o is especially better at vision and audio understanding compared to existing models. In line with our commitment to building AI safely and consistent with our voluntary commitments to the White House, we are sharing the GPT-4o System Card, which includes our Preparedness Framework evaluations. In this System Card, we provide a detailed look at GPT-4o's capabilities, limitations, and safety evaluations across multiple categories, focusing on speech-to-speech while also evaluating text and image capabilities, and measures we've implemented to ensure the model is safe and aligned. We also include third-party assessments on dangerous capabilities, as well as discussion of potential societal impacts of GPT-4o's text and vision capabilities.

cs.CL↗

PLaMo-100B: A Ground-Up Language Model Designed for Japanese Proficiency

We introduce PLaMo-100B, a large-scale language model designed for Japanese proficiency. The model was trained from scratch using 2 trillion tokens, with architecture such as QK Normalization and Z-Loss to ensure training stability during the training process. Post-training techniques, including Supervised Fine-Tuning and Direct Preference Optimization, were applied to refine the model's performance. Benchmark evaluations suggest that PLaMo-100B performs well, particularly in Japanese-specific tasks, achieving results that are competitive with frontier models like GPT-4. The base model is available at https://huggingface.co/pfnet/plamo-100b.

cs.CL↗

A new method of reconstructing images of gamma-ray telescopes applied to the LST-1 of CTAO

Imaging atmospheric Cherenkov telescopes (IACTs) are used to observe very high-energy photons from the ground. Gamma rays are indirectly detected through the Cherenkov light emitted by the air showers they induce. The new generation of experiments, in particular the Cherenkov Telescope Array Observatory (CTAO), sets ambitious goals for discoveries of new gamma-ray sources and precise measurements of the already discovered ones. To achieve these goals, both hardware and data analysis must employ cutting-edge techniques. This also applies to the LST-1, the first IACT built for the CTAO, which is currently taking data on the Canary island of La Palma. This paper introduces a new event reconstruction technique for IACT data, aiming to improve the image reconstruction quality and the discrimination between the signal and the background from misidentified hadrons and electrons. The technique models the development of the extensive air shower signal, recorded as a waveform per pixel, seen by CTAO telescopes' cameras. Model parameters are subsequently passed to random forest regressors and classifiers to extract information on the primary particle. The new reconstruction was applied to simulated data and to data from observations of the Crab Nebula performed by the LST-1. The event reconstruction method presented here shows promising performance improvements. The angular and energy resolution, and the sensitivity, are improved by 10 to 20% over most of the energy range. At low energy, improvements reach up to 22%, 47%, and 50%, respectively. A future extension of the method to stereoscopic analysis for telescope arrays will be the next important step.

astro-ph.HE↗

First joint oscillation analysis of Super-Kamiokande atmospheric and T2K accelerator neutrino data

The Super-Kamiokande and T2K collaborations present a joint measurement of neutrino oscillation parameters from their atmospheric and beam neutrino data. It uses a common interaction model for events overlapping in neutrino energy and correlated detector systematic uncertainties between the two datasets, which are found to be compatible. Using 3244.4 days of atmospheric data and a beam exposure of $19.7(16.3) \times 10^{20}$ protons on target in (anti)neutrino mode, the analysis finds a 1.9$σ$ exclusion of CP-conservation (defined as $J_{CP}=0$) and a preference for the normal mass ordering.

hep-ex↗

Exploring atmospheric neutrino oscillations at ESSnuSB

This study provides an analysis of atmospheric neutrino oscillations at the ESSnuSB far detector facility. The prospects of the two cylindrical Water Cherenkov detectors with a total fiducial mass of 540 kt are investigated over 10 years of data taking in the standard three-flavor oscillation scenario. We present the confidence intervals for the determination of mass ordering, $θ_{23}$ octant as well as for the precisions on $\sin^2θ_{23}$ and $|Δm_{31}^2|$. It is shown that mass ordering can be resolved by $3σ$ CL ($5σ$ CL) after 4 years (10 years) regardless of the true neutrino mass ordering. Correspondingly, the wrong $θ_{23}$ octant could be excluded by $3σ$ CL after 4 years (8 years) in the case where the true neutrino mass ordering is normal ordering (inverted ordering). The results presented in this work are complementary to the accelerator neutrino program in the ESSnuSB project.

hep-ex↗

Measurements of the branching fractions of $Ξ_{c}^{0}\toΞ^{0}π^{0}$, $Ξ_{c}^{0}\toΞ^{0}η$, and $Ξ_{c}^{0}\toΞ^{0}η^{\prime}$ and asymmetry parameter of $Ξ_{c}^{0}\toΞ^{0}π^{0}$

We present a study of $Ξ_{c}^{0}\toΞ^{0}π^{0}$, $Ξ_{c}^{0}\toΞ^{0}η$, and $Ξ_{c}^{0}\toΞ^{0}η^{\prime}$ decays using the Belle and Belle~II data samples, which have integrated luminosities of 980~$\mathrm{fb}^{-1}$ and 426~$\mathrm{fb}^{-1}$, respectively. We measure the following relative branching fractions $${\cal B}(Ξ_{c}^{0}\toΞ^{0}π^{0})/{\cal B}(Ξ_{c}^{0}\toΞ^{-}π^{+}) = 0.48 \pm 0.02 ({\rm stat}) \pm 0.03 ({\rm syst}) ,$$ $${\cal B}(Ξ_{c}^{0}\toΞ^{0}η)/{\cal B}(Ξ_{c}^{0}\toΞ^{-}π^{+}) = 0.11 \pm 0.01 ({\rm stat}) \pm 0.01 ({\rm syst}) ,$$ $${\cal B}(Ξ_{c}^{0}\toΞ^{0}η^{\prime})/{\cal B}(Ξ_{c}^{0}\toΞ^{-}π^{+}) = 0.08 \pm 0.02 ({\rm stat}) \pm 0.01 ({\rm syst}) $$ for the first time, where the uncertainties are statistical ($\rm stat$) and systematic ($\rm syst$). By multiplying by the branching fraction of the normalization mode, ${\mathcal B}(Ξ_{c}^{0}\toΞ^{-}π^{+})$, we obtain the following absolute branching fraction results $(6.9 \pm 0.3 ({\rm stat}) \pm 0.5 ({\rm syst}) \pm 1.3 ({\rm norm})) \times 10^{-3}$, $(1.6 \pm 0.2 ({\rm stat}) \pm 0.2 ({\rm syst}) \pm 0.3 ({\rm norm})) \times 10^{-3}$, and $(1.2 \pm 0.3 ({\rm stat}) \pm 0.1 ({\rm syst}) \pm 0.2 ({\rm norm})) \times 10^{-3}$, for $Ξ_{c}^{0}$ decays to $Ξ^{0}π^{0}$, $Ξ^{0}η$, and $Ξ^{0}η^{\prime}$ final states, respectively. The third errors are from the uncertainty on ${\mathcal B}(Ξ_{c}^{0}\toΞ^{-}π^{+})$. The asymmetry parameter for $Ξ_{c}^{0}\toΞ^{0}π^{0}$ is measured to be $α(Ξ_{c}^{0}\toΞ^{0}π^{0}) = -0.90\pm0.15({\rm stat})\pm0.23({\rm syst})$.

hep-ex↗

The Visual Monitoring Camera (VMC) on Mars Express: a new science instrument made from an old webcam orbiting Mars

The Visual Monitoring Camera (VMC) is a small imaging instrument onboard Mars Express with a field of view of ~40x30 degrees. The camera was initially intended to provide visual confirmation of the separation of the Beagle 2 lander and has similar technical specifications to a typical webcam of the 2000s. In 2007, a few years after the end of its original mission, VMC was turned on again to obtain full-disk images of Mars to be used for outreach purposes. As VMC obtained more images, the scientific potential of the camera became evident, and in 2018 the camera was given an upgraded status of a new scientific instrument, with science goals in the field of Martian atmosphere meteorology. The wide Field of View of the camera combined with the orbit of Mars Express enable the acquisition of full-disk images of the planet showing different local times, which for a long time has been rare among orbital missions around Mars. The small data volume of images also allows videos that show the atmospheric dynamics of dust and cloud systems to be obtained. This paper is intended to be the new reference paper for VMC as a scientific instrument, and thus provides an overview of the updated procedures to plan, command and execute science observations of the Martian atmosphere. These observations produce valuable science data that is calibrated and distributed to the community for scientific use.

astro-ph.IM↗