arXiv Science⌕ Search

arXiv subjects

:

Publications and source records attributed to :.

At least 109 records · Page 6Linked to original sources

Seedream 4.0: Toward Next-generation Multimodal Image Generation

We introduce Seedream 4.0, an efficient and high-performance multimodal image generation system that unifies text-to-image (T2I) synthesis, image editing, and multi-image composition within a single framework. We develop a highly efficient diffusion transformer with a powerful VAE which also can reduce the number of image tokens considerably. This allows for efficient training of our model, and enables it to fast generate native high-resolution images (e.g., 1K-4K). Seedream 4.0 is pretrained on billions of text-image pairs spanning diverse taxonomies and knowledge-centric concepts. Comprehensive data collection across hundreds of vertical scenarios, coupled with optimized strategies, ensures stable and large-scale training, with strong generalization. By incorporating a carefully fine-tuned VLM model, we perform multi-modal post-training for training both T2I and image editing tasks jointly. For inference acceleration, we integrate adversarial distillation, distribution matching, and quantization, as well as speculative decoding. It achieves an inference time of up to 1.8 seconds for generating a 2K image (without a LLM/VLM as PE model). Comprehensive evaluations reveal that Seedream 4.0 can achieve state-of-the-art results on both T2I and multimodal image editing. In particular, it demonstrates exceptional multimodal capabilities in complex tasks, including precise image editing and in-context reasoning, and also allows for multi-image reference, and can generate multiple output images. This extends traditional T2I systems into an more interactive and multidimensional creative tool, pushing the boundary of generative AI for both creativity and professional applications. We further scale our model and data as Seedream 4.5. Seedream 4.0 and Seedream 4.5 are accessible on Volcano Engine https://www.volcengine.com/experience/ark?launch=seedream.

cs.CV↗

Measurements of the mass difference $m(B^0)-m(B^+)$ and the energy dependence of the cross-section ratio $σ(e^+e^-\to B^0\bar{B}^0) / σ(e^+e^-\to B^+B^-)$ at Belle and Belle II

Using data samples collected by the Belle and Belle II experiments at the $Υ(4S)$ resonance with integrated luminosities of 571 fb$^{-1}$ and 365 fb$^{-1}$, respectively, we measure the pseudoscalar $B$-meson mass difference to be $m(B^0)-m(B^+) = (0.495\pm0.024\pm0.005)$ MeV/c$^2$. The results are based on a simultaneous fit to the variable $\tilde{M}_{bc}$, which is related to the $B$ momentum, for $B^0$ and $B^+$ candidates; and to the energy dependence of ${\cal R}=σ(e^+e^-\to B^0\bar{B}^0) / σ(e^+e^-\to B^+B^-)$, which is measured using changes in the average center-of-mass energy over the data taking periods. The phase-space hypothesis ${\cal R}=(p_{B^0}/p_{B^+})^3$, upon which previous measurements rely, is strongly disfavored by our fit; the measured mass-difference value for the phase-space hypothesis also differs significantly from our measurement. We constrain ${\cal R}$ in a broader energy range than covered by the direct measurement and extract the energy dependence of ${\cal R}$ in the range from the $B\bar{B}$ threshold up to 10.59 GeV. We interpret the results using a phenomenological model and constrain the parameters of the $B\bar{B}$ potential in the isovector channel.

hep-ex↗

Combining multiplexed functional data to improve variant classification

With the surge in the number of variants of uncertain significance (VUS) reported in ClinVar in recent years, there is an imperative to resolve VUS at scale. Multiplexed assays of variant effect (MAVEs), which allow the functional consequence of 100s to 1000s of genetic variants to be measured in a single experiment, are emerging as a powerful source of evidence which can be used in clinical gene variant classification. Increasingly, multiple published MAVEs are available for the same gene, sometimes measuring different aspects of variant impact. When multiple functional roles of a gene need to be considered, combining data from multiple MAVEs may provide a more comprehensive measure of the consequence of a genetic variant, which could impact variant classifications. Here, we provide guidance for combining such multiplexed functional data, incorporating a stepwise process from data curation and collection to model generation and validation. We demonstrate the potential and pitfalls of this approach by showing the integration of multiplexed functional data from five MAVEs for the gene TP53, two MAVEs for the gene LDLR and two MAVEs for PTEN. We also present a web applet that allows users to test various methods for combining score sets from multiple assays, calculate integrated functional scores for all variants, and assess whether combining data enables the application of stronger evidence for pathogenicity or benignity. By following these steps with appropriate guardrails, researchers can maximize the value of MAVEs, strengthen the functional evidence for clinical variant classification, and potentially uncover novel mechanisms of pathogenicity for clinically relevant genes.

q-bio.GN↗

NVIDIA Nemotron Nano V2 VL

We introduce Nemotron Nano V2 VL, the latest model of the Nemotron vision-language series designed for strong real-world document understanding, long video comprehension, and reasoning tasks. Nemotron Nano V2 VL delivers significant improvements over our previous model, Llama-3.1-Nemotron-Nano-VL-8B, across all vision and text domains through major enhancements in model architecture, datasets, and training recipes. Nemotron Nano V2 VL builds on Nemotron Nano V2, a hybrid Mamba-Transformer LLM, and innovative token reduction techniques to achieve higher inference throughput in long document and video scenarios. We are releasing model checkpoints in BF16, FP8, and FP4 formats and sharing large parts of our datasets, recipes and training code.

cs.LG↗

Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning

We present Isaac Lab, the natural successor to Isaac Gym, which extends the paradigm of GPU-native robotics simulation into the era of large-scale multi-modal learning. Isaac Lab combines high-fidelity GPU parallel physics, photorealistic rendering, and a modular, composable architecture for designing environments and training robot policies. Beyond physics and rendering, the framework integrates actuator models, multi-frequency sensor simulation, data collection pipelines, and domain randomization tools, unifying best practices for reinforcement and imitation learning at scale within a single extensible platform. We highlight its application to a diverse set of challenges, including whole-body control, cross-embodiment mobility, contact-rich and dexterous manipulation, and the integration of human demonstrations for skill acquisition. Finally, we discuss upcoming integration with the differentiable, GPU-accelerated Newton physics engine, which promises new opportunities for scalable, data-efficient, and gradient-based approaches to robot learning. We believe Isaac Lab's combination of advanced simulation capabilities, rich sensing, and data-center scale execution will help unlock the next generation of breakthroughs in robotics research.

cs.RO↗

Measurement of the $ D^{0}\rightarrow K^{-}π^{+}e^{+}e^{-} $ branching fraction and search for $ D^{0}\rightarrow π^{+}π^{-}e^{+}e^{-} $ and $D^{0}\rightarrow K^{+}K^{-}e^{+}e^{-} $ decays at Belle

We present a study of the rare charm meson decays $ D^{0}\rightarrow K^{+}K^{-}e^{+}e^{-} $, $ π^{+}π^{-}e^{+}e^{-} $, and $ K^{-}π^{+}e^{+}e^{-} $ using a 942 fb$^{-1}$ data set collected by the Belle detector at the KEKB asymmetric-energy $ e^{+}e^{-} $ collider. We use $ D^{0} $ candidates identified by the charge of the pion in $ D^{*} \rightarrow D^{0} π$ decays and normalize the branching fractions to $ D^{0} \rightarrow K^{-}π^{+}π^{-}π^{+} $ decays. The branching fraction for decay $ D^{0} \rightarrow K^{-}π^{+}e^{+}e^{-} $ is measured to be (39.6 $\pm$ 4.5 (stat) $\pm$ 2.9 (syst)) $\times$ $10^{-7}$, with the dielectron mass in the $ ρ/ω$ mass region $ 675 < m_{ee} < 875 $ MeV$/c^{2}$. We also search for $ D^{0}\rightarrow h^{-} h^{(\prime)+}e^{+}e^{-} $ ($ h^{(\prime)}=K,\,π$) decays with the dielectron mass near the $η$ and $ϕ$ resonances, and away from these resonances for the $ K^{+}K^{-}e^{+}e^{-} $ and $ π^{+}π^{-}e^{+}e^{-} $ modes. For these modes, we find no significant signals and set 90$\%$ confidence level upper limits on their branching fractions at the $\mathcal{O}$(10$^{-7}$) level.

hep-ex↗

Search for an Axion-Like Particle in $B\rightarrow K^{(*)} a (\rightarrowγγ)$ Decays at Belle

We report a search for an axion-like particle $a$ in $B\rightarrow K^{(*)} a (\rightarrowγγ)$ decays using data collected with the Belle detector at the KEKB asymmetric energy electron-positron collider. The search is based on a $711 \mathrm{fb^{-1}}$ data sample collected at the $Υ4S$ resonance energy, corresponding to a sample of $772\times10^6$ $Υ4S$ events. In this study, we search for the decay of the axion-like particle into a pair of photons, $a \rightarrow γγ$. We scan the two-photon invariant mass in the range $0.16\ \mathrm{GeV/}c^2-4.50\ \mathrm{GeV}/c^2$ for the $K$ modes and $0.16\ \mathrm{GeV/}c^2-4.20\ \mathrm{GeV}/c^2$ for the $K^{*}$ modes. No significant signal is observed in any of the modes, and 90\% confidence level upper limits are established on the coupling to the $W$ boson, $g_aW$, as a function of $a$ mass. The limits range from $3 \times 10^{-6} \mathrm{GeV}^{-1}$ to $3 \times 10^{-5} \mathrm{GeV}^{-1}$, improving the current constraints on $g_aW$ by a factor of two over the most stringent previous experimental results.

hep-ex↗

First Evidence of Solar Neutrino Interactions on $^{13}$C

The SNO+ Collaboration reports the first evidence of $^{8}\text{B}$ solar neutrinos interacting on $^{13}\text{C}$ nuclei. The charged current interaction proceeds through $^{13}\text{C} + ν_e \rightarrow {}^{13}\text{N} + e^-$ which is followed, with a 10 minute half-life, by ${}^{13}\text{N} \rightarrow {}^{13}\text{C} + e^+ +ν_e .$ The detection strategy is based on the delayed coincidence between the electron and the positron. Evidence for the charged current signal is presented with a significance of 4.2$σ$. Using the natural abundance of $^{13}\text{C}$ present in the scintillator, 5.7 tonnes of $^{13}\text{C}$ over 231 days of data were used in this analysis. The 5.6$^{+3.0}_{-2.3}$ observed events in the data set are consistent with the expectation of 4.7$^{+0.6}_{-1.3}$ events. This result is the second real-time measurement of CC interactions of $^{8}\text{B}$ neutrinos with nuclei and constitutes the lowest energy observation of neutrino interactions on $^{13}\text{C}$ generally. This enables the first direct measurement of the CC $ν_e$ reaction to the ground state of ${}^{13}\text{N}$, yielding an average cross section of $(16.1 ^{+8.5}_{-6.7} (\text{stat.}) ^{+1.6}_{-2.7} (\text{syst.}) )\times 10^{-43}$ cm$^{2}$ over the relevant $^{8}\text{B}$ solar neutrino energies.

nucl-ex↗

Unbinned measurement of thrust in $e^+e^-$ collisions at $\sqrt{s}$ = 91.2 GeV with ALEPH archived data

The strong coupling constant ($α_{S}$) is a fundamental parameter of quantum chromodynamics (QCD), the theory of the strong force. Some of the earliest precise constraints on $α_{S}$ came from measurements of event shape observables, such as thrust ($T$), using hadronic $Z$ boson decays produced in $e^+e^-$ collisions. However, recent work has revealed discrepancies between event-shape-based extractions of $α_{S}$ and values determined using other experimental methods. This work reexamines archived $e^+e^-$ data collected at a collision energy of $\sqrt{s}=91.2$ GeV by the ALEPH detector at the Large Electron-Positron Collider. Modern machine learning techniques are used to correct for detector effects in an unbinned manner, allowing the $T$ distribution to be measured with higher granularity than previous ALEPH measurements. The new measurement reveals a small but systematic shift towards larger values of $τ=1-T$, and the potential implications of this shift for $α_{S}$ extractions are illustrated by comparing to state-of-the-art theoretical calculations. In addition, the region of $-6<\logτ<-2$, where poorly-understood non-perturbative effects are large, is compared to modern parton shower Monte Carlo simulations. This measurement provides unique new inputs for $α_{S}$ extractions and also improves constraints on phenomenological models of QCD dynamics such as parton fragmentation and hadronization.

hep-ex↗

Joint neutrino oscillation analysis from the T2K and NOvA experiments

The landmark discovery that neutrinos have mass and can change type (or "flavor") as they propagate -- a process called neutrino oscillation -- has opened up a rich array of theoretical and experimental questions being actively pursued today. Neutrino oscillation remains the most powerful experimental tool for addressing many of these questions, including whether neutrinos violate charge-parity (CP) symmetry, which has possible connections to the unexplained preponderance of matter over antimatter in the universe. Oscillation measurements also probe the mass-squared differences between the different neutrino mass states ($Δm^2$), whether there are two light states and a heavier one (normal ordering) or vice versa (inverted ordering), and the structure of neutrino mass and flavor mixing. Here, we carry out the first joint analysis of data sets from NOvA and T2K, the two currently operating long-baseline neutrino oscillation experiments (hundreds of kilometers of neutrino travel distance), taking advantage of our complementary experimental designs and setting new constraints on several neutrino sector parameters. This analysis provides new precision on the $Δm^2_{32}$ mass difference, finding $2.43^{+0.04}_{-0.03}\ \left(-2.48^{+0.03}_{-0.04}\right)\times 10^{-3}~\mathrm{eV}^2$ in the normal (inverted) ordering, as well as a $3σ$ interval on $δ_{\rm CP}$ of $[-1.38π,\ 0.30π]$ $\left([-0.92π,\ -0.04π]\right)$ in the normal (inverted) ordering. The data show no strong preference for either mass ordering, but notably if inverted ordering were assumed true within the three-flavor mixing paradigm, then our results would provide evidence of CP symmetry violation in the lepton sector.

hep-ex↗

Measurement of the time-integrated $CP$ asymmetry in $D^0 \to K^0_{\rm S} K^0_{\rm S}$ decays using opposite-side flavor tagging at Belle and Belle II

We measure the time-integrated $CP$ asymmetry in $D^0 \to K^0_{\rm S} K^0_{\rm S}$ decays reconstructed in $e^+e^-\to c{\overline c}$ events collected by the Belle and Belle II experiments. The corresponding data samples have integrated luminosities of 980 and 428 fb${}^{-1}$, respectively. To infer the flavor of the $D^0$ meson, we exploit the correlation between the flavor of the reconstructed decay and the electric charges of particles reconstructed in the rest of the $e^+e^-\to c{\overline c}$ event. This results in a sample which is independent from any other previously used at Belle or Belle II. The result, $A_{CP}(D^0 \to K^0_{\rm S} K^0_{\rm S}) = (1.3 \pm 2.0 \pm 0.2)\%$, where the first uncertainty is statistical and the second systematic, is consistent with previous determinations and with $CP$ symmetry.

hep-ex↗

Search for lepton flavor-violating decay modes $B^0 \to K^{\ast 0}τ^\pm\ell^\mp$ ($\ell = e,μ$) with hadronic B-tagging at Belle and Belle II

We present the results of a search for the charged-lepton-flavor violating decays $B^0 \rightarrow K^{*0}τ^\pm \ell^{\mp}$, where $\ell^{\mp}$ is either an electron or a muon. The results are based on 365 fb$^{-1}$ and 711 fb$^{-1}$ datasets collected with the Belle II and Belle detectors, respectively. We use an exclusive hadronic $B$-tagging technique, and search for a signal decay in the system recoiling against a fully reconstructed $B$ meson. We find no evidence for $B^0 \rightarrow K^{*0}τ^\pm \ell^{\mp}$ decays and set upper limits on the branching fractions in the range of $(2.9-6.4)\times10^{-5}$ at 90% confidence level.

hep-ex↗

VHE $γ$-ray observations of bright BL Lacs with the Large-Sized Telescope prototype (LST-1) of the CTAO

Cherenkov Telescope Array Observatory (CTAO) is the next-generation ground-based gamma-ray observatory operating in the energy range from 20 GeV up to 300 TeV, with two sites in La Palma (Spain) and Paranal (Chile). It will consist of telescopes of three sizes, covering different parts of the large energy range. We report on the performance of Large-Sized Telescope prototype (LST-1) in the detection and characterization of extragalactic gamma-ray sources, with a focus on the reconstructed gamma-ray spectra and variability of classical bright BL Lacertae objects, which were observed during the early commissioning phase of the instrument. LST-1 data from known bright gamma-ray blazars - Markarian 421, Markarian 501, 1ES 1959+650, 1ES 0647+250, and PG 1553+113 - were collected between July 10, 2020, and May 23, 2022, covering a zenith angle range of 4 deg to 57 deg. The reconstructed light curves were analyzed using a Bayesian block algorithm to distinguish the different activity phases of each blazar. Simultaneous Fermi-LAT data were utilized to reconstruct the broadband $γ$-ray spectra for the sources during each activity phase. High-level reconstructed data in a format compatible with gammapy are provided together with measured light curves and spectral energy distributions (SEDs) for several bright blazars and an interpretation of the observed variability in long and short timescales. Simulations of historical flares are generated to evaluate the sensitivity of LST-1. This work represents the first milestone in monitoring bright BL Lacertae objects with a CTAO telescope.

astro-ph.HE↗

Towards an AI-Augmented Textbook

Textbooks are a cornerstone of education, but they have a fundamental limitation: they are a one-size-fits-all medium. Any new material or alternative representation requires arduous human effort, so that textbooks cannot be adapted in a scalable manner. We present an approach for transforming and augmenting textbooks using generative AI, adding layers of multiple representations and personalization while maintaining content integrity and quality. We refer to the system built with this approach as Learn Your Way. We report pedagogical evaluations of the different transformations and augmentations, and present the results of a a randomized control trial, highlighting the advantages of learning with Learn Your Way over regular textbook usage.

cs.CY↗

PLaMo 2 Technical Report

In this report, we introduce PLaMo 2, a series of Japanese-focused large language models featuring a hybrid Samba-based architecture that transitions to full attention via continual pre-training to support 32K token contexts. Training leverages extensive synthetic corpora to overcome data scarcity, while computational efficiency is achieved through weight reuse and structured pruning. This efficient pruning methodology produces an 8B model that achieves performance comparable to our previous 100B model. Post-training further refines the models using a pipeline of supervised fine-tuning (SFT) and direct preference optimization (DPO), enhanced by synthetic Japanese instruction data and model merging techniques. Optimized for inference using vLLM and quantization with minimal accuracy loss, the PLaMo 2 models achieve state-of-the-art results on Japanese benchmarks, outperforming similarly-sized open models in instruction-following, language fluency, and Japanese-specific knowledge.

cs.CL↗

Hunyuan3D-Omni: A Unified Framework for Controllable Generation of 3D Assets

Recent advances in 3D-native generative models have accelerated asset creation for games, film, and design. However, most methods still rely primarily on image or text conditioning and lack fine-grained, cross-modal controls, which limits controllability and practical adoption. To address this gap, we present Hunyuan3D-Omni, a unified framework for fine-grained, controllable 3D asset generation built on Hunyuan3D 2.1. In addition to images, Hunyuan3D-Omni accepts point clouds, voxels, bounding boxes, and skeletal pose priors as conditioning signals, enabling precise control over geometry, topology, and pose. Instead of separate heads for each modality, our model unifies all signals in a single cross-modal architecture. We train with a progressive, difficulty-aware sampling strategy that selects one control modality per example and biases sampling toward harder signals (e.g., skeletal pose) while downweighting easier ones (e.g., point clouds), encouraging robust multi-modal fusion and graceful handling of missing inputs. Experiments show that these additional controls improve generation accuracy, enable geometry-aware transformations, and increase robustness for production workflows.

cs.CV↗

Measurement of the time-integrated CP asymmetry in $D^{0}\rightarrow K^{0}_{S}K^{0}_{S}$ decays using Belle and Belle II data

We measure the time-integrated CP asymmetry in $D^{0} \rightarrow K^{0}_{S}K^{0}_{S}$ decays reconstructed in $e^{+}e^{-} \rightarrow c\overline{c}$ events collected by the Belle and Belle II experiments. The corresponding data samples have integrated luminosities of 980 fb$^{-1}$ and 428 fb$^{-1}$, respectively. The $D^{0}$ decays are required to originate from the $D^{*+} \rightarrow D^{0}π^{+}$ decay, which determines the charm flavor at production time. A control sample of $D^{0} \rightarrow K^{+}K^{-}$ decays is used to correct for production and detection asymmetries. The result, $(-1.4\pm1.3{\rm(stat)}\pm0.1{\rm (syst)})\%$, is consistent with previous determinations and with CP symmetry.

hep-ex↗

Observation of the decays $B^{+} \to Σ_{c}(2455)^{++} \overlineΞ_{c}^{-}$ and $B^{0} \to Σ_{c}(2455)^{0} \overlineΞ_{c}^{0}$

We report the first observation of the two-body baryonic decays $B^{+} \to Σ_{c}(2455)^{++} \overlineΞ_{c}^{-}$ and $B^{0} \to Σ_{c}(2455)^{0} \overlineΞ_{c}^{0}$ with significances of $7.3\,σ$ and $6.2\,σ$, respectively, including statistical and systematic uncertainties. The branching fractions are measured to be $\mathcal{B}(B^{+} \to Σ_{c}(2455)^{++} \overlineΞ_{c}^{-}) = (5.74 \pm 1.11 \pm 0.42_{-1.53}^{+2.47}) \times 10^{-4}$ and $\mathcal{B}(B^{0} \to Σ_{c}(2455)^{0} \overlineΞ_{c}^{0}) = (4.83 \pm 1.12 \pm 0.37_{-0.60}^{+0.72}) \times 10^{-4}$. The first and second uncertainties are statistical and systematic, respectively, while the third ones arise from the absolute branching fractions of $\overlineΞ_{c}^{-}$ or $\overlineΞ_{c}^{0}$ decays. The data samples used for this analysis have integrated luminosities of 711~$\mathrm{fb}^{-1}$ and 365~$\mathrm{fb}^{-1}$, and were collected at the $Υ(4S)$ resonance by the Belle and Belle~II detectors operating at the KEKB and SuperKEKB asymmetric-energy $e^{+}e^{-}$ colliders, respectively.

hep-ex↗