arXiv ScienceSearch

arXiv subjects

Jonas Spinner

Publications and source records attributed to Jonas Spinner.

14 recordsLinked to original sources

Neural Boltzmann Equations

The dynamics of particles in the early universe are described by Boltzmann equations, which involve high-dimensional phase-space integrals. Classical approaches use quadrature integration and evolve the system on a fixed momentum grid, which scales poorly to complicated systems and parameter scans, severely limiting the complexity of processes that can be studied. We introduce Neural Boltzmann Equations (NBEs), which combine three coupled concepts to overcome these limitations. First, particle properties are encoded in physics-inspired neural distribution functions, with parameters that can be predicted using neural networks, enabling efficient parameter scans. Second, phase-space integrals are evaluated with Monte Carlo, using importance sampling tools from collider physics. Third, we use the natural gradient method to evolve the system. After demonstrating the individual benefits of NBEs, we use the framework to perform a precision calculation of the effective number of relativistic neutrino degrees of freedom in the early universe.

hep-ph

Virtues and Vices of Equivariant Transformers

We study for the first time the benefit of Lorentz-equivariant transformers for large-size jet tagging and flavor tagging. To control their computing demands, we optimize all implementations for inference cost metrics. In our scaling studies, we find that Lorentz-equivariant networks outperform standard transformers, provided geometric features are relevant. This holds true in an idealized world as well as for limited resources. The conditional gain from Lorentz equivariance provides interesting input to the development of foundation models for LHC data.

hep-ph

Economical Jet Taggers -- Equivariant, Slim, and Quantized

Modern machine learning is transforming jet tagging at the LHC, but the leading transformer architectures are large, not particularly fast, and training-intensive. We present a slim version of the L-GATr tagger, reduce the number of parameters of jet-tagging transformers, and quantize them. We compare different quantization methods for standard and Lorentz-equivariant transformers and estimate their gains in resource efficiency. We find an order-of-magnitude reduction in energy cost for an moderate performance decrease, down to 1000-parameter taggers. This might be a step towards trigger-level jet tagging with small and quantized versions of the leading equivariant transformer architectures.

hep-ph

Generative Unfolding of Jets and Their Substructure

Unfolding, for example of distortions imparted by detectors, provides suitable and publishable representations of LHC data. Many methods for unbinned and high-dimensional unfolding using machine learning have been proposed, but no generative method scales to the several hundred dimensions necessary to fully characterize LHC collisions. This paper proposes a 3-stage generative unfolding framework that is capable of unfolding several hundred dimensions. It is effective to unfold the jet-level kinematics as well as the full substructure of light-flavor jets and of top jets, and is the first generative unfolding study to achieve high precision on high-dimensional jet substructure.

hep-ph

Forecasting Generative Amplification

Generative networks are perfect tools to enhance the speed and precision of LHC simulations. Especially when generating events beyond the size of the training dataset, it is important to understand their statistical precision. We present two complementary methods to estimate the amplification factor without large holdout datasets. Averaging amplification uses Bayesian networks or ensembling to estimate amplification from the precision of integrals over given phase-space volumes. Differential amplification uses hypothesis testing to quantify amplification without any resolution loss. Applied to state-of-the-art event generators, both methods indicate that amplification is already possible in specific regions of phase space.

hep-ph

Lorentz-Equivariance without Limitations

Lorentz Local Canonicalization (LLoCa) ensures exact Lorentz-equivariance for arbitrary neural networks with minimal computational overhead. For the LHC, it equivariantly predicts local reference frames for each particle and propagates any-order tensorial information between them. We apply it to graph networks and transformers. We showcase its cutting-edge performance on amplitude regression, end-to-end event generation, and jet tagging. For jet tagging, we introduce a large top tagging dataset to benchmark LLoCa versions of a range of established benchmark architectures and highlight the importance of symmetry breaking.

hep-ph

The SN 1987A Cooling Bound on Dark Matter Absorption in Electron Targets

We present new supernova (SN 1987A) cooling bounds on sub-MeV fermionic dark matter with effective couplings to electrons. These bounds probe the parameter space relevant for direct detection experiments in which dark matter can be absorbed by the target material, showing strong complementarity with indirect searches and constraints from dark matter overproduction. Crucially, our limits exclude the projected sensitivity regions of current and upcoming direct detection experiments. Since these conclusions are a priori not valid for light mediators, we extend our analysis to this case. We show that sub-GeV mediators can be produced resonantly both in supernova cores and in the early Universe, altering the SN 1987A analysis for effective couplings. Still, a combination of supernova cooling constraints and limits from dark matter overproduction excludes the entire parameter space relevant for direct detection in this case.

hep-ph

Lorentz Local Canonicalization: How to Make Any Network Lorentz-Equivariant

Lorentz-equivariant neural networks are becoming the leading architectures for high-energy physics. Current implementations rely on specialized layers, limiting architectural choices. We introduce Lorentz Local Canonicalization (LLoCa), a general framework that renders any backbone network exactly Lorentz-equivariant. Using equivariantly predicted local reference frames, we construct LLoCa-transformers and graph networks. We adapt a recent approach for geometric message passing to the non-compact Lorentz group, allowing propagation of space-time tensorial features. Data augmentation emerges from LLoCa as a special choice of reference frame. Our models achieve competitive and state-of-the-art accuracy on relevant particle physics tasks, while being $4\times$ faster and using $10\times$ fewer FLOPs.

stat.ML

Extrapolating Jet Radiation with Autoregressive Transformers

Generative networks are an exciting tool for fast LHC event fixed number of particles. Autoregressive transformers allow us to generate events containing variable numbers of particles, very much in line with the physics of QCD jet radiation, and offer the possibility to generalize to higher multiplicities. We show how transformers can learn a factorized likelihood for jet radiation and extrapolate in terms of the number of generated jets. For this extrapolation, bootstrapping training data and training with modifications of the likelihood loss can be used.

hep-ph

A Lorentz-Equivariant Transformer for All of the LHC

We show that the Lorentz-Equivariant Geometric Algebra Transformer (L-GATr) yields state-of-the-art performance for a wide range of machine learning tasks at the Large Hadron Collider. L-GATr represents data in a geometric algebra over space-time and is equivariant under Lorentz transformations. The underlying architecture is a versatile and scalable transformer, which is able to break symmetries if needed. We demonstrate the power of L-GATr for amplitude regression and jet classification, and then benchmark it as the first Lorentz-equivariant generative network. For all three LHC tasks, we find significant improvements over previous architectures.

hep-ph

Lorentz-Equivariant Geometric Algebra Transformers for High-Energy Physics

Extracting scientific understanding from particle-physics experiments requires solving diverse learning problems with high precision and good data efficiency. We propose the Lorentz Geometric Algebra Transformer (L-GATr), a new multi-purpose architecture for high-energy physics. L-GATr represents high-energy data in a geometric algebra over four-dimensional space-time and is equivariant under Lorentz transformations, the symmetry group of relativistic kinematics. At the same time, the architecture is a Transformer, which makes it versatile and scalable to large systems. L-GATr is first demonstrated on regression and classification tasks from particle physics. We then construct the first Lorentz-equivariant generative model: a continuous normalizing flow based on an L-GATr network, trained with Riemannian flow matching. Across our experiments, L-GATr is on par with or outperforms strong domain-specific baselines.

physics.data-an

Supernova Limits on Muonic Dark Forces

Proto-neutron stars formed during core-collapse supernovae are hot and dense environments that contain a sizable population of muons. If these interact with new long-lived particles with masses up to roughly 100 MeV, the latter can be produced and escape from the stellar plasma, causing an excessive energy loss constrained by observations of SN 1987A. In this article we calculate the emission of light dark fermions that are coupled to leptons via a new massive vector boson, and determine the resulting constraints on the general parameter space. We apply these limits to the gauged $L_\mu-L_\tau$ model with dark fermions, and show that the SN 1987A constraints exclude a significant portion of the parameter space targeted by future experiments. We also extend our analysis to generic effective four-fermion operators that couple dark fermions to muons, electrons, or neutrinos. We find that SN 1987A cooling probes a new-physics scale up to $\sim7$ TeV, which is an order of magnitude larger than current bounds from laboratory experiments.

hep-ph

Jet Diffusion versus JetGPT -- Modern Networks for the LHC

We introduce two diffusion models and an autoregressive transformer for LHC physics simulations. Bayesian versions allow us to control the networks and capture training uncertainties. After illustrating their different density estimation methods for simple toy models, we discuss their advantages for Z plus jets event generation. While diffusion networks excel through their precision, the transformer scales best with the phase space dimensionality. Given the different training and evaluation speed, we expect LHC physics to benefit from dedicated use cases for normalizing flows, diffusion models, and autoregressive transformers.

hep-ph

The Axion-Higgs Portal

The phenomenology of axions and axion-like particles strongly depends on their couplings to Standard Model particles. The focus of this paper is the phenomenology of the unique dimension six operator respecting the shift symmetry: the axion-Higgs portal. We compare constraints from Higgs physics, flavor violating and radiative meson decays, bounds from atomic spectroscopy searching for fifth forces and astrophysical observables. In contrast to the QCD axion, axions interacting through the axion-Higgs portal are stable and can provide a dark matter candidate for any axion mass. We derive the parameter space for which freeze-out and freeze-in production as well as the misalignment mechanism can reproduce the observed relic abundance and compare the results with the phenomenological constraints. For comparison we also derive Higgs, flavor and spectroscopy constraints and the parameter space for which the scalar Higgs portal without derivative interactions can explain dark matter.

hep-ph