arXiv ScienceSearch

arXiv subjects

Mike Williams

Publications and source records attributed to Mike Williams.

At least 19 recordsLinked to original sources

Real-time lepton identification at LHCb in Run 3 using Lipschitz neural networks

The LHCb physics program in Run 3 relies critically on the efficient real-time selection of events containing muons and electrons, which are key signatures in a wide range of heavy-flavor and exotic decay processes. The LHCb Run 3 detector now operates with a fully software-based trigger that processes the complete detector readout at the LHC bunch-crossing rate, with the first trigger stage executed on GPUs. In this environment, particle-identification algorithms must achieve high efficiency and background rejection while satisfying stringent constraints on throughput and memory footprint. We present algorithms for muon and electron identification in the LHCb Run 3 GPU trigger based on Lipschitz-constrained neural networks. Separate networks are developed for muons and electrons and are trained using simulated events. Their performance is evaluated relative to the previous baseline algorithms, demonstrating improved discrimination across a wide range of kinematic regions while remaining compatible with the requirements of real-time GPU execution.

hep-ex

Decorrelation of neural networks from particle lifetimes in the LHCb topological $b$ trigger

The LHCb topological beauty trigger is the primary set of algorithms for selecting collision events containing $b$-hadrons in the fully software-based LHCb trigger. The algorithms apply monotonic Lipschitz neural networks (NNs) to select vertices of charged particles consistent with the distinct topology of a $b$ decay, i.e., those with large lifetimes and transverse momentum. Many analyses of the events recorded require that the selection must be unbiased with respect to the $b$-hadron lifetime at large lifetimes. Accurate reconstruction is challenging in busier detector environments, in which several visible proton-proton collisions occur simultaneously per bunch crossing, such that misassociation of decay products can result in vertices with artificially large measured lifetimes. This paper presents two approaches to mitigate correlations between NN scores and candidate lifetimes at large lifetime, and evaluates the performance of the resulting models.

hep-ex

Seeing the Forest Through the Trees: Knowledge Retrieval for Streamlining Particle Physics Analysis

Generative Large Language Models (LLMs) are a promising approach to structuring knowledge contained within the corpora of research literature produced by large-scale and long-running scientific collaborations. Within experimental particle physics, such structured knowledge bases could expedite methodological and editorial review. Complementarily, within the broader scientific community, generative LLM systems grounded in published work could make for reliable companions allowing non-experts to analyze open-access data. Techniques such as Retrieval Augmented Generation (RAG) rely on semantically matching localized text chunks, but struggle to maintain coherent context when relevant information spans multiple segments, leading to a fragmented representation devoid of global cross-document information. Here, we utilize the hierarchical organization of experimental physics articles to build a tree representation of the corpus, and present the SciTreeRAG system that uses this structure to create contexts that are more focused and contextually rich than standard RAG. Additionally, we develop methods for using LLMs to transform the unstructured corpus into a structured knowledge graph representation. We then implement SciGraphRAG, a retrieval system that leverages this knowledge graph to access global cross-document relationships eluding standard RAG, thereby encapsulating domain-specific connections and expertise. We demonstrate proof-of-concept implementations using the corpus of the LHCb experiment at CERN.

hep-ex

The Future of Artificial Intelligence and the Mathematical and Physical Sciences (AI+MPS)

This community paper developed out of the NSF Workshop on the Future of Artificial Intelligence (AI) and the Mathematical and Physics Sciences (MPS), which was held in March 2025 with the goal of understanding how the MPS domains (Astronomy, Chemistry, Materials Research, Mathematical Sciences, and Physics) can best capitalize on, and contribute to, the future of AI. We present here a summary and snapshot of the MPS community's perspective, as of Spring/Summer 2025, in a rapidly developing field. The link between AI and MPS is becoming increasingly inextricable; now is a crucial moment to strengthen the link between AI and Science by pursuing a strategy that proactively and thoughtfully leverages the potential of AI for scientific discovery and optimizes opportunities to impact the development of AI by applying concepts from fundamental science. To achieve this, we propose activities and strategic priorities that: (1) enable AI+MPS research in both directions; (2) build up an interdisciplinary community of AI+MPS researchers; and (3) foster education and workforce development in AI for MPS researchers and students. We conclude with a summary of suggested priorities for funding agencies, educational institutions, and individual researchers to help position the MPS community to be a leader in, and take full advantage of, the transformative potential of AI+MPS.

cs.AI

The DNA of nuclear models: How AI predicts nuclear masses

Obtaining high-precision predictions of nuclear masses, or equivalently nuclear binding energies, $E_b$, remains an important goal in nuclear-physics research. Recently, many AI-based tools have shown promising results on this task, some achieving precision that surpasses the best physics models. However, the utility of these AI models remains in question given that predictions are only useful where measurements do not exist, which inherently requires extrapolation away from the training (and testing) samples. Since AI models are largely black boxes, the reliability of such an extrapolation is difficult to assess. We present an AI model that not only achieves cutting-edge precision for $E_b$, but does so in an interpretable manner. For example, we find that (and explain why) the most important dimensions of its internal representation form a double helix, where the analog of the hydrogen bonds in DNA here link the number of protons and neutrons found in the most stable nucleus of each isotopic chain. Furthermore, we show that the AI prediction of $E_b$ can be factorized and ordered hierarchically, with the most important terms corresponding to well-known symbolic models (such as the famous liquid drop). Remarkably, the improvement of the AI model over symbolic ones can almost entirely be attributed to an observation made by Jaffe in 1969 based on the structure of most known nuclear ground states. The end result is a fully interpretable data-driven model of nuclear masses based on physics deduced by AI.

nucl-th

A covariant description of the interactions of axion-like particles and hadrons

We present a covariant framework for analyzing the interactions and decay rates of axion-like particles (ALPs) that couple to both gluons and quarks. We identify combinations of couplings that are invariant under quark-field redefinitions, and use them to obtain physical expressions for the prominent decay rates of such ALPs, which are compared with previous calculations for scenarios where ALPs couple exclusively to quarks or to gluons. Our framework can be used to obtain ALP decay rates for arbitrary ALP couplings to gluons and quarks across a broad range of ALP masses.

hep-ph

From Neurons to Neutrons: A Case Study in Interpretability

Mechanistic Interpretability (MI) promises a path toward fully understanding how neural networks make their predictions. Prior work demonstrates that even when trained to perform simple arithmetic, models can implement a variety of algorithms (sometimes concurrently) depending on initialization and hyperparameters. Does this mean neuron-level interpretability techniques have limited applicability? We argue that high-dimensional neural networks can learn low-dimensional representations of their training data that are useful beyond simply making good predictions. Such representations can be understood through the mechanistic interpretability lens and provide insights that are surprisingly faithful to human-derived domain knowledge. This indicates that such approaches to interpretability can be useful for deriving a new understanding of a problem from models trained to solve it. As a case study, we extract nuclear physics concepts by studying models trained to reproduce nuclear data.

cs.LG

Applications of Lipschitz neural networks to the Run 3 LHCb trigger system

The operating conditions defining the current data taking campaign at the Large Hadron Collider, known as Run 3, present unparalleled challenges for the real-time data acquisition workflow of the LHCb experiment at CERN. To address the anticipated surge in luminosity and consequent event rate, the LHCb experiment is transitioning to a fully software-based trigger system. This evolution necessitated innovations in hardware configurations, software paradigms, and algorithmic design. A significant advancement is the integration of monotonic Lipschitz neural networks into the LHCb trigger system. These deep learning models offer certified robustness against detector instabilities, and the ability to encode domain-specific inductive biases. Such properties are crucial for the inclusive heavy-flavour triggers and, most notably, for the topological triggers designed to inclusively select $b$-hadron candidates by exploiting the unique kinematic and decay topologies of beauty decays. This paper describes the recent progress in integrating Lipschitz neural networks into the topological triggers, highlighting the resulting enhanced sensitivity to highly displaced multi-body candidates produced within the LHCb acceptance.

hep-ex

Probing axion-like particles at the Electron-Ion Collider

The Electron-Ion Collider~(EIC), a forthcoming powerful high-luminosity facility, represents an exciting opportunity to explore new physics. In this article, we study the potential of the EIC to probe the coupling between axion-like particles~(ALPs) and photons in coherent scattering. The ALPs can be produced via photon fusion and decay back to two photons inside the EIC detector. In a prompt-decay search, we find that the EIC can set the most stringent bound for $m_a \lesssim 20\,\GeV$ and probe the effective scales $\Lambda \lesssim 10^{5}\,$GeV. In a displaced-vertex search, which requires adopting an EM calorimeter technology that provides directionality, the EIC could probe ALPs with $m_a \lesssim 1\,\GeV$ at effective scales $\Lambda \lesssim 10^{7}\,\GeV$. Combining the two search strategies, the EIC can probe a significant portion of unexplored parameter space in the $0.2 < m_a <20\,\GeV$ mass range.

hep-ph

Development of the Topological Trigger for LHCb Run 3

The data-taking conditions expected in Run 3 of the LHCb experiment at CERN are unprecedented and challenging for the software and computing systems. Despite that, the LHCb collaboration pioneers the use of a software-only trigger system to cope with the increased event rate efficiently. The beauty physics programme of LHCb is heavily reliant on topological triggers. These are devoted to selecting beauty-hadron candidates inclusively, based on the characteristic decay topology and kinematic properties expected from beauty decays. The following proceeding describes the current progress of the Run 3 implementation of the topological triggers using Lipschitz monotonic neural networks. This architecture offers robustness under varying detector conditions and sensitivity to long-lived candidates, improving the possibility of discovering New Physics at LHCb.

hep-ex

NuCLR: Nuclear Co-Learned Representations

We introduce Nuclear Co-Learned Representations (NuCLR), a deep learning model that predicts various nuclear observables, including binding and decay energies, and nuclear charge radii. The model is trained using a multi-task approach with shared representations and obtains state-of-the-art performance, achieving levels of precision that are crucial for understanding fundamental phenomena in nuclear (astro)physics. We also report an intriguing finding that the learned representation of NuCLR exhibits the prominent emergence of crucial aspects of the nuclear shell model, namely the shell structure, including the well-known magic numbers, and the Pauli Exclusion Principle. This suggests that the model is capable of capturing the underlying physical principles and that our approach has the potential to offer valuable insights into nuclear theory.

nucl-th

Snowmass 2021 Dark Matter Complementarity Report

The fundamental nature of Dark Matter is a central theme of the Snowmass 2021 process, extending across all Frontiers. In the last decade, advances in detector technology, analysis techniques and theoretical modeling have enabled a new generation of experiments and searches while broadening the types of candidates we can pursue. Over the next decade, there is great potential for discoveries that would transform our understanding of dark matter. In the following, we outline a road map for discovery developed in collaboration among the Frontiers. A strong portfolio of experiments that delves deep, searches wide, and harnesses the complementarity between techniques is key to tackling this complicated problem, requiring expertise, results, and planning from all Frontiers of the Snowmass 2021 process.

hep-ex

Snowmass 2021 Cross Frontier Report: Dark Matter Complementarity (Extended Version)

The fundamental nature of Dark Matter is a central theme of the Snowmass 2021 process, extending across all frontiers. In the last decade, advances in detector technology, analysis techniques and theoretical modeling have enabled a new generation of experiments and searches while broadening the types of candidates we can pursue. Over the next decade, there is great potential for discoveries that would transform our understanding of dark matter. In the following, we outline a road map for discovery developed in collaboration among the frontiers. A strong portfolio of experiments that delves deep, searches wide, and harnesses the complementarity between techniques is key to tackling this complicated problem, requiring expertise, results, and planning from all Frontiers of the Snowmass 2021 process.

hep-ph

Finding NEEMo: Geometric Fitting using Neural Estimation of the Energy Mover's Distance

A novel neural architecture was recently developed that enforces an exact upper bound on the Lipschitz constant of the model by constraining the norm of its weights in a minimal way, resulting in higher expressiveness compared to other techniques. We present a new and interesting direction for this architecture: estimation of the Wasserstein metric (Earth Mover's Distance) in optimal transport by employing the Kantorovich-Rubinstein duality to enable its use in geometric fitting applications. Specifically, we focus on the field of high-energy particle physics, where it has been shown that a metric for the space of particle-collider events can be defined based on the Wasserstein metric, referred to as the Energy Mover's Distance (EMD). This metrization has the potential to revolutionize data-driven collider phenomenology. The work presented here represents a major step towards realizing this goal by providing a differentiable way of directly calculating the EMD. We show how the flexibility that our approach enables can be used to develop novel clustering algorithms.

stat.ML

Dark Sector Physics at High-Intensity Experiments

Is Dark Matter part of a Dark Sector? The possibility of a dark sector neutral under Standard Model (SM) forces furnishes an attractive explanation for the existence of Dark Matter (DM), and is a compelling new-physics direction to explore in its own right, with potential relevance to fundamental questions as varied as neutrino masses, the hierarchy problem, and the Universe's matter-antimatter asymmetry. Because dark sectors are generically weakly coupled to ordinary matter, and because they can naturally have MeV-to-GeV masses and respect the symmetries of the SM, they are only mildly constrained by high-energy collider data and precision atomic measurements. Yet upcoming and proposed intensity-frontier experiments will offer an unprecedented window into the physics of dark sectors, highlighted as a Priority Research Direction in the 2018 Dark Matter New Initiatives (DMNI) BRN report. Support for this program -- in the form of dark-sector analyses at multi-purpose experiments, realization of the intensity-frontier experiments receiving DMNI funds, an expansion of DMNI support to explore the full breadth of DM and visible final-state signatures (especially long-lived particles) called for in the BRN report, and support for a robust dark-sector theory effort -- will enable comprehensive exploration of low-mass thermal DM milestones, and greatly enhance the potential of intensity-frontier experiments to discover dark-sector particles decaying back to SM particles.

hep-ph

Axial vectors in DarkCast

In this work, we explore new spin-1 states with axial couplings to the standard model fermions. We develop a data-driven method to estimate their hadronic decay rates based on data from $\tau$ decays and using SU(3)$_{\rm flavor}$ symmetry. We derive the current and future experimental constraints for several benchmark models. Our framework is generic and can be used for models with arbitrary vectorial and axial couplings to quarks. We have made our calculations publicly available by incorporating them into the DarkCast package, see https://gitlab.com/darkcast/releases.

hep-ph

Experiments and Facilities for Accelerator-Based Dark Sector Searches

This paper provides an overview of experiments and facilities for accelerator-based dark matter searches as part of the US Community Study on the Future of Particle Physics (Snowmass 2021). Companion white papers to this paper present the physics drivers: thermal dark matter, visible dark portals, and new flavors and rich dark sectors.

hep-ex

Towards Understanding Grokking: An Effective Theory of Representation Learning

We aim to understand grokking, a phenomenon where models generalize long after overfitting their training set. We present both a microscopic analysis anchored by an effective theory and a macroscopic analysis of phase diagrams describing learning performance across hyperparameters. We find that generalization originates from structured representations whose training dynamics and dependence on training set size can be predicted by our effective theory in a toy setting. We observe empirically the presence of four learning phases: comprehension, grokking, memorization, and confusion. We find representation learning to occur only in a "Goldilocks zone" (including comprehension and grokking) between memorization and confusion. We find on transformers the grokking phase stays closer to the memorization phase (compared to the comprehension phase), leading to delayed generalization. The Goldilocks phase is reminiscent of "intelligence from starvation" in Darwinian evolution, where resource limitations drive discovery of more efficient solutions. This study not only provides intuitive explanations of the origin of grokking, but also highlights the usefulness of physics-inspired tools, e.g., effective theories and phase diagrams, for understanding deep learning.

cs.LG