arXiv ScienceSearch

arXiv subjects

Ryan Coffee

Publications and source records attributed to Ryan Coffee.

At least 19 recordsLinked to original sources

Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap

One year ago, the AISLE roadmap argued that autonomous laboratories operated as isolated islands and proposed a grassroots network organized around five critical dimensions. The field has since moved faster than anticipated. Multi-agent systems have produced experimentally validated hypotheses, self-driving laboratories have grown more interoperable and orchestrated, reasoning-trained and domain foundation models have raised the capability ceiling, and the Genesis Mission has placed autonomous experimentation at the center of U.S. federal science strategy, with industry emerging as a primary actor. Progress has met a sobering counter-current, including a corrected flagship discovery result, benchmarks showing that agents which rival experts on closed-ended questions still complete only a fraction of open-ended research, and fabricated citations surfacing at leading venues. We read this as the defining tension of the field. Producing a candidate discovery is no longer the hard part, but verifying it is, and this asymmetry now limits autonomous science more than raw model capability. We update the roadmap around seven dimensions, revisiting the original five and elevating two former cross-cutting concerns, trust, verification, and reproducibility, and safety, security, and governance, to first-class status. We assess the original milestones (M1 through M14) as achieved, partially achieved, reframed, or open, add four new milestones (M15 through M18), and scope the path forward to a two-year horizon. The first year concentrates on interfaces, protocol adoption, and the scaffolding of verification, and the second targets federation, zero-trust coordination, and governance. Throughout, we position the grassroots network as the interoperability fabric that lets national programs, international initiatives, and commercial platforms connect rather than re-silo.

cs.DC

FPGA-Accelerated Real-Time Diagnostics at DIII-D Using the SLAC Neural Network Library for ML Inference

In this work, we demonstrate the deployment of a hardware-accelerated machine learning (ML) inference system integrated into a real-time processing at the DIII-D tokamak fusion reactor. The team has successfully deployed an AMD/Xilinx KCU1500 field-programmable gate array (FPGA) into the realtime Plasma Control System (PCS) nodes that receives the live Beam Emission Spectroscopy (BES) signal used for Edge Localized Mode (ELM) forecasting. The FPGA hosts a dense neural network using the SLAC Neural Network Library (SNL) that has been trained to infer the likelihood of disruptive ELM conditions. This likelihood then feeds a separate plasma controller that uses Resonant Magnetic Perturbation coils to suppress the predicted disruptive condition. The SNL allows for on-the-fly updates of the neural network weights and biases without requiring full hardware resynthesis for the FPGA. Judicious design of the neural-network architecture can further allow for the hot-swapping of multiple classification tasks to be executed on the single FPGA, significantly enhancing the real-time adaptability of the system for context-aware control strategies that respond in real-time to evolving reactor conditions. These adaptive weights naturally support continuous model refinement and seamless task switching during live experimental operation. This use case is chosen as a high rate signal processing example that can serve as a template for general ML-based reactor diagnostic processing for active reactor control systems. We see this as an essential development for achieving reactor relevant operation in future continuous operation fusion devices.

physics.plasm-ph

Advanced Control of Electron Beams: Tailoring X-ray Production with Programmable Laser Shaping

Leveraging the full scientific capabilities of next-generation high-repetition-rate free-electron lasers requires programmable control over electron-beam properties at their source. The photoinjector drive laser defines the electron beam's initial six-dimensional phase-space distribution, yet has historically been limited to Gaussian or static flat-top profiles, with most manipulation occurring downstream. Here we demonstrate software-programmable ultraviolet pulse shaping at the LCLS-II photoinjector as a source-level actuator that complements traditional accelerator controls. Using a coupled architecture combining dispersion-controlled nonlinear frequency conversion with spatial-light-modulator spectral shaping, we generate user-defined temporal structures and observe their imprint on electron bunches through high-resolution time-domain diagnostics. Laser-imposed multi-peaked modulation persists through acceleration, magnetic compression, and undulator transport with shot-to-shot repeatability, producing clearly resolved current structure in the compressed beam. Variance-based reconstruction from transverse deflecting cavity measurements reveals structured X-ray emission profiles exhibiting temporal features consistent with the programmed laser waveform. By providing rapid, software-controlled reconfiguration of electron-beam initial conditions, this source-level control approach establishes a programmable upstream actuator for future adaptive optimization and autonomous facility operation at high-repetition-rate light sources.

physics.acc-ph

Machine Learning on Heterogeneous, Edge, and Quantum Hardware for Particle Physics (ML-HEQUPP)

The next generation of particle physics experiments will face a new era of challenges in data acquisition, due to unprecedented data rates and volumes along with extreme environments and operational constraints. Harnessing this data for scientific discovery demands real-time inference and decision-making, intelligent data reduction, and efficient processing architectures beyond current capabilities. Crucial to the success of this experimental paradigm are several emerging technologies, such as artificial intelligence and machine learning (AI/ML), silicon microelectronics, and the advent of quantum algorithms and processing. Their intersection includes areas of research such as low-power and low-latency devices for edge computing, heterogeneous accelerator systems, reconfigurable hardware, novel codesign and synthesis strategies, readout for cryogenic or high-radiation environments, and analog computing. This white paper presents a community-driven vision to identify and prioritize research and development opportunities in hardware-based ML systems and corresponding physics applications, contributing towards a successful transition to the new data frontier of fundamental science.

physics.ins-det

Upstream Laser-based Longitudinal Enhancement of Relativistic Photoelectrons

Controlling the longitudinal phase space of high-brightness relativistic electron beams is crucial for advancing a broad spectrum of charged-particle-based instrumentation and scientific frontiers. A generalized method for achieving this control involves manipulating the photoemission laser's temporal distribution at the picosecond level, a long-standing technical challenge. Recent developments in laser shaping have enabled the creation of high-power, picosecond-scale symmetrical and asymmetrical temporal profiles, capable of fine-tuning complex space-charge dynamics and external field effects in relativistic charged-particle beams. Here, we demonstrate that rather than deviations from theorized, idealized laser distributions, a controlled asymmetry can be harnessed to counteract accelerator-induced distortions. By implementing spatiotemporal shaping of the ultraviolet photocathode laser at the LCLS-II superconducting injector, we achieve deterministic control over the longitudinal phase space without downstream corrections. We find that this optical asymmetry induces a self-linearizing effect across both low (40 pC) and high (80 pC) charge regimes, effectively suppressing nonlinear compression and energy chirp. Consequently, this approach is expected to preserve a low emittance comparable to that of ideal flattop or regular Gaussian profiles, while delivering superior current uniformity and shot-to-shot stability. These results establish spatiotemporal laser shaping as a compact, generalizable tool for directly optimizing beam brightness at the source.

physics.optics

FPGA-Accelerated Real-Time Beam Emission Spectroscopy Diagnostics at DIII-D Using the SLAC Neural Network Library for ML Inference

Achieving reliable real-time control of tokamak plasmas is essential for sustaining high-performance operation in next-generation fusion reactors. A major challenge is the accurate and timely prediction of edge-localized modes (ELMs), especially in high-confinement regimes such as wide-pedestal quiescent H-mode. We present a hardware-accelerated machine learning (ML) inference system integrated into the RTSTAB processing node of the DIII-D real-time diagnostic and control infrastructure. The system uses an AMD/Xilinx KCU1500 FPGA to enable ultra low latency plasma state classification and ELM forecasting. Input features come from real-time Beam Emission Spectroscopy (BES), and the ML model is implemented as a dense neural network using the SLAC Neural Network Library (SNL). A key capability is SNL dynamic parameter loading, which allows on-the-fly updates of neural network weights and biases without hardware resynthesis. This enables multiple classification tasks on a single FPGA design and supports adaptive control strategies that respond to evolving plasma conditions. By decoupling inference from fixed-weight configurations, the system supports continuous model refinement and seamless task switching during live operation. The SNL-based inference engine is fully integrated with the FPGA in the DIII-D RTSTAB Plasma Control System (PCS), improving ELM avoidance, confinement, and operational stability. These results show the feasibility of embedding dynamically reconfigurable FPGA-based ML inference into real-time fusion diagnostic pipelines, providing a scalable and resilient path toward intelligent and autonomous plasma control in future magnetic confinement fusion devices.

physics.plasm-ph

The LCLStream Ecosystem for Multi-Institutional Dataset Exploration

We describe a new end-to-end experimental data streaming framework designed from the ground up to support new types of applications -- AI training, extremely high-rate X-ray time-of-flight analysis, crystal structure determination with distributed processing, and custom data science applications and visualizers yet to be created. Throughout, we use design choices merging cloud microservices with traditional HPC batch execution models for security and flexibility. This project makes a unique contribution to the DOE Integrated Research Infrastructure (IRI) landscape. By creating a flexible, API-driven data request service, we address a significant need for high-speed data streaming sources for the X-ray science data analysis community. With the combination of data request API, mutual authentication web security framework, job queue system, high-rate data buffer, and complementary nature to facility infrastructure, the LCLStreamer framework has prototyped and implemented several new paradigms critical for future generation experiments.

cs.IR

A Grassroots Network and Community Roadmap for Interconnected Autonomous Science Laboratories for Accelerated Discovery

Scientific discovery is being revolutionized by AI and autonomous systems, yet current autonomous laboratories remain isolated islands unable to collaborate across institutions. We present the Autonomous Interconnected Science Lab Ecosystem (AISLE), a grassroots network transforming fragmented capabilities into a unified system that shorten the path from ideation to innovation to impact and accelerates discovery from decades to months. AISLE addresses five critical dimensions: (1) cross-institutional equipment orchestration, (2) intelligent data management with FAIR compliance, (3) AI-agent driven orchestration grounded in scientific principles, (4) interoperable agent communication interfaces, and (5) AI/ML-integrated scientific education. By connecting autonomous agents across institutional boundaries, autonomous science can unlock research spaces inaccessible to traditional approaches while democratizing cutting-edge technologies. This paradigm shift toward collaborative autonomous science promises breakthroughs in sustainable energy, materials development, and public health.

cs.CY

A Hybrid Neural Architecture: Online Attosecond X-ray Characterization

The emergence of high-repetition-rate x-ray free-electron lasers, such as SLAC's LCLS-II, serve as our canonical example for autonomous controls that necessitate high-throughput diagnostics paired with streaming computational pipelines capable of single-shot analysis with extremely low latency. We present the Deterministic Characterization with an Integrated Parallelizable Hybrid Resolver architecture, a hybrid machine learning framework designed for fast, accurate analysis of XFEL diagnostics using angular streaking-based sinogram images. This architecture integrates convolutional neural networks and bidirectional long short-term memory models to denoise input, identify x-ray sub-spike features, and extract sub-spike relative delays with sub-30 attosecond temporal resolution. Deployed on low-latency hardware, it achieves over 10~kHz throughput with \SI{168.3}{\micro\second} inference latency, indicating scalability to 14~kHz with FPGA integration. By transforming regression tasks into classification problems and leveraging optimized error encoding, we achieve high precision with low-latency performance that is critical for real-time streaming event selection and experimental control feedback signals. This represents a key development in real-time control pipelines for next-generation autonomous science, generally, and high repetition-rate x-ray experiments in particular.

physics.app-ph

Implementation of a framework for deploying AI inference engines in FPGAs

The LCLS2 Free Electron Laser FEL will generate xray pulses to beamline experiments at up to 1Mhz These experimentals will require new ultrahigh rate UHR detectors that can operate at rates above 100 kHz and generate data throughputs upwards of 1 TBs a data velocity which requires prohibitively large investments in storage infrastructure Machine Learning has demonstrated the potential to digest large datasets to extract relevant insights however current implementations show latencies that are too high for realtime data reduction objectives SLAC has endeavored on the creation of a software framework which translates MLs structures for deployment on Field Programmable Gate Arrays FPGAs deployed at the Edge of the data chain close to the instrumentation This framework leverages Xilinxs HLS framework presenting an API modeled after the open source Keras interface to the TensorFlow library This SLAC Neural Network Library SNL framework is designed with a streaming data approach optimizing the data flow between layers while minimizing the buffer data buffering requirements The goal is to ensure the highest possible framerate while keeping the maximum latency constrained to the needs of the experiment Our framework is designed to ensure the RTL implementation of the network layers supporting full redeployment of weights and biases without requiring resynthesis after training The ability to reduce the precision of the implemented networks through quantization is necessary to optimize the use of both DSP and memory resources in the FPGA We currently have a preliminary version of the toolset and are experimenting with both general purpose example networks and networks being designed for specific LCLS2 experiments.

physics.ins-det

Automatic Identification of Edge Localized Modes in the DIII-D Tokamak

Fusion power production in tokamaks uses discharge configurations that risk producing strong Type I Edge Localized Modes. The largest of these modes will likely increase impurities in the plasma and potentially damage plasma facing components such as the protective heat and waste divertor. Machine learning-based prediction and control may provide for online mitigation of these damaging modes before they grow too large to suppress. To that end, large labeled datasets are required for supervised training of machine learning models. We present an algorithm that achieves 97.7% precision when automatically labeling Edge Localized Modes in the large DIII-D tokamak discharge database. The algorithm has no user controlled parameters and is largely robust to tokamak and plasma configuration changes. This automatically-labeled database of events can subsequently feed future training of machine learning models aimed at autonomous Edge Localized Mode control and suppression.

physics.plasm-ph

Automating Dislocation Characterization in 3D Dark Field X-ray Microscopy

Mechanical properties in crystals are strongly correlated to the arrangement of 1D line defects, termed dislocations. Recently, Dark field X-ray Microscopy (DFXM) has emerged as a new tool to image and interpret dislocations within crystals using multidimensional scans. However, the methods required to reconstruct meaningful dislocation information from high-dimensional DFXM scans are still nascent and require significant manual oversight (i.e. \textit{supervision}). In this work, we present a new relatively unsupervised method that extracts dislocation-specific information (features) from a 3D dataset ($x$, $y$, $\phi$) using Gram-Schmidt orthogonalization to represent the large dataset as an array of 3-component feature vectors for each position, corresponding to the weak-beam conditions and the strong-beam condition. This method offers key opportunities to significantly reduce dataset size while preserving only the crystallographic information that is important for data reconstruction.

cond-mat.mtrl-sci

An Online Dynamic Amplitude-Correcting Gradient Estimation Technique to Align X-ray Focusing Optics

High-brightness X-ray pulses, as generated at synchrotrons and X-ray free electron lasers (XFEL), are used in a variety of scientific experiments. Many experimental testbeds require optical equipment, e.g Compound Refractive Lenses (CRLs), to be precisely aligned and focused. The lateral alignment of CRLs to a beamline requires precise positioning along four axes: two translational, and the two rotational. At a synchrotron, alignment is often accomplished manually. However, XFEL beamlines present a beam brightness that fluctuates in time, making manual alignment a time-consuming endeavor. Automation using classic stochastic methods often fail, given the errant gradient estimates. We present an online correction based on the combination of a generalized finite difference stencil and a time-dependent sampling pattern. Error expectation is analyzed, and efficacy is demonstrated. We provide a proof of concept by laterally aligning optics on a simulated XFEL beamline.

physics.ins-det

Applying Bayesian Inference and deterministic anisotropy to retrieve the molecular structure $|\Psi(\boldsymbol{R})|^2$ distribution from gas-phase diffraction experiments

Currently, our general approach to retrieving molecular structures from ultrafast gas-phase diffraction heavily relies on complex ab initio electronic or vibrational excited state simulations to make conclusive interpretations. Without such simulations, inverting this measurement for the structural probability distribution is typically intractable. This creates a so-called inverse problem. In this work, we develop a broadly applicable method that addresses this inverse problem by approximating the molecular frame structure $|\Psi(\boldsymbol{R}, t)|^2$ distribution independent of these complex simulations. We retrieve the vibronic ground state $|\Psi(\boldsymbol{R})|^2$ for both simulated stretched NO$_2$ and measured N$_2$O. From measured N$_2$O, we observe 40 mAngstroms coordinate-space resolution from 3.75 inverse Angstroms reciprocal space range and poor signal-to-noise, a 50X improvement over traditional Fourier transform methods. In simulated NO$_2$, typical to high signal-to-noise levels predict 100--1000X resolution improvements, down to 0.1 mAngstroms. By directly measuring the width of $|\Psi(\boldsymbol{R})|^2$, we open ultrafast gas-phase diffraction capabilities to measurements beyond current analysis approaches. This method has the potential to effectively turn gas-phase ultrafast diffraction into a discovery-oriented technique to probe systems that are prohibitively difficult to simulate.

physics.atom-ph

fairDMS: Rapid Model Training by Data and Model Reuse

Extracting actionable information rapidly from data produced by instruments such as the Linac Coherent Light Source (LCLS-II) and Advanced Photon Source Upgrade (APS-U) is becoming ever more challenging due to high (up to TB/s) data rates. Conventional physics-based information retrieval methods are hard-pressed to detect interesting events fast enough to enable timely focusing on a rare event or correction of an error. Machine learning~(ML) methods that learn cheap surrogate classifiers present a promising alternative, but can fail catastrophically when changes in instrument or sample result in degradation in ML performance. To overcome such difficulties, we present a new data storage and ML model training architecture designed to organize large volumes of data and models so that when model degradation is detected, prior models and/or data can be queried rapidly and a more suitable model retrieved and fine-tuned for new conditions. We show that our approach can achieve up to 100x data labelling speedup compared to the current state-of-the-art, 200x improvement in training speed, and 92x speedup in-terms of end-to-end model updating time.

cs.LG

Bridging Data Center AI Systems with Edge Computing for Actionable Information Retrieval

Extremely high data rates at modern synchrotron and X-ray free-electron laser light source beamlines motivate the use of machine learning methods for data reduction, feature detection, and other purposes. Regardless of the application, the basic concept is the same: data collected in early stages of an experiment, data from past similar experiments, and/or data simulated for the upcoming experiment are used to train machine learning models that, in effect, learn specific characteristics of those data; these models are then used to process subsequent data more efficiently than would general-purpose models that lack knowledge of the specific dataset or data class. Thus, a key challenge is to be able to train models with sufficient rapidity that they can be deployed and used within useful timescales. We describe here how specialized data center AI (DCAI) systems can be used for this purpose through a geographically distributed workflow. Experiments show that although there are data movement cost and service overhead to use remote DCAI systems for DNN training, the turnaround time is still less than 1/30 of using a locally deploy-able GPU.

cs.LG

Attosecond Coherent Electron Motion in Auger-Meitner Decay

In quantum systems, coherent superpositions of electronic states evolve on ultrafast timescales (few femtosecond to attosecond, 1 as = 0.001 fs = 10^{-18} s), leading to a time dependent charge density. Here we exploit the first attosecond soft x-ray pulses produced by an x-ray free-electron laser to induce a coherent core-hole excitation in nitric oxide. Using an additional circularly polarized infrared laser pulse we create a clock to time-resolve the electron dynamics, and demonstrate control of the coherent electron motion by tuning the photon energy of the x-ray pulse. Core-excited states offer a fundamental test bed for studying coherent electron dynamics in highly excited and strongly correlated matter.

physics.chem-ph

Coherent X-rays with Tunable Time-Dependent Polarization

We describe a method for producing high power, coherent x-ray pulses from a free electron laser with femtosecond scale periodic temporal modulation of the polarization vector. This approach relies on the generation of a temporal intensity modulation after self seeding either by modulating the seed intensity or the beam current. After generating a coherent temporally modulated $s$-polarization pulse, the electron beam is delayed by half a modulation period and sent into a short orthogonally oriented undulator, serving as a $p$-polarization afterburner. We provide simulations of three configurations for realizing this polarization switching, namely, enhanced self seeding with an intensity modulation generated by 2 color self seeding, enhanced self seeding of a current modulated bunch, and regular self seeding of a current modulated bunch. Start to end simulations for the Linac Coherent Light Source-II are provided for the latter.

physics.acc-ph