arXiv ScienceSearch

arXiv subjects

Aaron Philip

Publications and source records attributed to Aaron Philip.

4 recordsLinked to original sources

Bayesian Inference for Extracting Barrier Distributions from Fusion Excitation Functions

Barrier distributions encode rich information about the structure and dynamics of fusing nuclei, but extracting them from experimental fusion cross sections requires an estimate of the fusion excitation function's second derivative. In this work we approach the task of extracting barrier distributions with uncertainty estimates from sparse experimental measurements as a Bayesian inference problem. We introduce a method based on AutoBNN, an interpretable Bayesian machine learning framework, to provide a robust statistical approach for analyzing fusion excitation functions. Benchmarking against Gaussian process regression on simulated excitation functions that span a wide range of realistic experimental conditions, we find that AutoBNN more faithfully recovers the underlying barrier distribution and reports well-calibrated uncertainties. We then apply the AutoBNN method to four experimentally measured heavy-ion fusion reactions where it mitigates spurious above-barrier structure and constrains existing predictions. Alongside these results, we have developed a user-friendly software implementation of our method, facilitating its application to future heavy and light-ion fusion experiments.

nucl-th

Extracting Barrier Distributions from Fusion Cross Sections

Studying fusion cross sections provides insight into the fusion process, details about the internal structure of heavier nuclear systems, and a window into astrophysical processes. Barrier distributions, extracted from fusion excitation functions, are immensely useful for comparing theoretical model predictions and experimental results. Extracting this barrier distribution from the measured cross-section data amounts to taking the second derivative of the energy-weighted cross section. In practice, barrier distributions are highly sensitive to the quality of collected experimental data and the choice of step size when using standard point difference schemes. In this work, we explore Bayesian methods for extracting a posterior distribution over barrier distributions that could reasonably describe experimental data. We benchmark Gaussian processes and recently developed Bayesian machine learning inference algorithms against realistic simulated data generated from a simple model of fusion excitation functions. We find that Gaussian processes often exhibit aliasing at higher energies of the barrier distribution. We demonstrate that the BNN architectures can more faithfully recover the barrier distribution with quantified uncertainties at all energies, while also identifying key regions of high uncertainty and model discrepancy to determine precisely where additional experiments would be maximally impactful. We use our conclusions to calibrate models to measured experimental data. All methods are comparatively robust to data sparsity and irregularity, but we find that the single most important factor dictating the fidelity of all models is the relative size of experimental uncertainties. We release an open-source version of our analysis and a user-friendly implementation of our method to encourage its future usage for experimental analysis.

nucl-th

Dissecting Query-Key Interaction in Vision Transformers

Self-attention in vision transformers is often thought to perform perceptual grouping where tokens attend to other tokens with similar embeddings, which could correspond to semantically similar features of an object. However, attending to dissimilar tokens can be beneficial by providing contextual information. We propose to analyze the query-key interaction by the singular value decomposition of the interaction matrix (i.e. ${\textbf{W}_q}^\top\textbf{W}_k$). We find that in many ViTs, especially those with classification training objectives, early layers attend more to similar tokens, while late layers show increased attention to dissimilar tokens, providing evidence corresponding to perceptual grouping and contextualization, respectively. Many of these interactions between features represented by singular vectors are interpretable and semantic, such as attention between relevant objects, between parts of an object, or between the foreground and background. This offers a novel perspective on interpreting the attention mechanism, which contributes to understanding how transformer models utilize context and salient features when processing images.

cs.CV

Using Machine Learning Hamiltonians To Compute Molecular Motor Barrier Heights

Machine Learning Inter-atomic Potentials (MLIPs) have become a common tool in use by computational chemists due to their combination of accuracy and speed. Yet, it is still not clear how well these tools behave at or near transitions states found in complex molecules. Here we investigate the applicability of MLIPs in evaluating the transition barrier of two, complex, molecular motor systems: a 1st generation Feringa motor and the 9c alkene 2nd generation Feringa motor. We compared paths generated with the Hierarchically Interacting Particle Neural Network (HIP-NN), the PM3 semi-empirical quantum method (SEQM), PM3 interfaced with HIP-NN (SEQM+HIP-NN), and Density Functional Theory calculations. We found that using SEQM+HIP-NN to generate cheap, realistic pathway guesses then refining the intermediates with DFT allowed us to cheaply find realistic reaction paths and energy barriers matching experiment, providing evidence that deep learning can be used for high precision computational tasks such as transition path sampling while also suggesting potential application to high throughput screening.

physics.chem-ph