arXiv ScienceSearch

arXiv subjects

Alejandro Strachan

Publications and source records attributed to Alejandro Strachan.

At least 19 recordsLinked to original sources

Microstructure-Aware Deep Learning Bridges Atomistics to Macroscale for Shock-to-Detonation Prediction

The shock-to-detonation transition in energetic materials is governed by coupled processes spanning Angstroms to millimeters and femtoseconds to microseconds, where traditional multiscale models fail due to the lack of scale separation. We address this grand challenge by directly bridging large-scale molecular dynamics (MD) simulations with continuum finite-element (FE) models using MISTnetX, a convolutional deep neural network. Trained on MD simulations of shock propagation through complex microstructures, MISTnetX captures shock-microstructure interactions, hotspot formation, and the transition to deflagration, supplying critical sub-grid information to FE simulations of mechanics, shocks, thermal transport, and chemistry. Applied to a synthetic but realistic nanostructured plastic-bonded RDX composite, MISTnetX enables parameter-free prediction of the full run-to-detonation transition.

cond-mat.mtrl-sci

A Data-Driven Parametric Reduced-Order Chemical Kinetics Model Derived from Atomistic Simulations

Coarse-grained modeling in molecular simulations serves not only to extend accessible time and length scales beyond atomistic limits, but also to reduce high-dimensional chemical data to low-dimensional representations that expose the underlying latent structure. In the context of energetic materials, reduced-order chemical kinetics models are essential for describing thermally driven decomposition, deflagration, and detonation. Recent data-driven approaches based on machine learning and dimensionality reduction have shown promise for constructing such models directly from atomistic simulations; however, when reaction pathways vary strongly with thermodynamic conditions, these methods can produce latent representations that are difficult to interpret physically or extrapolate reliably. Here, we introduce a parametric, temperature-dependent autoencoder framework that learns a unified reduced-order description of chemical decomposition across a wide range of temperatures within a single model. Physical interpretability is enforced through non-negativity constraints and a softmax activation, enabling the latent variables to be directly associated with additive chemical components and their relative contributions. Reaction kinetics and heat-release parameters are optimized simultaneously within the neural-network architecture, providing a self-consistent coupling between chemical evolution and energetics. The proposed approach yields significantly improved reconstruction accuracy compared to a state-of-the-art dimensionality-reduction method, as quantified by reductions in mean-squared error, while preserving a physically meaningful latent representation. These results demonstrate that parametric, interpretable machine-learning models can provide robust reduced-order chemical kinetics suitable for multiscale modeling of complex reactive systems.

physics.chem-ph

Evaluating LLM-generated code for domain-specific languages: molecular dynamics with LAMMPS

Large language models (LLMs) are changing the way researchers interact with code and data in scientific computing. While their ability to generate general-purpose code is well established, their effectiveness in producing scientifically valid scripts for domain-specific language (DSLs) remains largely unexplored. We propose an evaluation procedure that enables domain experts to assess the validity of LLM-generated input files for LAMMPS, a widely used molecular dynamics (MD) code, without requiring deep familiarity with its syntax. The evaluation procedure combines a normalization step that produces canonical input files with an extensible parser for syntax analysis, followed by a reduced-cost execution stage and accuracy checks that isolate common errors before running costly simulations. We apply the pipeline to eight state-of-the-art LLMs across three prompts of increasing complexity. The parser pass rate has improved from 74% to 91% over the past year, but scientific accuracy on coupled multi-step workflows remains limited. Across all 80 scripts evaluated on the most complex prompt, only one was fully correct as generated. We further package the automated stages as a reusable agentic skill that LLMs can invoke during script generation; in a small-scale demonstration, this skill helped two models produce five fully correct scripts out of six across the same three prompts, including the hardest one. The pipeline highlights both the limitations of current LLMs in generating scientific DSLs and a practical path toward integrating them into domain-specific computational ecosystems.

cs.SE

Nuclear Quantum Effects in Multi-Step Condensed Matter Chemistry: A Path Integral Molecular Dynamics Study of Thermal Decomposition

Nuclear quantum effects (NQEs) are often central to a predictive understanding of chemical reactions and rates. While their incorporation in gas-phase reactions is well established, studies involving condensed matter often neglect or approximate such effects. To clarify the role of NQEs in multi-step, multi-molecular reactions in a molecular crystal, we compare atomistic simulations of the thermal decomposition of the energetic material TATB using path integral molecular dynamics (PIMD), the more approximate quantum thermal bath (QTB), and classical MD (ClMD). PIMD samples the quantum canonical distribution by representing each atom as a string of beads (replicas), while QTB uses a frequency-dependent thermostat to reproduce the Bose-Einstein distribution. We find that PIMD results in faster chemical decomposition of the TATB crystal compared to ClMD, as the initial steps involve hydrogen transfer processes. Interestingly, some of the subsequent reactions (e.g. the formation of N2) occur on identical timescales. The PIMD simulations also predict a reduction in overall activation energy by ~8% as compared to the classical result. As observed in model systems and simple unimolecular gas-phase reactions, the QTB significantly overestimates quantum acceleration of chemical reactions and the reduction in activation energy. A comparison of the kinetic energy operator in PIMD and the centroid dynamics provides insight into the physics behind the differences between the QTB and PIMD results.

physics.chem-ph

Physics-constrained Gaussian Processes for Predicting Shockwave Hugoniot Curves

A physics-constrained Gaussian Process regression framework is developed for predicting shocked material states and their associated uncertainties along the Hugoniot curve using data from a small number of shockwave simulations. The proposed Gaussian process is constrained by the Rankine-Hugoniot jump conditions between the various shocked material states to construct a thermodynamically consistent covariance function. This leads to the formulation of an optimization problem over a small number of interpretable hyperparameters and enables the identification of regime transitions, from a leading elastic wave to trailing plastic and phase transformation waves. Shock Hugoniots are an important measure for understanding material behavior under extreme conditions, including for the development of equations of state and determining material properties such as the Hugoniot Elastic Limit, but they are costly to generate through large-scale molecular dynamics simulations or shock experiments. Under these constraints, the proposed methodology establishes Hugoniot curves from a limited number of molecular dynamics simulations. We consider silicon carbide as a representative material and Molecular Dynamics simulations are performed using a reverse ballistic approach. The framework reproduces the Hugoniot curve with satisfactory accuracy while also quantifying the uncertainty in the predictions using the Gaussian Process posterior. These uncertain Hugoniot predictions can then be used to calibrate equation of state models, estimate material properties, or inform future experimental and/or simulation campaigns.

cs.CE

Multi-Fidelity Predictive Model for Shock Response of Energetic Materials Using Conditional U-Net

Mapping microstructure to properties is central to materials science. Perhaps most famously, the Hall-Petch relationship relates average grain size to strength. More challenging has been deriving relationships for properties that depend on subtle microstructural features and not average properties. One such example is the initiation of energetic materials under dynamical loading, dominated by energy localization on microstructural features such as pores, cracks, and interfaces. We propose a conditional convolutional neural network to predict the shock-induced temperature field as a function of shock strength, for a wide range of microstructures, and obtained via two different simulation methods. The proposed model, denoted MISTnet2, significantly extends prior work that was limited to a single shock strength, model, and type of microstructure. MISTnet2 can contribute to bridging atomistics with coarse-grain simulations and enable first principles predictions of detonation initiation and safety of this class of materials.

cond-mat.mtrl-sci

Active Learning Discovery of High Temperature Oxidation Resistant Refractory Complex Concentrated Alloys

Refractory complex concentrated alloys (RCCAs) are of significant interest for advanced high-temperature applications, owing to their broad compositional range and potential for attractive mechanical properties and oxidation resistance. However, their compositional complexity poses significant challenges to conventional alloy discovery methodologies. In this study, an active learning framework is introduced that integrates Gaussian process regression with Bayesian global optimization to accelerate identification of oxidation-resistant RCCAs. Focusing on aluminum-containing quaternary systems, alloy and oxide descriptors were used to predict oxidation performance at 1000$^\circ$C. Beginning with a dataset of 81 experimentally validated RCCAs, this framework was used to iteratively select alloy batches (five alloys per batch) with optimization based on a balance between exploration and exploitation to minimize associated experimental costs. After six iterations, two alloys were identified (nominal Al$_{30}$Mo$_5$Ti$_{15}$Cr$_{50}$ and Al$_{40}$Mo$_5$Ti$_{30}$Cr$_{25}$) that exhibited specific mass gains less than 1 mg/cm$^2$ at 1000$^\circ$C in air. Both of these alloys formed adherent external $\alpha$-Al$_2$O$_3$ scales and exhibited parabolic oxidation kinetics consistent with diffusion-limited scale growth. Furthermore, our multiobjective analysis demonstrates that these alloys simultaneously achieve high specific hardness ($>0.12$ HV$_{0.5}$m$^3$/kg) and thermal expansion compatibility with thermal barrier coating systems, positioning them as promising bond coat candidates. This work underscores the efficacy of active learning in traversing complex compositional landscapes, and offers a scalable strategy for the development of advanced materials suitable for extreme environments.

cond-mat.mtrl-sci

Machine learning descriptors for predicting the high temperature oxidation of refractory complex concentrated alloys

Refractory Complex Concentrated Alloys (RCCAs) can exhibit exceptional high-temperature strength, making such alloys promising candidates for high-temperature structural applications. However, current RCCAs do not possess the high-temperature oxidation resistance required to survive in oxidizing environments for more than a few hours at or above 1000$^\circ$C, without relying primarily on an environmental barrier coating. Here, we present a machine-learning framework designed to predict the oxidation-induced specific mass changes of RCCAs exposed for 24 h at 1000$^\circ$C in air, in order to support the search for oxidation-resistant alloys over a wide range of compositions. A database was constructed of experimental specific mass change data, upon oxidation at 900-1000$^\circ$C for 24 h in air, for 77 compositions comprised of simple elements, binary alloys, and higher-order elemental systems. We then developed a Gaussian Process Regression (GPR) model with physics-informed descriptors based on oxidation products, capturing the fundamental chemistry of oxide formation and stability. Application of this GPR model to the database yielded a MAE (mean absolute error) test score of 5.78 mg/cm$^2$, which was a significant improvement in accuracy relative to models only utilizing traditional alloy-based descriptors. Our model was used to screen over 5,100 quaternary RCCAs, revealing compositions with significantly lower predicted specific mass changes compared to existing literature sources. Overall, this work establishes a versatile and efficient strategy to accelerate the discovery of next-generation RCCAs with enhanced resistance to extreme environments.

cond-mat.mtrl-sci

Bridging the Synthesizability Gap in Perovskites by Combining Computations, Literature Data, and PU Learning

Among emerging energy materials, halide and chalcogenide perovskites have garnered significant attention over the last decade owing to the abundance of their constituent species, low manufacturing costs, and their highly tunable composition-structure-property space. Navigating the vast perovskite compositional landscape is possible using density functional theory (DFT) computations, but they are not easily extended to predictions of the synthesizability of new materials and their properties. As a result, only a limited number of compositions identified to have desirable optoelectronic properties from these calculations have been realized experimentally. One way to bridge this gap is by learning from the experimental literature about how the perovskite composition-structure space relates to their likelihood of laboratory synthesis. Here, we present our efforts in combining high-throughput DFT data with experimental labels collected from the literature to train classifier models employing various materials descriptors to forecast the synthesizability of any given perovskite compound. Our framework utilizes the positive and unlabeled (PU) learning strategy and makes probabilistic estimates of the synthesis likelihood based on DFT- computed energies and the prior existence of similar synthesized compounds. Our data and models can be readily accessed via a Findable, Accessible, Interoperable, and Reproducible (FAIR) nanoHUB tool.

cond-mat.mtrl-sci

Exploration of Hexagonal, Layered Carbides and Nitrides as Ultra-High Temperature Ceramics

Layered, hexagonal crystal structures, like zeta and eta phases, play an important role in ultra-high temperature ceramics, often significantly increasing toughness of carbide composites. Despite their importance open questions remain about their structure, stability, and compositional pervasiveness. We use high-throughput density functional theory to characterize the thermodynamic stability and elastic constants of layered carbides and nitrides M$_{n+1}$X$_{n}$ with $n$ = 1, 2, and 3, $M$ = Ta, Ti, Hf, Zr, Nb, Mo, V, W, Sc, Cr, Mn and $X$ = C, N. The stacking sequences explored are inspired by the possible use of MXenes as precursors to enable relatively low temperature processing of high-temperature ceramics. We identified 67 new hexagonal, layered materials with thermal stability comparable or better than previously observed zeta phases. To assess their potential for high temperature applications, we used machine learning and physics-based models with DFT inputs to predict their melting temperatures and discovered several candidates on par with the current state of the art zeta-like phases and five with predicted melting temperatures above 2500 K. The findings expand the range of chemistries and structures for high-temperature applications.

cond-mat.mtrl-sci

Predictive models for strain energy in condensed phase reactions

Molecular modeling of thermally activated chemistry in condensed phases is essential to understand polymerization, depolymerization, and other processing steps of molecular materials. Current methods typically combine molecular dynamics (MD) simulations to describe short-time relaxation with a stochastic description of predetermined chemical reactions. Possible reactions are often selected on the basis of geometric criteria, such as a capture distance between reactive atoms. Although these simulations have provided valuable insight, the approximations used to determine possible reactions often lead to significant molecular strain and unrealistic structures. We show that the local molecular environment surrounding the reactive site plays a crucial role in determining the resulting molecular strain energy and, in turn, the associated reaction rates. We develop a graph neural network capable of predicting the strain energy associated with a cyclization reaction from the pre-reaction, local, molecular environment surrounding the reactive site. The model is trained on a large dataset of condensed-phase reactions during the activation of polyacrylonitrile (PAN) obtained from MD simulations and can be used to adjust relative reaction rates in condensed systems and advance our understanding of thermally activated chemical processes in complex materials

cond-mat.mtrl-sci

A collaborative digital twin built on FAIR data and compute infrastructure

The integration of machine learning with automated experimentation in self-driving laboratories (SDL) offers a powerful approach to accelerate discovery and optimization tasks in science and engineering applications. When supported by findable, accessible, interoperable, and reusable (FAIR) data infrastructure, SDLs with overlapping interests can collaborate more effectively. This work presents a distributed SDL implementation built on nanoHUB services for online simulation and FAIR data management. In this framework, geographically dispersed collaborators conducting independent optimization tasks contribute raw experimental data to a shared central database. These researchers can then benefit from analysis tools and machine learning models that automatically update as additional data become available. New data points are submitted through a simple web interface and automatically processed using a nanoHUB Sim2L, which extracts derived quantities and indexes all inputs and outputs in a FAIR data repository called ResultsDB. A separate nanoHUB workflow enables sequential optimization using active learning, where researchers define the optimization objective, and machine learning models are trained on-the-fly with all existing data, guiding the selection of future experiments. Inspired by the concept of ``frugal twin", the optimization task seeks to find the optimal recipe to combine food dyes to achieve the desired target color. With easily accessible and inexpensive materials, researchers and students can set up their own experiments, share data with collaborators, and explore the combination of FAIR data, predictive ML models, and sequential optimization. The tools introduced are generally applicable and can easily be extended to other optimization problems.

cs.AI

Spall strength of symmetric tilt grain boundaries in 6H silicon carbide

Characterizing microstructural effects on the dynamical response of materials is challenging due to the extreme conditions and the short timescales involved. For example, little is known about how grain boundary characteristics affect spall strength. This study explores 6H-SiC bicrystals under shock waves via large-scale molecular dynamics simulations. We focused on symmetric tilt grain boundaries with a wide range of misorientations and found that spall strength and dynamical fracture surface energy are strongly affected by the grain boundary microstructure, especially the excess free volume. Grain boundary energy also plays a considerable role. As expected, low-angle grain boundaries tend to have higher spallation strengths. We also extracted cohesive models for the dynamical strength of bulk systems and grain boundaries that can be used in continuum simulations.

cond-mat.mtrl-sci

Harnessing Machine Learning for Quantum-Accurate Predictions of Non-Equilibrium Behavior in 2D Materials

Accurately predicting the non-equilibrium mechanical properties of two-dimensional (2D) materials is essential for understanding their deformation, thermo-mechanical properties, and failure mechanisms. In this study, we parameterize and evaluate two machine learning (ML) interatomic potentials, SNAP and Allegro, for modeling the non-equilibrium behavior of monolayer MoSe2. Using a density functional theory (DFT) derived dataset, we systematically compare their accuracy and transferability against the physics-based Tersoff force field. Our results show that SNAP and Allegro significantly outperform Tersoff, achieving near-DFT accuracy while maintaining computational efficiency. Allegro surpasses SNAP in both accuracy and efficiency due to its advanced neural network architecture. Both ML potentials demonstrate strong transferability, accurately predicting out-of-sample properties such as surface stability, inversion domain formation, and fracture toughness. Unlike Tersoff, SNAP and Allegro reliably model temperature-dependent edge stabilities and phase transformation pathways, aligning closely with DFT benchmarks. Notably, their fracture toughness predictions closely match experimental measurements, reinforcing their suitability for large-scale simulations of mechanical failure in 2D materials. This study establishes ML-based force fields as a powerful alternative to traditional potentials for modeling non-equilibrium mechanical properties in 2D materials.

cond-mat.mtrl-sci

Data Fusion of Deep Learned Molecular Embeddings for Property Prediction

Data-driven approaches such as deep learning can result in predictive models for material properties with exceptional accuracy and efficiency. However, in many applications, data is sparse, severely limiting their accuracy and applicability. To improve predictions, techniques such as transfer learning and multitask learning have been used. The performance of multitask learning models depends on the strength of the underlying correlations between tasks and the completeness of the data set. Standard multitask models tend to underperform when trained on sparse data sets with weakly correlated properties. To address this gap, we fuse deep-learned embeddings generated by independent pretrained single-task models, resulting in a multitask model that inherits rich, property-specific representations. By reusing (rather than retraining) these embeddings, the resulting fused model outperforms standard multitask models and can be extended with fewer trainable parameters. We demonstrate this technique on a widely used benchmark data set of quantum chemistry data for small molecules as well as a newly compiled sparse data set of experimental data collected from literature and our own quantum chemistry and thermochemical calculations.

cs.LG

Modeling Framework to Predict Melting Dynamics at Microstructural Defects in TNT-HMX High Explosive Composites

Many high explosive (HE) formulations are composite materials whose microstructure is understood to impact functional characteristics. Interfaces are known to mediate the formation of hot spots that control their safety and initiation. To study such processes at molecular scales, we developed all-atom force fields (FFs) for Octol, a prototypical HE formulation comprised of TNT (2,4,6-trinitrotoluene) and HMX (octahydro-1,3,5,7-tetranitro-1,3,5,7-tetrazocine). We extended a FF for TNT and recasted it in a form that can be readily combined with a well-established FF for HMX. The resulting FF was extensively validated against experimental results and density functional theory calculations. We applied the new combined TNT-HMX FF to predict and rank surface and interface energies, which indicate that there is an energetic driver for coarsening of microstructural grains in TNT-HMX composites. Finally, we assess the impact of several microstructural environments on the dynamic melting of TNT crystal under ultrafast thermal loading. We find that both free surfaces and planar material interfaces are effective nucleation points for TNT melting. However, MD simulations show that TNT crystal is prone to superheating by at least 50 K on sub-nanosecond timescales and that the degree of superheating is inversely correlated with surface and interface energy. The modeling framework presented here will enable future studies on hot spot formation processes in accident scenarios that are governed by strong coupling between microstructural interfaces, material mechanics, momentum and energy transport, phase transitions, and chemistry.

cond-mat.mtrl-sci

Exploring the defect landscape and dopability of chalcogenide perovskite BaZrS3

BaZrS3 is a chalcogenide perovskite that has shown promise as a photovoltaic absorber, but its performance is limited because of defects and impurities that have a direct influence on carrier concentrations. Functional dopants that show lower donor-type or acceptor-type formation energies than naturally occurring defects can help tune the optoelectronic properties of BaZrS3. In this work, we applied first principles computations to comprehensively investigate the defect landscape of BaZrS3, including all intrinsic defects and a set of selected impurities and dopants. BaZrS3 intrinsically exhibits n-type equilibrium conductivity under both S-poor and S-rich conditions, which remains largely unchanged in the presence of O and H impurities. La and Nb dopants created stable donor-type defects, which made BaZrS3 even more n-type, whereas As and P dopants formed amphoteric defects with relatively high formation energies. This work highlights the difficulty of creating p-type BaZrS3 owing to the low formation energies of donor defects, both intrinsic and extrinsic. Defect formation energies were also used to compute expected defect concentrations and make comparisons with experimentally reported values. Our dataset of defects in BaZrS3 paves the path for training machine learning models to subsequently perform larger-scale prediction and screening of defects and dopants across many chalcogenide perovskites, including cation-site or anion-site alloys.

cond-mat.mtrl-sci

Thermodynamic Fidelity of Generative Models for Ising System

Machine learning has become a central technique for modeling in science and engineering, either complementing or as surrogates to physics-based models. Significant efforts have recently been devoted to models capable of predicting field quantities but the limitations of current state-of-the-art models in describing complex physics are not well understood. We characterize the ability of generative diffusion models and generative adversarial networks (GAN) to describe the Ising model. We find diffusion models trained using equilibrium configurations obtained using Metropolis Monte Carlo for a range of temperatures around the critical temperature can capture average thermodynamic variables across the phase transformation and extrapolate to higher and lower temperatures. The model also captures the overall trends of physical properties associated with fluctuations (specific heat and susceptibility) except at the non-ergodic low temperatures and non-trivial scale-free correlations at the critical temperature, albeit with some difference in the critical exponent compared to Monte Carlo simulations. GANs perform more poorly on thermodynamic properties and are susceptible to mode-collapse without careful training. This investigation highlights the potential and limitations of generative models in capturing the complex phenomena associated with certain physical systems.

cond-mat.stat-mech