Search NASASearch

SEARCH · Search NASA

Results for “machine learning potential”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Segmentation and Classification of Fission as Pores in Reactor Irradiated Annular U–10Zr Metallic Fuel Using Machine Learning Models

Metallic fuels, particularly U—10Zr, are promising candidates for next-generation sodium-cooled fast reactors. Irradiation of nuclear fuels in reactors can lead to the formation of solid and gas fission product which subsequently forms microstructural pores, deteriorating fuel performance. Due to the massive amount of pores and complex phases formed, a quantitative description of fission gas pores is not yet available, preventing the development of microstructure-informed fuel performance modeling for fuel qualification. This paper applied a pre-trained deep learning model to ~10,260 high magnification scanning electron microscopy images. This method increased the accuracy of fission gas pore segmentation and allows statistical features to be extracted which cannot be achieved manually. A pre-trained decision tree model worked on the segemenation results and further classified the pores into different categories to produce a correlation between the pores, movement of lanthanides, and temperature gradient during irradiation. Finally, this paper emphasizes the potentials of machine learning models to accelerate fuel research, development, and qualification for advanced reactors.

36 MATERIALS SCIENCE

Predictive analytics of selections of russet potatoes

We explore the application of machine learning algorithms specifically to enhance the selection process of Russet potato (Solanum tuberosum L.) clones in breeding trials by predicting their suitability for advancement. This study addresses the challenge of efficiently identifying high-yield, disease-resistant, and climate-resilient potato varieties that meet processing industry standards. Leveraging manually collected data from trials in the state of Oregon, we investigate the potential of a wide variety of state-of-the-art binary classification models. The dataset includes 1086 clones, with data on 38 attributes recorded for each clone, focusing on yield, size, appearance, and frying characteristics, with several control varieties planted consistently across four Oregon regions from 2013 to 2021. We conduct a comprehensive analysis of the dataset that includes preprocessing, feature engineering, and imputation to address missing values. We focus on several key metrics such as accuracy, F1-score, and Matthews correlation coefficient (MCC) for model evaluation. The top-performing models, namely a feedforward neural network classifier (Neural Net), a histogram-based gradient boosting classifier (HGBC), and a support vector machine classifier (SVM), demonstrate consistent and significant results. To further validate our findings, we conducted a simulation study using the aims, data-generating mechanisms, estimands, methods, and performance measures (ADEMP) framework, simulating different data-generating scenarios to assess model robustness and performance through true positive, true negative, false positive, and false negative distributions, area under the receiver operating characteristic curve (AUC-ROC) and MCC. The simulation results highlight that non-linear models like SVM and HGBC consistently show higher AUC-ROC and MCC than logistic regression, thus outperforming the traditional linear model across various distributions, and emphasizing the importance of model selection and tuning in agricultural trials. Variable selection further enhances model performance and identifies influential features in predicting trial outcomes. The findings emphasize the potential of machine learning in streamlining the selection process for potato varieties, offering benefits such as increased efficiency, substantial cost savings, and judicious resource utilization. Our study contributes insights into precision agriculture and showcases the relevance of advanced technologies for informed decision-making in breeding programs.

60 APPLIED LIFE SCIENCES

Prediction of electric and magnetic fields from spectral data using machine learning algorithms for Doppler-free saturation spectroscopy diagnostics

The prediction of electric and magnetic field amplitudes from atomic spectral data is critical for plasma control in fusion devices such as tokamaks. Conventional approaches that rely on physics-based models are computationally expensive and unsuitable for real-time applications. In this work, we develop and benchmark three machine learning algorithms—simulation-based inference (SBI), fully connected neural networks (FCNN), and histogram-based gradient boosting regression (GBR-Hist)—to infer field intensities directly from Doppler-free saturation spectroscopy (DFSS) spectra. Synthetic datasets of spectra were generated using the EZSSS code and evaluated both with and without added Poisson noise to mimic experimental conditions. We find that SBI achieves the highest accuracy and robustness, FCNN provides a strong balance of accuracy and computational efficiency for real-time applications, and GBR-Hist offers the fastest inference but is more sensitive to noise. Furthermore, these results demonstrate the potential of machine learning to accelerate DFSS analysis and enhance its utility for plasma diagnostics and control.

Doppler-free saturation spectroscopy

A multimodal large language model for materials science

Understanding and predicting the properties of inorganic materials is crucial for accelerating advancements in materials science and driving applications in energy, electronics and beyond. Integrating material structure data with language-based information through multimodal large language models (LLMs) offers great potential to support these efforts by enhancing human–artificial intelligence interaction. However, a key challenge lies in integrating atomic structures at full resolution into LLMs. In this work, we introduce MatterChat, a versatile structure-aware multimodal LLM that unifies material structural data and textual inputs into a single cohesive model. MatterChat uses a bridging module to effectively align a pretrained universal machine learning interatomic potential with a pretrained LLM, reducing training costs and enhancing flexibility. Our results demonstrate that MatterChat greatly improves performance in material property prediction and human–artificial intelligence interaction, surpassing general-purpose LLMs such as GPT-4. We also demonstrate its usefulness in applications such as more advanced scientific reasoning and step-by-step material synthesis.

Tang, Yingheng [Lawrence Berkeley National Laborat

Next-Generation Materials Design: Quantum Mechanics and Data-Driven Modeling

The future of materials design is rapidly advancing through the combination of quantum mechanics and data-driven modeling. These approaches integrate quantum principles with advanced data analysis, enabling precise insights into material behavior. This talk will highlight recent progress in using these methods for computational design, particularly in high-entropy alloy catalysts, emphasizing the role of hierarchical machine-learning architectures for accurate predictions. Additionally, I will discuss our work on developing machine learning interatomic potentials (MLPs) for single-element metals, metal oxides, and alloys under extreme conditions, focusing on melting behavior and phase properties at high temperatures and pressures. We have also refined our MLP models to capture dynamic surface interactions, such as CO2 and CO adsorption on MgO, using both static and molecular dynamics simulations. These models maintain high accuracy while significantly reducing computational costs compared to first-principles calculations. By enabling efficient and accurate simulations, this work supports broader community adoption, optimizes datasets for materials discovery, and extends the accessible time, size, and environmental conditions beyond the limits of experiments and traditional simulations.

machine learning

Melting curves of atomic hydrogen and deuterium calculated using path-integral Monte Carlo

We calculate the melting line of atomic hydrogen and deuterium up to 900 GPa with path-integral Monte Carlo using a machine-learned interatomic potential. We improve upon previous simulations of melting by treating the electrons with reptation quantum Monte Carlo, and by performing solid and liquid simulations using isothermal-isobaric path-integral Monte Carlo. Here, the resulting melting line for atomic hydrogen is higher than previous estimates. There is a small but resolvable decrease in the melting temperature as pressure is increased, which can be attributed to quantum effects.

08 HYDROGEN

Machine learning-assisted identification of potential sources of bias in measurements of prompt-fission neutron spectra

Unrecognized sources of uncertainty (USU) can bias the reported mean and/or covariance of experimental nuclear data. These biases, in turn, can propagate through evaluated nuclear data to application simulations or may poorly inform nuclear theory that is fitted to the experimental data. Such unknown sources of bias must be tied to the inherent physical constituents of the measurements such as the characteristics of a detector response or a background reduction technique. Here, in this article, a sparse Bayesian learning model is used to support experts in their efforts to identify and characterize USU in experimental prompt fission neutron spectra (PFNS) for spontaneous fissioning of 252 Cf by linking observed biases to features of the measurement system. Three different bias components were found. The first acts as a verification case for the algorithm as it identifies a bias coming from a well-known source related to the use of 6 Li in the neutron detection system. The second two cases demonstrate how this method can benefit the evaluation of experimental nuclear data by identifying, quantifying, and relating unknown biases to potential causes.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

HydraGNN v5.0

HydraGNN v5.0 expands the code base into a more portable, scalable, and flexible framework for scientific graph learning, with particular strength in atomistic machine-learning interatomic potentials and large-scale distributed training. The release adds Fully Sharded Data Parallel (FSDP) support alongside existing DDP and DeepSpeed paths, including FSDP-aware checkpointing and optimizer integration, and introduces a configurable multi-precision training workflow supporting FP32, BF16, and FP64 across GPUs and Intel XPUs. For atomistic modeling, HydraGNN v5.0 strengthens its MLIP capabilities through dynamic graph construction at every forward pass, energy-conserving force prediction via automatic differentiation, and per-atom energy loss formulations, while extending EGNN models to properly handle periodic boundary conditions. The release also broadens model expressiveness through graph-level attribute conditioning, adds new multi-task and model-parallel extensions such as MACE support and encoder/decoder branch optimization, and expands application coverage with integrated examples for datasets including OC25, Nabla2-DFT, QCML, Open Polymers 2026, and OPF. In parallel, HydraGNN v5.0 improves production readiness through performance optimizations for large-scale runs, stratified sampling and linear-regression preprocessing utilities, and tested installation scripts for DOE supercomputers including Frontier, Aurora, Perlmutter, and Andes. Overall, the release advances HydraGNN as a robust software platform for scalable graph neural networks across materials science, chemistry, and scientific machine learning workflows

Lupo Pasini, Massimiliano [Oak Ridge National Labo

Revisiting point defect thermodynamics in group IVB and VB transition metal carbides

We present a comprehensive re-examination of point defect thermodynamics in group IVB and VB transition metal carbides (TMCs) with the rocksalt structure using a combination of density functional theory (DFT) calculations and a statistical mechanical Wagner-Schottky model within the canonical ensemble. The most stable configurations of point defects were discovered using basin-hopping global optimization, driven by either a machine learning interatomic potential (MLIP) or DFT. A key finding is the identification of previously unreported dicarbon antisites—a C–C dimer occupying a metal site—as the structural (constitutional) defects on the carbon-rich side of stoichiometry in all group IVB and VB TMCs except TaC. Furthermore, dicarbon antisite-containing thermal defect complexes, such as quadruple and interbranch defects, can dominate in TMCs under specific stoichiometric and temperature conditions. In conclusion, by incorporating dicarbon antisites into the defect landscape, this work provides a revised understanding of the thermodynamics of point defects in TMCs.

Carbides

Protonation Dynamics of Confined Ethanol–Water Mixtures in H-ZSM-5 from Machine Learning-Driven Metadynamics

Zeolites are indispensable heterogeneous catalysts in industrial chemical processes, valued for their strong Brønsted acidity, well-defined microporous frameworks, and tunable pore structures. Their catalytic activity arises primarily from Brønsted acid sites (BAS), typically present as bridging hydroxyl groups (Si–OH–Al). Under aqueous reaction conditions, these protons interact dynamically with water and alcohol molecules, leading to complex solvation and protonation behavior within confined pores. In this study, we investigate the protonation equilibrium occurring between ethanol and water at the BAS of acidic zeolites under varying hydration levels, i.e., C2H5OH–(H2O)n, n=1–4. Local structure was analyzed through an adaptive-learning global optimization algorithm, while enhanced sampling molecular dynamics simulations with Well-Tempered Metadynamics (WMetaD) and machine learning interatomic potentials (MLPs) provide free-energy surfaces (FES) at variable hydration levels. The results reveal a strong dependence of proton localization on the degree of hydration. At low hydration (1 water molecule), the proton resides predominantly on ethanol; with 2 water molecules, it shifts toward water, and at higher hydration (3 or more water molecules), it becomes extensively delocalized over the water cluster. These findings underscore the critical role of solvation in modulating acid site behavior and suggest that a minimum of three water molecules is necessary to fully stabilize the proton on water within the zeolite framework. This solvation threshold has significant implications for catalytic processes, particularly in biomass conversion reactions where alcohol protonation is a key step in dehydration mechanisms.

machine learning

Role of Wadsley Defects and Cation Disorder to Enhance MoNb 12 O 33 Diffusion

Wadsley-Roth (WR) niobates have emerged as high-rate anode materials that can combine rapid ionic diffusion with good electronic conductivity. WR compounds have been defect-enhanced by limited annealing, however, such materials often contain multiple types of defects. In particular, both Wadsley defects (variable block size) and transition metal disorder have the potential to modify transport rates, however the corresponding effects are not well understood mechanistically. Here, MoNb 12 O 33 (MNO) was calcined at two different temperatures to compare a defect-rich condition (MNO-800) with a proximal order-rich condition (MNO-900) as assessed through XRD, XANES, EXAFS, and STEM characterizations. Galvanostatically cycled lithium half cells of MNO-800 exhibited additional capacity (307 mAhg −1 at 0.1C, 4.66% higher) and improved high-rate capacity of 200 mAhg −1 at 10C. ICI-based overpotential analysis identified solid state diffusion as the dominant rate limiting process where MNO-800 correspondingly exhibited ∼3X faster capacity-weighted diffusivity. A machine-learning interatomic potential was trained to density functional theory and then applied with molecular dynamics (MLIP-MD) to examine the possible roles of Wadsley defects and transition metal disorder. For both defect-types, Li was found to populate and activate fast diffusion paths from window sites at lower extents of lithiation as compared to the order-rich model.

defect

Shadow molecular dynamics for flexible multipole models

Shadow molecular dynamics provide an efficient and stable atomistic simulation framework for flexible charge models with long-range electrostatic interactions. Shadow molecular dynamics simulations are driven by approximate “shadow” Born–Oppenheimer potentials for which the exact charges and forces are directly accessible without relying on costly (and approximate) iterative solvers. While previous implementations have been limited to atomic monopole charge distributions, we extend this approach to flexible multipole models. We derive detailed expressions for the shadow energy functions, potentials, and force terms, explicitly incorporating monopole–monopole, dipole–monopole, and dipole–dipole interactions. In our formulation, both atomic monopoles and atomic dipoles are treated as extended dynamical variables alongside the propagation of the nuclear degrees of freedom. We demonstrate that introducing the additional dipole degrees of freedom preserves the stability and accuracy previously seen in monopole-only shadow molecular dynamics simulations. In addition, we present a shadow molecular dynamics scheme where the monopole charges are held fixed while the dipoles remain flexible. Our extended shadow dynamics provide a framework for stable, computationally efficient, and versatile molecular dynamics simulations involving long-range interactions between flexible multipoles. This is of particular current interest in combination with machine-learned interatomic potentials, including long-range electrostatic interactions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Multistage nucleation pathway in LiF molten salt mirrors the crystal–melt interface structure

Despite over a century of studies, fundamental questions remain about the processes governing crystal nucleation from melts or solutions. Research over the past three decades has presented mounting evidence for kinetic pathways of crystal nucleation that are more complex than envisioned by the simplest forms of classical theory. Such observations have been presented for colloidal and elemental systems with covalent and metallic bonding. Despite the technological and geochemical importance of molten salts, similar studies for these ionically bonded systems are currently lacking. Here we develop a machine learning interatomic potential for a model ionic system: LiF. The potential features quantum-level accuracy for both liquid and multiple solid polymorphs over wide temperature and pressure ranges and accurately reproduces experimentally measured properties. Thanks to the efficiency of the potential, which enables microsecond-scale molecular dynamics simulations, induction times for nucleation of LiF solids from their melts are computed over a range of undercoolings. With the aid of a set of robust local order parameters established here, the simulations reveal that homogeneous crystal nucleation in undercooled melts preferentially initiates from liquid regions showing slow dynamics and high bond orientational order simultaneously, and the second-shell order of both precritical nuclei and the surface of postcritical nuclei is dominated by hexagonal close packing and body-centered cubic local structure, even though the nucleus core is dominated by face-centered cubic structure corresponding to the stable rocksalt crystal structure. Finally, we establish a connection between the crystallization pathway and the equilibrium crystal-melt interface structure.

Applied Physical Sciences

Role of electron correlation on the adenine dimer interaction for non-equilibrium geometries: a benchmark Quantum Monte Carlo study

The accurate description of non-covalent interactions is critical for understanding the structure, dynamics, and eventual function of biomolecules. The adenine dimer serves as a benchmark system for computational methods due to its role in nucleic acid structures and its rich conformational landscape. In this study, we employ benchmark diffusion quantum Monte Carlo (DMC) methods to investigate the relative energies and role of electron correlation on a set of adenine dimer conformations generated via a search of the potential energy landscape using the global optimizer algorithm. Relative DMC energies are compared against a wide range of density functional theory (DFT) approximation results. We find that although most of the DFT functionals perform well for low-energy structures, their accuracy varies significantly for higher-energy conformations, including stacked and T-shaped structures. A large fraction of the variation is due to the treatment of the van der Waals interaction. BLYP, B3LYP, and PBE0 significantly improve with added D4 dispersion, while the recent r2SCAN-D4 and ωB97M-V functionals show the least scatter and closest agreement with the DMC. These findings highlight the delicate nature of these interactions in biomolecular systems and provide guidance for simulations of their structure and dynamics and for the development of machine learned interatomic potentials.

Washburn, Laurel [ORNL] (ORCID:0000000324179335)

Modeling graphene sheet growth and dynamical matrix calculations using molecular dynamics

Molecular dynamics (MD) has been an incredibly useful tool to model physical processes that were synthesized experimentally but not fully understood. MD, through the use of semi-empirical inter-atomic potentials, has allowed understanding of different physical processes in materials science. Yet as well as providing useful insights into materials science, molecular dynamics has a wider range of usability. In this report, I will be detailing how MD can be used to study graphene formation from a carbon liquid which requires high temperatures and pressures. Beyond this, I will describe the usefulness of MD for understanding the physics for phonon transport quantum sensors. To do this, MD was employed to determine the dynamical matrix by treating atoms as coupled oscillators. An accurate understanding of the dynamical matrix of a system is required to calculate the non-equilibrium Green’s function used to describe the phonon transport within phonon wave-guides. I found that, across multiple pressures and temperatures, randomly placed carbon atoms will show evidence of pent-first formation with semi-empirical models. Density functional theory (DFT), on the other hand, was too computationally expensive to use for full scale MD simulations, but we have the possibility of training a machine learned interatomic potential to approximate DFT for carbon in the environments being studied for pent-first graphene sheet formation.

36 MATERIALS SCIENCE

Machine-learned quantum molecular dynamics calculations of warm dense equation of state and ionic transport coefficients of deuterated water

White dwarf models require accurate equations of state and ionic transport coefficients in the warm dense matter regime, where kinetic theory models and tabulated equations of state are often inaccurate. In this work, spectral-partitioned density functional theory and machine-learned interatomic potentials are combined to perform large-scale, first-principles quantum molecular dynamics simulations of deuterated water (D 2 O) near the principal Hugoniot. This approach retains Kohn-Sham accuracy while achieving orders-of-magnitude speedup, yielding converged equation of state and transport properties over a broad pressure and temperature range. The results reveal the thermodynamic conditions under which ionic transport models for interdiffusivity and shear viscosity converge and identify those in closest agreement with density functional theory benchmarks at temperatures in the warm dense matter regime. The present framework extends first-principles transport calculations to higher temperatures than previously achieved, and provides an efficient, scalable, and general approach for studying transport properties in complex multicomponent mixtures.

79 ASTRONOMY AND ASTROPHYSICS

How Silica Surface Chemistry Modulates Interfacial Water: Insights from Machine Learning Molecular Dynamics

Controlling water structure and dynamics at silica interfaces are central to a wide range of technologies, including protective oxide layers for solar water splitting and nanoporous membranes. In this work, we develop a machine learning interatomic potential, trained via active learning, to achieve ab initio accuracy for water confined between hydroxylated silica surfaces over a range of silanol coverages and slit widths. We find that partially hydroxylated surfaces (50 and 75% OH) support stronger water−surface hydrogen bonding and more extended interfacial density profiles than fully hydroxylated (100% OH) surfaces, indicating that increasing OH coverage does not necessarily strengthen interfacial hydrogenbond networks. Translational diffusion decreases approximately linearly with slit width and OH coverage, whereas rotational dynamics respond nonlinearly. In particular, at the smallest slit width of 5 Å, 75% OH coverage produces an enhanced local tetrahedral ordered interfacial network that strongly suppresses reorientation, while 100% coverage yields a crowded, disordered interfacial layer that also hinders rotation. In contrast, the 50% OH coverage is sufficiently sparse that it does not markedly alter water structure or dynamics under confinement. These results show that coupled control of pore size and surface chemistry enables nonlinear tuning of interfacial water structure and transport, providing a design strategy for optimizing porous silica for either enhanced interfacial stability and controlled reactivity or rapid and selective transport.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Machine‐Learning‐Driven Exploration of Surface Reconstructions of Reduced Rutile TiO 2

Abstract Titanium dioxide (TiO 2 ) is widely used as a catalyst support due to its stability, tunable electronic properties, and surface oxygen vacancies, which are crucial for catalytic processes such as the reverse water‐gas shift (RWGS) reaction. Reduced TiO 2 surfaces undergo complex surface reconstructions that endow unique properties but are computationally challenging to describe. In this study, we utilize machine‐learning interatomic potentials (MLIPs) integrated with an active‐learning workflow to efficiently explore reduced rutile TiO 2 surfaces. This approach enabled the prediction of a phase diagram as a function of oxygen chemical potential, revealing a variety of reconstructed phases, including a previously unreported subsurface shear plane structure. We further investigate the electronic properties of these surfaces and validate our results by comparing experimental and theoretical high‐resolution transmission electron microscopy (HRTEM). Our findings provide new insights into how extreme surface reductions influence the structural and electronic properties of TiO 2 , with potential implications for catalyst design.

Lee, Yonghyuk [Chemistry and Biochemistry Universi