Search NASASearch

SEARCH · Search NASA

Results for “machine learning potentials”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

A review of displacement cascade simulations using molecular dynamics emphasizing interatomic potentials for TPBAR components

This review explores molecular dynamics simulations for studying radiation damage in Tritium Producing Burnable Absorber Rod (TPBAR) materials, emphasizing the role of interatomic potentials in displacement cascades. Recent machine learning potentials (MLPs), trained on quantum data, enhance prediction accuracy over traditional models like EAM. We highlight temperature, PKA energy, and composition effects on damage evolution in TPBAR components, recommending suitable potentials and discussing advancements for materials in extreme radiation environments.

36 MATERIALS SCIENCE

Benchmarking the performance of uncertainty quantification methods for neural network-based interatomic potentials

Machine-learned interatomic potentials (ML-IAPs) continue to gain popularity as accurate, computationally efficient replacements for traditional, physics-based interatomic potentials and expensive ab initio methods. Uncertainty quantification (UQ) of ML-IAPs is a growing area of research as UQ is critical in many applications of IAPs, such as developing curated datasets, active learning-based data augmentation, self-improving models, and estimating the uncertainty of molecular dynamics simulations. In this paper, we construct and benchmark a series of different neural network potentials (NNPs) with varying network architectures to determine the performance of these models with respect to both the mean and uncertainty calibration error. Each NNP method is specifically designed to predict either epistemic or aleatoric uncertainty with particular focus on the differences in behavior between the epistemic and aleatoric uncertainty estimates. We benchmark these methods using multiple datasets common in the ML-IAP literature. The results show that the aleatoric uncertainty from single-shot model architectures is a competitive alternative to ensemble-based epistemic uncertainty predictions in regions of sufficient data-density. However, in regions where the representative data is sparse, aleatoric uncertainty models tend to overpredict and epistemic methods tend to underpredict the actual model error. We conclude that the type of UQ is crucial when discussing performance of probabilistic model results as different methods have different performance characteristics depending on the regime in which they are evaluated. Therefore, the type of UQ method should be carefully evaluated against both the data characteristics and requirements for the intended application.

97 MATHEMATICS AND COMPUTING

Constraining Galaxy-Halo connection using machine learning

We investigate the potential of machine learning (ML) methods to model small-scale galaxy clustering for constraining Halo Occupation Distribution (HOD) parameters. Our analysis reveals that while many ML algorithms report good statistical fits, they often yield likelihood contours that are significantly biased in both mean values and variances relative to the true model parameters. This highlights the importance of careful data processing and algorithm selection in ML applications for galaxy clustering, as even seemingly robust methods can lead to biased results if not applied correctly. ML tools offer a promising approach to exploring the HOD parameter space with significantly reduced computational costs compared to traditional brute-force methods if their robustness is established. Using our ANN-based pipeline, we successfully recreate some standard results from recent literature. Properly restricting the HOD parameter space, transforming the training data, and carefully selecting ML algorithms are essential for achieving unbiased and robust predictions. Among the methods tested, artificial neural networks (ANNs) outperform random forests (RF) and ridge regression in predicting clustering statistics, when the HOD prior space is appropriately restricted. We demonstrate these findings using the projected two-point correlation function (w p (r p )), angular multipoles of the correlation function (ξ ℓ (r)), and the void probability function (VPF) of Luminous Red Galaxies from Dark Energy Spectroscopic Instrument mocks. Our results show that while combining w p (r p ) and VPF improves parameter constraints, adding the multipoles ξ 0 , ξ 2 , and ξ 4 to w p (r p ) does not significantly improve the constraints.

cosmology

From bulk to surface: Structure and dynamics of amorphous alumina from deep potential molecular dynamics

Understanding the atomic-scale structure and dynamics of amorphous oxide surfaces is essential for interpreting their chemical reactivity, mechanical stability, and interfacial behavior, yet direct experimental characterization remains challenging. We employ Deep Potential (DP) molecular dynamics to generate large-scale, ab initio -quality models of amorphous Al 2 O 3 bulk glasses and melt-quenched free surfaces, enabling a quantitative analysis of both structure and relaxation dynamics with statistical confidence inaccessible to direct ab initio simulation. The trained DP model reproduces experimental liquid and glass structure, captures the cooling-rate dependence of the bulk glass transition, and corrects systematic biases in the polyhedral populations predicted by widely used classical force fields. At the free surface, mass density recovers to bulk values over ~10 Å, while local coordination requires a slightly wider subsurface region to fully converge. The outermost layer is oxygen-enriched, exhibits altered polyhedral connectivity with contracted Al–O bonds, and hosts a broad population of under-coordinated motifs (notably AlO 3 and OAl 2 ) whose abundances are governed by glass stability. These under-coordinated surface motifs exhibit distinct vibrational signatures and occur as locally paired Lewis acid and Brønsted base sites consistent with bond-valence compensation, yet remain spatially dispersed rather than aggregating into extended clusters. Despite this pronounced structural heterogeneity, surface relaxation and the glass-transition temperature remain comparable to their bulk counterparts, suggesting that the disordered surface is kinetically stable once formed. Together, these results establish a molecular-level picture of amorphous alumina surfaces and demonstrate the capability of machine-learned potentials to resolve structure–property relationships in disordered oxide interfaces.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

NNL.Fe.qSNAP-ZBL.2024.1: A Fe Spectral Neighbor Analysis Potential for Radiation Damage Simulations

The NNL.Fe.qSNAP-ZBL.2024.1 machine-learned potential (MLP) has been generated to support the development of an elemental body-centered cubic (BCC) Fe athermal recombination corrected neutron damage model and simulations of primary recoil atom (PRA) cascades in BCC Fe. This MLP is a quadratic spectral neighbor analysis potential (qSNAP) hybridized with the universal Ziegler-Beirsack-Littmark (ZBL) potential at short-range and is named according to Naval Nuclear Laboratory MLP naming conventions (NNL.material-system.MLP-type.year.version). Training set calculations for Fe are presented along with the subsequent MLP fitting procedure. A key criterion of the fitting procedure is that ZBL describes the short-range interaction with minimal impact on the MLP. The MLP is compared to density functional theory (DFT) predicted properties relevant to radiation damage simulation, including threshold displacement energies, for validation. The NNL.Fe.qSNAP-ZBL.2024.1 potential is considered suitable for molecular dynamics (MD) simulations of radiation defects up to 800 K and PRA cascades in BCC Fe up to around 10 keV. The potential can additionally be used on a limited basis for recoils of 10–20 keV, within which range the emergence of structures outside the training set in cascade simulations may cause system instabilities.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Iron Impurity Impairs the CO 2 Capture Performance of MgO: Insights from Microscopy and Machine Learning Molecular Dynamics

Magnesium oxide (MgO) is a promising sorbent for direct air capture (DAC) of carbon dioxide. Iron (Fe) is a common impurity in naturally occurring MgO and minerals used to produce MgO, yet a molecular-scale understanding of Fe-doping effects on carbonation is lacking. Here, in this study, we observed reduced carbonation performance in Fe-doped MgO experimentally. The energetics of adsorbing a (bi)carbonate ion on pristine and Fe-doped MgO(001) surfaces were further investigated using ab initio and machine learning potential molecular dynamics coupled with metadynamics simulations. Both pristine and Fe-doped surfaces exhibited a basic (OH – ) hydration layer, where the (bi)carbonate ion adsorption is thermodynamically favorable. However, the dissolution of surface Fe had smaller energy barriers and was more favorable than Mg. Leached Fe likely neutralized the near-surface basicity, yielding reduced reactivity on Fe-doped MgO. Our observations offer critical insights for material selection and emphasize the importance of evaluating the geologic origin of earth materials used for DAC.

36 MATERIALS SCIENCE

Facet-dependent structure and dissociation of water at pristine IrO 2 /water interfaces

Understanding the microscopic structure of water at metal oxide interfaces is crucial for advancing electrocatalysis. IrO 2 , specifically, has shown exceptional activity for electrochemical water oxidation, but we currently lack a fundamental understanding of how the surface structure of IrO 2 impacts water reactivity. In this work, we developed a machine learning potential trained to first-principles accuracy for modeling IrO 2 /water interfaces across different facets: (110), (100), (101), and (001). Using extensive machine learning molecular dynamics simulations, we investigated the spontaneous dissociation of water molecules at these interfaces. Our results reveal a distinct dissociation probability trend: (110) > (100) ≈ (101) > (001), which we attribute primarily to the reaction thermodynamics of surface water dissociation. A strong correlation is observed between the surface Ir–O bond distances and the dissociation probabilities, highlighting the role of surface geometry in modulating reactivity. As a consequence, the interfacial solvation structures and hydrogen bonding environments are dynamically tuned by the varying water dissociation capabilities across facets. This work elucidates how water dissociation energetics depend on surface orientation and interfacial structure, offering atomistic insights into manipulating reaction chemistry at electrocatalytic interfaces.

organic

The seventh blind test of crystal structure prediction: structure ranking methods

A seventh blind test of crystal structure prediction has been organized by the Cambridge Crystallographic Data Centre. The results are presented in two parts, with this second part focusing on methods for ranking crystal structures in order of stability. The exercise involved standardized sets of structures seeded from a range of structure generation methods. Participants from 22 groups applied several periodic DFT-D methods, machine learned potentials, force fields derived from empirical data or quantum chemical calculations, and various combinations of the above. In addition, one non-energy-based scoring function was used. Results showed that periodic DFT-D methods overall agreed with experimental data within expected error margins, while one machine learned model, applying system-specific AIMnet potentials, agreed with experiment in many cases demonstrating promise as an efficient alternative to DFT-based methods. For target XXXII, a consensus was reached across periodic DFT methods, with consistently high predicted energies of experimental forms relative to the global minimum (above 4 kJ mol −1 at both low and ambient temperatures) suggesting a more stable polymorph is likely not yet observed. The calculation of free energies at ambient temperatures offered improvement of predictions only in some cases (for targets XXVII and XXXI). Several avenues for future research have been suggested, highlighting the need for greater efficiency considering the vast amounts of resources utilized in many cases.

Chemistry

Development of an interatomic potential for the Ta–Li system

A new interatomic potential for the Ta–Li system is introduced to facilitate the study of phase stability, mechanical properties and non-equilibrium dynamics after Li implantation in Ta. Here, this potential is based on a generalization of the embedded atom method (GEAM) and includes contributions from embedding energy, explicit two- and three-body interactions, and nonlocal many-body interaction terms. The parameters of the potential are optimized using energies and atomic forces for a wide range of configurations obtained from ab initio density functional theory (DFT) calculations. The potential is rigorously validated across a range of physical properties, including elastic constants, equations of state, phonon dispersion curves, point defect properties, and melting temperatures for different compositions. Although our potential is trained on a small dataset, its accuracy is comparable to that of available machine learning potentials for Li and Ta. Our simulations show that at temperatures below 500 K, Li atoms in Ta–Li alloys form clusters separated by Ta-rich domains, and we find no evidence of ordered phase formation. For Li concentrations below a few percent, Li atoms preferentially segregate to surfaces and grain boundaries. However, in alloys containing more than ~10% Li, the accumulation of Li in symmetric-tilt grain boundaries can lead to one of the following effects: formation of amorphous-like regions, changes in grain boundary structural units, or lateral movement of the grain boundary.

GEAM potential

Hierarchical Reinforcement Learning of a Short-Range Bond-Order Potential for Silica: Analytic Embedding of Coordination with Classical Efficiency

Reinforcement learning (RL) has recently emerged as a data-efficient strategy to parametrize short-range interatomic potentials. Building on our past RL optimization of pairwise silica models, we extend the framework to a bond-order (Tersoff-type) potential that provides an analytic embedding of local coordination through a three-body term. A hierarchical RL workflow combining continuous-action Monte Carlo Tree Search and property-based rewards efficiently explores the 26-dimensional parameter space, sequentially optimizing lattice parameters, densities, angles, and cohesive energies of 21 silica polymorphs. The resulting models, Q-Tersoff and ML-Tersoff, reproduce the energetic ordering of low-energy phases and capture the angular correlations and amorphous structure factors of silica with improved fidelity over pairwise force fields, while remaining orders of magnitude faster than high-dimensional machine-learned potentials. Both models underperform for elastic constants and high-energy frameworks, delineating the limits of the current analytic form. The approach establishes a general and interpretable route to angle-aware, short-range potentials that bridge physics-based and machine-learned descriptions of silicate materials.

36 MATERIALS SCIENCE

Search for active and inactive ion insertion sites in organic crystalline materials

The position of mobile active and inactive ions, specifically ion insertion sites, within organic crystals, significantly affects the properties of organic materials used for energy storage and ionic transport. Identifying the positions of these atomic (and ionic) sites in an organic crystal is challenging, especially when the element has low X-ray scattering power, such as lithium (Li) and hydrogen, which are difficult to detect by powder X-ray diffraction. First-principles calculations, exemplified by density functional theory (DFT), are very practical for confirming the relative stability of ion positions in materials. However, the lack of effective strategies to identify ion sites in these organic crystalline frameworks renders this task extremely challenging. This work presents two algorithms: the (i) efficient location of ion insertion sites from extrema in electrostatic local potential and charge density (ELIISE), and the (ii) ElectRostatic InsertioN (ERIN), which leverage charge density and electrostatic potential fields accessed from first-principles calculations, combined with the Simultaneous Ion Insertion and Evaluation (SIIE) workflow. SIIE inserts all ions simultaneously—to determine ion positions in organic crystals. We demonstrate that these methods accurately reproduce known ion positions in 16 organic materials and identify previously overlooked low-energy sites in tetralithium 2,6-naphthalenedicarboxylate (Li4NDC), an organic electrode material, highlighting the importance of inserting all ions simultaneously, as in the SIIE workflow. These algorithms are also integrated with off-the-shelf machine learning potentials, yielding promising results comparable to first-principles findings.

Gopidi, Harshan Reddy [Argonne National Laboratory

Computing chemical potentials with machine-learning-accelerated simulations to accurately predict thermodynamic properties of molten salts

The successful design and deployment of next-generation nuclear technologies heavily rely on thermodynamic data for relevant molten salt systems. However, the lack of accurate force fields and efficient methods has limited the quality of thermodynamic predictions from atomistic simulations. Here we propose an efficient free energy framework for computing chemical potentials, which is the central free energy quantity behind many thermodynamic properties. We accelerate our simulations without sacrificing accuracy by using machine learning interatomic potentials trained on density functional theory (DFT) data. Using lithium chloride as our model system, we compute chemical potentials with DFT-accuracy for solid and liquid phases by transmuting ions into noninteracting particles. Notably, in the liquid phase, we demonstrate consistency whether we transmute one ion pair or the entire system into ideal gas particles. By locating the temperature where the chemical potential of solid and liquid phases cross, we predict a melting point of 880 ± 18 K for lithium chloride, which is remarkably close to the experimental value of 883 K. With this successful demonstration, we lay the foundation for high-throughput thermodynamic predictions of many properties that can be derived from the chemical potentials of the minority and majority components in molten salts.

Gibson, Luke D. [Oak Ridge National Laboratory (OR

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE

Cartesian equivariant representations for learning and understanding molecular orbitals

Qualitative and quantitative orbital properties such as bonding/antibonding character, localization, and orbital energies are critical to how chemists understand reactivity, catalysis, and excited-state behavior. Despite this, representations of orbitals in deep learning models have been very underdeveloped relative to representations of molecular geometries and Hamiltonians. Here, we apply state-of-the-art equivariant deep learning architectures to the task of assigning global labels to orbitals, namely energies characterizations, given the molecular coefficients from Hartree–Fock or density functional theory. The architecture we have developed, the Cartesian Equivariant Orbital Network (CEONET), shows how molecular orbital coefficients are readily featurized as equivariant node features common to all graph-based machine-learned potentials. We find that CEONET performs well at predicting difficult quantitative labels such as the orbital energy and orbital entropy. Furthermore, we find that the CEONET representation provides an intuitive latent space for differentiating orbital character for the qualitative assignment of e.g. bonding or antibonding character. In addition to providing a useful representation for further integrating deep learning with electronic structure theory, we expect CEONET to be useful for automatizing and interpreting the results of advanced electronic structure methods such as complete active space self-consistent field theory. In particular, the ability of CEONET to infer multireference character via the orbital entropy paves the way toward the machine-learned selection of active spaces.

chemical reactions

Phase transitions and dimensional cross-over in layered confined solids

The nature of solid phases and cross-over of order–disorder phase transitions from two-dimensional (2D) layers to three-dimensional (3D) bulk in confined atomic systems remain largely unexplained. To this end, we consider noble gases and aluminum confined between graphene sheets at different pressures and temperatures. Using crystal structure search methods and molecular dynamics based on machine-learned potentials with quantum-mechanical accuracy, we identify structures of multilayer confined solids that deviate from simple close packing. Upon heating, we find that confined 2D monolayers melt according to the two-step continuous Kosterlitz–Thouless–Halperin–Nelson–Young theory. However, multilayer solids transition continuously into an intermediate layered-hexatic phase before melting discontinuously into an isotropic liquid. This intermediate phase persists at least up to 12 layers studied here. This change can be qualitatively understood based on the cross-over from 2D topological defects toward 3D ones during melting as the number of layers increases.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Machine learning-accelerated path integral molecular dynamics simulations of reactive organic electrolytes

Hydrogen bonded electrolytes that exhibit accelerated proton transport via sequential reactive hops have drawn interest for their promise in clean energy applications. Molecular dynamics simulations of these electrolytes offer the opportunity to uncover microscopic mechanistic details that could be used to design and tune the properties of candidate electrolyte technologies. However, accurately modeling the proton transfer reactions and transport properties that give rise to high charge conductivites in these electrolytes proves computationally challenging because of the need to perform lengthy condensed phase simulations, treating both the electronic and nuclear degrees of freedom quantum mechanically. In this paper, we demonstrate that such a modeling task can be efficiently achieved with the use of density functional theory (DFT)-trained machine learning potentials (MLP) to accelerate path integral molecular dynamics (PIMD) simulations. We highlight the practical utility of this approach by using it to benchmark how closely PIMD simulations employing different DFT exchange–correlation functionals reproduce the composition-dependent densities, diffusion coefficients, and electrical conductivities of mixtures consisting of imidazole and levulinic acid. Even with the speedup afforded by our MLPs, PIMD simulations remain quite expensive. Furthermore, in order to render PIMD more computationally tractable, we introduce and benchmark the accuracy of a ring polymer contraction approach that leverages a computationally efficient short-range MLP to accelerate our PIMD simulations by an additional factor of four.

Chemical bonding

Fragme∩t: An Open‐Source Framework for Multiscale Quantum Chemistry Based on Fragmentation

Fragment-based quantum chemistry offers a means to circumvent the nonlinear computational scaling of conventional electronic structure calculations, by partitioning a large calculation into smaller subsystems then considering the many-body interactions between them. Variants of this approach have been used to parameterize classical force fields and machine learning potentials, applications that benefit from interoperability between quantum chemistry codes. However, there is a dearth of software that provides interoperability yet is purpose-built to handle the combinatorial complexity of fragment-based calculations. To fill this void we introduce “Fragme∩t”, an open-source software application that provides a tool for community validation of fragment-based methods, a platform for developing new approximations, and a framework for analyzing many-body interactions. Fragme∩t includes algorithms for automatic fragment generation and structure modification, and for distance- and energy-based screening of the requisite subsystems. Checkpointing, database management, and parallelization are handled internally and results are archived in a portable database. Interfaces to various quantum chemistry engines are easy to write and exist already for Q-Chem, PySCF, xTB, Orca, CP2K, MRCC, Psi4, NWChem, GAMESS, and MOPAC. Applications reported here demonstrate parallel efficiencies around 96% on more than 1000 processors but also showcase that the code can handle large-scale protein fragmentation using only workstation hardware, all with a codebase that is designed to be usable by non-experts. Fragme∩t conforms to modern software engineering best practices and is built upon well established technologies including Python, SQLite, and Ray. The source code is available under the Apache 2.0 license.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Porosity in nuclear graphite and its impact on nuclear reactor science and criticality safety applications

Porosity in nuclear-grade graphite significantly influences its low-energy neutron scattering, yet its effect on underlying phonon properties remains debated. Here, this work integrates inelastic and small-angle neutron scattering (INS/SANS) experiments, advanced atomistic simulations with a novel machine-learned potential (DeepMD), total cross-section measurements, and neutronics calculations (SCALE, MCNP, OpenMC) to investigate porosity’s impact on neutron thermalization. INS measurements on diverse graphite grades reveal no discernible porosity effect on phonon spectra, which align with crystalline graphite. Conversely, total cross-section data below ≈10 meV show increased scattering attributable to SANS. Our DeepMD simulations demonstrate that realistic micropores do not distort phonon spectra, challenging the assumptions in current ENDF/B-VIII.1 porosity thermal scattering laws (TSLs). These TSLs, based on random atom removal, produce unphysical phonon spectra and inflate inelastic cross-sections. Augmenting a crystalline TSL with an SANS component accurately captures experimental total cross-sections. Neutronics benchmarks (ICSBEP/IRPhE) show ENDF porosity TSLs unphysically increase neutron multiplication factor, keff. Crucially, incorporating SANS physics (NCrystal/OpenMC) indicates accurately modeled porosity negligibly affects keff, reactor physics, or criticality safety.

Critical benchmarks