Search NASASearch

SEARCH · Search NASA

Results for “Model Uncertainty”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

963 records · Page 17

Advanced Model Development for Large Eddy Simulation of Oxy-Combustion and Supercritical Carbon Dioxide Power Cycles

A joint experimental and numerical study is performed to observe the characteristics of a supercritical carbon dioxide turbulent mixing layer in the presence of strong nonlinearities in the thermodynamic and transport properties. A bespoke experimental setup is designed and employed for this purpose and provides insight into macroscopic mixing behavior. The mixing is experimentally observed using two techniques: shadowgraphy and spontaneous Raman scattering. Qualitative and quantitative intensity fields obtained via these techniques yield instantaneous and mean density data. Spanwise temperature data is also collected using analogue resistance temperature detectors. These measurements are used to quantify the level of mixed material within the field. The experimental data are supplemented by a companion high-fidelity numerical study. The numerical results are obtained through fully resolved, three-dimensional direct numerical simulation. The numerical dataset permits observation of the near-field mixing characteristics, which are difficult to measure experimentally due to the rapid dynamics and sharp thermophysical gradients in this area. Qualitative field visualizations are presented, followed by quantitative mixed material results and observations regarding thermodynamic property trends at select locations within the field. One-dimensional spectra of the turbulent kinetic energy and solenoidal dissipation are provided to observe the spectral characteristics of the flow. Reynolds stress anisotropy is analyzed graphically through anisotropy invariance maps (Lumley triangles). The mixing quantification, spectral data and anisotropy analysis of a flow at these thermodynamic conditions represent the main outcomes of the work.

20 FOSSIL-FUELED POWER PLANTS

Elastic-plastic models for multi-site damage

This paper presents recent developments in advanced analysis methods for the computation of stress site damage. The method of solution is based on the p-version of the finite element method. Its implementation was designed to permit extraction of linear stress intensity factors using a superconvergent extraction method (known as the contour integral method) and evaluation of the J-integral following an elastic-plastic analysis. Coarse meshes are adequate for obtaining accurate results supported by p-convergence data. The elastic-plastic analysis is based on the deformation theory of plasticity and the von Mises yield criterion. The model problem consists of an aluminum plate with six equally spaced holes and a crack emanating from each hole. The cracks are of different sizes. The panel is subjected to a remote tensile load. Experimental results are available for the panel. The plasticity analysis provided the same limit load as the experimentally determined load. The results of elastic-plastic analysis were compared with the results of linear elastic analysis in an effort to evaluate how plastic zone sizes influence the crack growth rates. The onset of net-section yielding was determined also. The results show that crack growth rate is accelerated by the presence of adjacent damage, and the critical crack size is shorter when the effects of plasticity are taken into consideration. This work also addresses the effects of alternative stress-strain laws: The elastic-ideally-plastic material model is compared against the Ramberg-Osgood model.

Actis, Ricardo L.

Mathematical Model of a Regenerative Fuel Cell for System Optimization

This thesis developed a system-level optimization model of a regenerative fuel cell (RFC) system for long-duration, off-world energy storage applications. Prior RFC design studies have typically been limited to reduced parameter sets and simplified constraints due to computational limitations relative to the number of relevant degrees of freedom. As a result, important nonlinear interactions between subsystems have not been fully captured. This work began to address that gap by developing a higher-fidelity, nonlinear optimization framework that incorporates a broader set of design variables and coupled constraints, enabling a multidimensional model that captures the coupled behavior of RFC subsystems and demonstrates the feasibility of applying optimization to such systems. An expanded system-level optimization approach was established that captures interactions between electrochemical performance, structural requirements, and storage design. This enabled a more comprehensive evaluation of trade-offs than conventional formulations. The model integrates four coupled subsystems: a fuel cell, an electrolyzer, reactant gas, and high-pressure storage tanks, and was formulated to accommodate a wide range of mission parameters, including operational time and required output power. It incorporates constraints on available solar array power, reactant mass balance between production and consumption, and pressure-dependent storage requirements. To enable reliable convergence, the optimization problem was reformulated to reduce dimensionality and improve numerical stability, with subsystem models organized for efficient evaluation. Problem dimensionality was reduced by consolidating lower-level design variables into higher-level representative quantities, and subsystem behavior was evaluated within the optimization loop. A multi-start initialization strategy was employed to mitigate sensitivity to local minima and improve solution quality, while nonlinear relationships were solved using robust numerical methods. The results showed that convergence was achieved across a range of required output power values. Specific energy reached a maximum at a critical mission power level, where the electrolyzer power matched the available solar input and operated near its voltage and current density limits. Beyond this point, further increases in required power resulted in less mass-efficient operation, increasing total system mass and reducing overall performance. The developed model represents an advancement in RFC system-level optimization by enabling analysis of a broader and more tightly coupled design space than previous considerations. While convergence behavior and computational cost remain challenges, the methods introduced improve solvability and allow inclusion of additional design variables with minimal loss of physical fidelity. However, the numerical results should not be interpreted as definitive design recommendations, as the model includes simplifying assumptions and omits several higher-order effects. Future work should extend this framework by incorporating additional subsystems and loss mechanisms, such as thermal management, parasitic power consumption, and reactant losses, to improve fidelity and ensure more representative design conclusions.

Electrochemistry

Characterizing the subglacial environment of lower Thwaites Glacier using radar modeling

Understanding the spatial heterogeneity beneath Thwaites Glacier, West Antarctica, is vital to projecting its impact on future sea levels. Radar-echo sounding (RES) is commonly used to infer subglacial conditions, but these data can be challenging to interpret. We assess basal heterogeneity across Thwaites Glacier by comparing RES returns to a radar backscattering simulator for over 400 km of RES data. The modeled variations in bed returned power exhibited a strong correlation with actual RES data in 40% of our simulated flight segments, which we consider evidence for a relatively homogeneous glacier bed. Other sites (40%) demonstrated improved fit quality when hydrology or substrate transitions were introduced in the bed material model. The remaining simulated segments (20%) were diagnosed as having more complex basal heterogeneity. The spatial distribution of complex heterogeneity appears to coincide with asymmetric patterns in the RES specularity content, which has been interpreted in previous studies as a signature for channelized hydrology. Conversely, the homogeneous substrate locations coincide with areas of fast-moving ice in western Thwaites. Our simulation method can isolate power variations induced by material heterogeneity vs topography, which is an important limitation of existing RES analysis methods.

Antarctic glaciology

A SIMULINK Environment for Flight Dynamics and Control Analysis - Application to the DHC-2 'Beaver', Part 1. Implementation of a Model Library in SIMULINK

The design of advanced Automatic Aircraft Control Systems (AACS's) can be improved upon considerably if the designer can access all models and tools required for control system design and analysis through a graphical user-interface, from within one software environment. This MSc-thesis presents the first step in the development of such an environment, which is currently being done at the Section for Stability and Control of Delft University of Technology, Faculty of Aerospace Engineering. The environment is implemented within the commercially available software package MATLAB/SIMULINK. The report consists of two parts. Part I gives a detailed description of the AACS design environment. The heart of this environment is formed by the SIMULINK implementation of a nonlinear aircraft model in block-diagram format. The model has been worked out for the old laboratory aircraft of the Faculty, the De Havilland DHC-2 'Beaver', but due to its modular structure, it can easily be adapted for other aircraft. Part I also describes MATLAB programs which can be applied for finding steady-state trimmed-flight conditions and for linearization of the aircraft model, and it shows how the built-in simulation routines of SIMULINK have been used for open-loop analysis of the aircraft dynamics. Apart from the implementation of the models and tools, a thorough treatment of the theoretical backgrounds is presented. Part II of this report presents a part of an autopilot design process for the 'Beaver' aircraft, which clearly demonstrates the power and flexibility of the AACS design environment from part I. Evaluations of all longitudinal and lateral control laws by means of nonlinear simulations are treated in detail. The AACS design environment from part I proved to be a very useful tool for designing the control laws of the 'Beaver' autopilot within a very tight time-schedule. The autopilot design process itself will be used as a guideline for future AACS research at the Faculty of Aerospace Engineering. Flight tests of the 'Beaver' autopilot, done after evaluating the control laws in the SIMULINK package, proved to be quite successful. In the future, the AACS design package will evolve into a standardized, integrated design environment which can be applied to virtually any type of aircraft. The AACS design cycle will be shortened further by developing tools for automatically porting control laws from the MATLAB/SIMULINK environment to a piloted real-time flight simulator and the Flight Control Computers of the aircraft.

Aircraft Modelling

Computing the Critical Temperature of the Affine-Transformed $D=3$ Ising Model Using Masked Autoregressive Flow

The simple Ising model provides a rich environment to build and study lattice field theories. As part of an ongoing project to construct a conformal field theory (CFT) on an arbitrarily curved manifold, in this work we develop methods to measure the critical temperature $β_c$ of the affine-transformed Ising model on the face-centered cubic (FCC) lattice. The main challenge in this endeavor is finding a computationally efficient and accurate method of interpolating and extrapolating Monte Carlo observables with respect to coupling coefficients and temperature. Herein, we compare two such methods. A traditional statistical approach uses the multiple histogram (MH) method, while a newer machine learning approach uses a masked autoregressive flow (MAF) to estimate the underlying probability density function of a set of observables. While the MH method is specifically designed to interpolate and extrapolate Monte Carlo observables, we find that MAF is a viable alternative for measuring $β_c$ with a computational cost that scales more favorably. Furthermore, we comment on additional advantages of MAF relevant to our work, such as extrapolating in system volume.

Svenson, Kai [Texas U.]

A Transient Hydrodynamic Model of Screen Channel Liquid Acquisition Devices for In-Space Cryogenic Propellant Management

Screen channel liquid acquisition devices (LADs) will play a crucial role in future deep space travel. It is essential that vapor-free delivery of propellants during tank-to-tank transfer is ensured to maximize yield from storage tanks and prevent potential combustion instabilities. The screen channel LAD utilizes a fine screen wire mesh that can separate phases in a low Bond number (i.e. microgravity) environment using surface tension forces. This study presents the development and verification of a new model for transient screen compliance, one of the influential factors for screen channel LAD design. Screen compliance is crucial during LAD channel outflow transients because the slight deflection of the screen can provide needed mass to satisfy rapid outflow demands and reduce the pressure difference across the screen. The model is successfully verified against CFD simulations. In addition, the characteristic speed for the governing screen compliance equations is derived which allows for numerical stability criteria to be established. As shown in this study, the transient maximum pressure difference across the screen can greatly exceed the steady state maximum pressure difference across the screen in many cases.

Hydrodynamics Simulations

Earth Observations and Integrative Models in Support of Food and Water Security

Global food production depends upon many factors that Earth observing satellites routinely measure about water, energy, weather, and ecosystems. Increasingly sophisticated, publicly-available satellite data products can improve efficiencies in resource management and provide earlier indication of environmental disruption. Satellite remote sensing provides a consistent, long-term record that can be used effectively to detect large-scale features over time, such as a developing drought. Accuracy and capabilities have increased along with the range of Earth observations and derived products that can support food security decisions with actionable information. This paper highlights major capabilities facilitated by satellite observations and physical models that have been developed and validated using remotely-sensed observations. Although we primarily focus on variables relevant to agriculture, we also include a brief description of the growing use of Earth observations in support of aquaculture and fisheries.

Water Resources

A quality-agnostic combinatoric cost estimation model for large-format directed energy deposition metal additive manufacturing

Directed energy deposition (DED) additive manufacturing (AM) processes are amenable to synergistic combination into multi-process AM systems due to similar requirements for automation and energy sources. This work analyzes the economic performance of such DED AM systems from a quality-agnostic combinatoric standpoint with a model that calculates lowest-cost system combinations based on part geometry and process performance metrics. Common DED AM systems research focuses on a single process and does not consider the process, system, and application in the context of all possible system combinations (e.g., the combined set of process selection(s), motion system(s), and process hardware), leading to limited applicability of the resulting DED AM systems to cost-sensitive components such as those found in energy generation applications. The model developed herein incorporates the capital, material, and energy costs associated with DED AM system combinations into a predictive tool for estimating part and system cost, the output of which is intended to guide deployment of finite research and development resources towards DED AM system combinations with the lowest costs and greatest likelihood of economic impact. The DED AM systems identified by this framework may enable domestic production of the large conventionally cast and forged components necessary for energy generation.

Shanafield, Alexandra [ORNL]

Phase-field modeling of stored-energy-driven grain growth with intra-granular variation in dislocation density

Abstract We present a phase-field (PF) model to simulate the microstructure evolution occurring in polycrystalline materials with a variation in the intra-granular dislocation density. The model accounts for two mechanisms that lead to the grain boundary migration: the driving force due to capillarity and that due to the stored energy arising from a spatially varying dislocation density. In addition to the order parameters that distinguish regions occupied by different grains, we introduce dislocation density fields that describe spatial variation of the dislocation density. We assume that the dislocation density decays as a function of the distance the grain boundary has migrated. To demonstrate and parameterize the model, we simulate microstructure evolution in two dimensions, for which the initial microstructure is based on real-time experimental data. Additionally, we applied the model to study the effect of a cyclic heat treatment (CHT) on the microstructure evolution. Specifically, we simulated stored-energy-driven grain growth during three thermal cycles, as well as grain growth without stored energy that serves as a baseline for comparison. We showed that the microstructure evolution proceeded much faster when the stored energy was considered. A non-self-similar evolution was observed in this case, while a nearly self-similar evolution was found when the microstructure evolution is driven solely by capillarity. These results suggest a possible mechanism for the initiation of abnormal grain growth during CHT. Finally, we demonstrate an integrated experimental-computational workflow that utilizes the experimental measurements to inform the PF model and its parameterization, which provides a foundation for the development of future simulation tools capable of quantitative prediction of microstructure evolution during non-isothermal heat treatment.

Materials Science

Impact of Crystalline Phases on Low-Activity Waste Glass Durability: Insights from PCT and VHT

During vitrification of nuclear wastes, slow cooling along the container centerline promotes crystalline phase formation, which can alter residual glass composition and reduce chemical durability. This study investigates the effects of crystalline phases on the chemical durability of low-activity waste (LAW) borosilicate glasses using the product consistency test (PCT) and vapor hydration test (VHT) on container centerline cooled (CCC) samples. A preliminary model (R2 = 0.88) was developed to predict CCC PCT responses based on glass composition, PCT data from quenched glasses, and measured crystal fractions. Using the latest LAW glass dataset, the feasibility of predictive modeling is evaluated, limitations in current data and methods are identified, and challenges for improving model accuracy are discussed to guide future data collection and model development.

borosilicate glass

Energy-Optimal Vehicle Longitudinal Motion Control via Pontryagin’s Minimum Principle and Ultra-Local Model

Longitudinal vehicle motion control is essential for enhancing performance and optimizing a vehicle’s energy usage. However, it remains a challenging task due to the nonlinear and uncertain nature of vehicle dynamics, along with varying driving conditions. This paper presents a novel ultra-local optimal control approach based on Pontryagin’s Minimum Principle (PMP) that circumvents the need for detailed system identification by employing an ultra-local model. The control objective is to minimize the total energy consumption under boundary conditions while ensuring smooth traction force generation. The proposed approach is evaluated using a high-fidelity vehicle model in three representative scenarios: (i) nominal driving, (ii) a change in tire road friction coefficient (TRFC) from 0.5 to 0.65 and road slope from 0% to 5% during the maneuver, with target velocity unchanged, and (iii) a change in target velocity from 20 m/s to 0 m/s during the maneuver, while maintaining nominal TRFC and slope conditions. The simulation results demonstrate that the proposed method delivers robust performance, effectively balancing consumption and tracking accuracy in all tested scenarios.

Waleed khan, Muhammad [The University of Texas at

Dark energy survey year 3 results: likelihood-free, simulation-based w CDM inference with neural compression of weak-lensing map statistics

We present simulation-based cosmological wcold dark matter (wCDM) inference using dark energy survey year 3 weak-lensing maps, via neural data compression of weak-lensing map summary statistics: power spectra, peak counts, and direct map-level compression/inference with convolutional neural networks (CNN). Using simulation-based inference, also known as likelihood-free or implicit inference, we use forward-modelled mock data to estimate posterior probability distributions of unknown parameters. This approach allows all statistical assumptions and uncertainties to be propagated through the forward-modelled mock data; these include sky masks, non-Gaussian shape noise, shape measurement bias, source galaxy clustering, photometric redshift uncertainty, intrinsic galaxy alignments, non-Gaussian density fields, neutrinos, and non-linear summary statistics. We include a series of tests to validate our inference results. This paper also describes the Gower Street simulation suite: 791 full-sky pkdgrav3 dark matter simulations, with cosmological model parameters sampled with a mixed active-learning strategy, from which we construct over 3000 mock dark energy survey lensing data sets. For wCDM inference, for which we allow –1 < w < –$\frac{1}{3}$⁠, our most constraining result uses power spectra combined with map-level (CNN) inference. Using gravitational lensing data only, this map-level combination gives Ω m = 0.283$^{+0.020}_{–0.027}$⁠, S 8 = 0.804$^{+0.025}_{–0.017⁠}$, and w < –0.80 (with a 68 per cent credible interval); compared to the power spectrum inference, this is more than a factor of two improvement in dark energy parameter (Ω⁠ DE , w⁠) precision.

79 ASTRONOMY AND ASTROPHYSICS

Developing a Pyrolysis Gas Thermal Blocking Model for Reentry Demise

In NASA’s Object Reentry Survival Analysis Tool (ORSAT), aerodynamic drag and aerothermal heating coefficients are computed for each of the free-molecular, continuum, and transitional flow regimes using analytical and semi-analytical methods. These heating coefficients were derived for typical metallic materials that melt and do not have a strong gas-phase contribution to the flow in the boundary layer. Modern satellites typically feature fiber-reinforced polymer (FRP) components, such as solar array booms, facesheets of sandwich panels, or overwraps for composite-overwrapped pressure vessels (COPV). These FRP materials do not behave the same as metals in the reentry environment, but instead will pyrolyze and develop significant volumes of gas into the boundary layer. Accurately predicting the reentry demise of FRP components is critical to assessing the reentry casualty risk for modern spacecraft. Research in recent years has shown that this demisability can depend heavily on how the expulsion of gaseous pyrolysis products through the outer surface of the material affects the heat flux at the surface. The ODPO has been developing a reduced-order model of the effect of pyrolysis gas blowing on the heat flux based on correlations between a blowing factor and a non-dimensional heat flux to be incorporated in the upcoming version 7.3 of the Object Reentry Survivability Analysis Tool (ORSAT). This presentation discusses the progress of this development project and the challenges remaining for generalizing the model across families of FRP materials.

Benton Greene

Prime Time for Model-Predictive Control? Assessing the Technical and Market Readiness of Advanced Controls in Buildings

Despite three decades of extensive research and field testing that have consistently validated the benefits of Model Predictive Control (MPC) in building applications, the technology has seen limited market adoption. This paper evaluates the readiness of MPC for widespread deployment, showcases recent demonstrations and field tests across diverse building types, including residential, small commercial, large commercial, and campus settings. Our results demonstrate that MPC can optimize system operations to achieve load shifting, minimize curtailment of on-site generation, and reduce energy costs by up to 80 %, while maintaining or improving occupant comfort. We also show that MPC can effectively control large assets, such as MW-sized thermal storage systems, and respond to dynamic pricing signals. However, achieving scale remains difficult due to labor-intensive workflows, reliance on a “PhD-in-the-loop” for MPC design and maintenance, susceptibility to fragile data infrastructure, and persistent workforce education and acceptance barriers. To bridge this gap, we outline a transition from bespoke, labor intensive prototypes toward streamlined, segment-targeted deployment strategies that leverage model templates, semantic tools, and generative AI. By automating control configuration and reducing engineering effort, these recommendations provide a pathway for transforming successful research demonstrations into scalable, market ready solutions for MPC-based controls.

Pritoni, Marco

CFD Modeling of Bi-Directional PMD inside Cryogenic Propellant Tanks Onboard Parabolic Flights

Future cryogenic propulsion systems will require efficient methods with which to transfer cryogenic propellants from a depot storage tank to a customer receiver tank to minimize cost and maximize reusability. The Reduced Gravity Cryogenic Transfer project is currently developing advanced cryogenic fluid management technology and developing and validating new numerical models for three phases of transfer: line chilldown, tank chilldown, and tank fill. Additionally, multiple liquid nitrogen (LN 2 ) parabolic flight transfer rigs are being designed by universities and NASA to investigate the gravitational sensitivities that exist in these three technologies. In order to maximize the collection of low-g data during flights, it is required to extract as much (LN 2 as possible from the supply tank, despite variable gravity levels. The purpose of this paper is to present computational fluid dynamics (CFD) volume of fluid simulations of (LN 2 behavior in the supply tank onboard parabolic flights to validate the optimal design of a bi-directional propellant management device (PMD) using the commercial software FLOW-3D. A parametric study is conducted on the effects of gravity level, fill level, pore size, open area, thickness, and type of baffle on PMD performance. Based on results, the PMD as designed exceeds the targeted expulsion efficiency.

Jason Hartwig

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE

Model-based, in-situ, non-destructive qualification and certification of parts made by autonomous additive manufacturing

To address the significant productivity challenges associated with the qualification and certification (Q&C) tasks of additively manufactured (AM) parts, which have traditionally relied on rigorous post‐build inspection and testing, we propose an integrated framework that combines model‐based qualification and certification (MBQ&C) with autonomous additive manufacturing (AAM). MBQ&C employs high‐fidelity predictive models, developed within the Integrated Computational Materials Engineering (ICME) paradigm, to simulate process–structure–property–performance relationships for assessing a part’s fitness for use. Since predictive models are commonly machine learning (ML)-based or reduced-order surrogates of validated physics models, they run efficiently, enabling timely inference. In parallel, the self-driving AAM utilises ML-based adaptive, closed‐loop control strategies to avoid, mitigate, or repair defects and anomalies during fabrication, thereby increasing the likelihood of producing acceptable parts. A key feature of the combined AAM-MBQ&C framework is that predictive models explicitly incorporate defects or anomalies that persist after the build, using instance-specific data captured via in-situ sensing. This customisation enables a build‐specific assessment of fitness for use, rather than relying on nominal or generic parameters. Such individualised evaluation provides a robust basis for Q&C-related acceptance decisions relating to each build. Additionally, the rapid solution capabilities of ML or reduced-order models enable the determination of a part’s suitability for service shortly after build completion. As the framework matures, it has the potential to substantially reduce reliance on conventional point‐design approaches—such as time‐consuming post‐build computed tomography scanning and costly destructive testing. Thus, the AAM-MBQ&C framework represents a transformative, scalable strategy for quality assurance of AM components, as parts produced within a stable, validated, and certified envelope can be certified with reduced testing. Key benefits include: (1) significant gains in Q&C productivity through efficient, model-centric assessment; (2) performance-based classification of defects into critical and non-critical categories; (3) the ability to predict potential deviations in the performance of parts affected by real-time, adaptive process control interventions relative to those produced under a certified process, and (4) the enabling of virtual Q&C for service environments that are difficult, hazardous, or impractical to access or reproduce experimentally. Collectively, these capabilities strengthen the business case for AM, particularly for high‐consequence and mission‐critical applications. Finally, although this work focuses on powder-based AM, the proposed techniques could be extended to AM processes employing alternative feedstock forms.

Gunasegaram, Dayalan