Search NASASearch

SEARCH · Search NASA

Results for “Architectural nanomaterials”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

92 records · Page 5

Used Nuclear Fuel Management Using the Next Generation System Analysis Model

The U.S. Department of Energy (DOE) is leading the National effort to manage the back end of the nuclear fuel cycle, encompassing the safe transportation, storage/staging, and/or eventual disposal of used nuclear fuel (UNF) and high-level radioactive waste. The Next Generation System Analysis Model (NGSAM) is DOE’s discrete-event, agent-based simulation tool designed to model the full life cycle of UNF from reactor discharge to final disposal. NGSAM supports the DOE Office of Spent Fuel and High-Level Waste Disposition by enabling a detailed, scenario-based analysis of logistics, infrastructure, and shipping strategies. NGSAM replaces legacy models with a modern, flexible platform built on Repast Simphony and enhanced by the Process Analysis Tool. NGSAM simulates the movement and interaction of individual fuel assemblies with system components such as canisters, casks, railcars, and facilities. The model integrates with the Java Transportation Operations Model to plan and execute transportation scenarios, supporting both constrained and unconstrained resource allocation. Key features include customizable allocation and acceptance algorithms, detailed facility-level operations, and a Quick Edit tool for rapid scenario adjustments. NGSAM supports multimodal transportation modeling (e.g. rail, road, barge) and provides comprehensive cost, schedule, and infrastructure data. NGSAM utilizes data from sources such as DOE’s STANDARDS UNF database and DOE’s Stakeholder Tool for Assessing Radioactive Transportation, while also allowing user-defined inputs for scenario customization. NGSAM enables stakeholders to evaluate complex UNF management strategies, assess system performance under varying assumptions, and inform decision making for future infrastructure investments. Its modular architecture and integration with other Integrated Waste Management System tools make it a critical asset for planning the safe and efficient disposition of the Nation’s growing UNF inventory.

Craig, Brian [Argonne National Laboratory (ANL)]

Maximizing dynamic range and performance of anatase TiO 2 ECRAM through structure and programming

Here, in this study, we investigate the structure-dependent modulation characteristics of all-solid-state three-terminal electrochemical random-access memory (ECRAM) based on an anatase Li x TiO 2 channel. By directly comparing “asymmetric” and “symmetric” ECRAM device architectures, we reveal significant insight into the impact of a non-zero gate-drain open-circuit voltage and its influence on voltage vs. current-controlled gating. We also explore the impact of potentiation/depression write parameters on the symmetry, linearity, and dynamic range of the device response. Together, initial results from optimizing structure and programming approaches yielded unprecedented G max /G min ratios of >1,000 for ECRAM and hundreds of tunable memory states with excellent linearity and symmetry. Simulations based on these ECRAM devices further illustrate the promise of this analog memory technology, achieving near 2% classification error in the MNIST digit recognition benchmark for a range of training parameters compared to a theoretical best of 1.66% and outperforming other device models extracted from the literature.

AIHWKit

CatTestHub: A benchmarking database of experimental heterogeneous catalysis for evaluating advanced materials

The ability to quantitatively compare newly evolving catalytic materials and technologies is hindered by the widespread availability of catalytic data collected in a consistent manner. While certain catalytic chemistries have been widely studied across decades of scientific research, quantitative comparisons based on literature information is hindered by variability in reaction conditions, types of reported data, and reporting procedures. Here, we present CatTestHub, an open-access database dedicated to benchmarking experimental heterogeneous catalysis data. Combining systematically reported catalytic activity data for selected probe chemistries, with relevant material characterization and reactor configuration information, the database provides a collection of catalytic benchmarks for distinct classes of active site functionality. Through key choices in data access, availability, and traceability, CatTestHub seeks to balance the fundamental information needs of chemical catalysis and the FAIR data design principles. Details of the database architecture and the means through which to navigate it are presented, highlighting examples of catalytic insights readily drawn from the available benchmarking data. In its current iteration, CatTestHub spans over 250 unique experimental data points, collected over 24 solid catalysts, that facilitated the turnover of 3 distinct catalytic chemistries. Here, a roadmap is presented through which to expand the open-access platform that serves as a community wide benchmark, primarily through continuous addition of kinetic information on select catalytic systems by members of the heterogeneous catalysis community at large.

Benchmark

High Sulfur Loading and Capacity Retention in Bilayer Garnet Sulfurized‐Polyacrylonitrile/Lithium‐Metal Batteries with Gel Polymer Electrolytes

The cubic‐garnet (Li 7 La 3 Zr 2 O 12 , LLZO) lithium–sulfur battery shows great promise in the pursuit of achieving high energy densities. The sulfur used in the cathodes is abundant, inexpensive, and possesses high specific capacity. In addition, LLZO displays excellent chemical stability with Li metal; however, the instabilities in the sulfur cathode/LLZO interface can lead to performance degradation that limits the development of these batteries. Therefore, it is critical to resolve these interfacial challenges to achieve stable cycling. Here, an innovative gel polymer buffer layer to stabilize the sulfur cathode/LLZO interface is created. Employing a thin bilayer LLZO (dense/porous) architecture as a solid electrolyte and significantly high sulfur loading of 5.2 mg cm −2 , stable cycling is achieved with a high initial discharge capacity of 1542 mAh g −1 (discharge current density of 0.87 mA cm −2 ) and an average discharge capacity of 1218 mAh g −1 (discharge current density of 1.74 mA cm −2 ) with 80% capacity retention over 265 cycles, at room temperature (22 °C) and without applied pressure. Achieving such stability with high sulfur loading is a major step in the development of potentially commercial garnet lithium–sulfur batteries.

25 ENERGY STORAGE

Maintenance strategy, structural design, and site layout of the ST-E1 fusion power plant

An effective fusion reactor maintenance scheme enables safe operations and short downtimes. This in turn leads to high availability, which is critical to the commercial viability of a power-producing plant. In tokamak-based fusion power plants, the chosen maintenance approach has a significant impact on the spatial design of the tokamak, as well as the surrounding infrastructure, and therefore needs to be considered from the outset. Tokamak Energy has developed a pre-concept design of a fusion power plant, ST-E1. This work describes the major drivers and constraints that have been considered, presents the tokamak architecture and chosen maintenance regime, and discusses how this enables the plant’s two-phased approach to demonstrating commercial operations. It also shows the implications for the design of other systems areas, in particular the machine structural arrangement and bioshield and hot cell layout. The reactor core segmentation and removal scheme replaces entire toroidal segments radially through a large vacuum port, along a single axis only. The result is a change-tolerant machine and plant layout that can accommodate the evolving designs of the tokamak.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Preliminary Plan to Inform Testing of a Heat Exchanger Test Article

This report presents a preliminary plan to guide the qualification testing of advanced heat exchanger (HX) components for nuclear-to-industrial heat transfer applications. The objective is to establish a defensible, physics-based methodology that integrates computational modeling, targeted experimentation, and in-service inspection considerations to demonstrate component performance and reliability under representative reactor conditions. The analysis identifies Sodium-cooled Fast Reactor (SFR) and High-Temperature Gas-cooled Reactor (HTGR) systems as reference configurations in terms of temperature, pressure, and chemical environment. Within these operating envelopes, dominant degradation mechanisms— including creep–fatigue interaction, flow-induced vibration, corrosion, and diffusion-bond deterioration—were evaluated to define test requirements. A comprehensive computationalexperimental framework is proposed to support life prediction and qualification activities. The framework couples high-fidelity structural-mechanics, thermal-hydraulic, and fluid-structure interaction models with accelerated degradation testing to produce a traceable linkage between microstructural evolution, mechanical performance, and remaining useful life (RUL). The approach adheres to established Verification, Validation, and Uncertainty Quantification (VVUQ) standards (ASME V&V 10/20; NUREG-2152) and incorporates a digital-twin architecture for continuous model refinement through data assimilation. The plan further outlines testing methodologies, including pre-test analyses, test-loop design parameters, and sensor placement strategies that maximize information yield while maintaining mechanistic fidelity. Complementary sections describe in-service inspection (ISI), on-line monitoring (OLM), and structural-health-monitoring (SHM) techniques applicable to compact HX geometries typical of advanced reactors. Collectively, these activities establish the technical foundation for demonstrating 40-60-year equivalent service life of advanced heat exchangers in support of the U.S. Department of Energy’s Advanced Reactor and Integrated Energy Systems programs. The forthcoming phase will execute the defined pre-test analyses, initiate hardware fabrication, and implement the integrated testing campaign to validate the proposed qualification methodology.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Solid neon as a noise-resilient host for electron qubits above 100 mK

Solid neon can be used as a solid host for single-electron qubits. At temperatures of around 10 mK, electron-on-solid-neon charge qubits exhibit long coherence times and high operation fidelities. However, a systematic characterization of the noise features of such systems is needed for the development of scalable quantum information architectures. Here, in this work, we show that solid neon can be used as a noise-resilient host for electron qubits above 100 mK. We examine the resilience of solid neon against charge and thermal noise when electron-on-solid-neon charge qubits are operated away from the charge-insensitive sweet spot and at elevated temperatures. We show that the extracted high-frequency charge noise density of electron-on-solid-neon qubits, projected as voltage fluctuations on nearby electrodes, is between 10 −4 μV 2 Hz −1 and 10 −6 μV 2 Hz −1 at 0.01 MHz to 1 MHz, which is comparable to common semiconductor hosts. We also show that the electron-on-solid-neon charge qubits operating at frequencies of around 5 GHz can maintain echo coherence times of over 1 μs at temperatures up to 400 mK.

42 ENGINEERING

Solvent-Mediated Control of Nanocellulose Dispersion: An Integrated Computational and Experimental Investigation

Fibrillated cellulose derived from forestry feedstocks represents a renewable and high-strength materials platform for circular bioeconomies. However, its practical implementation is hindered by the irreversible aggregation of nanocellulose architectures, including cellulose nanofibers (CNFs). Solvent-based dispersion offers a simple and practical route to prevent CNF aggregation. Here, in this work, we integrate classical and enhanced sampling molecular dynamics (MD) simulations with experimental suspension rheology and atomic force microscopy (AFM) to elucidate how solvent environments tune CNF–CNF interactions and dispersion stability. CNF–CNF contact free energies computed from MD simulations reveal reduced aggregation in acetone/water, γ-valerolactone (GVL)/water, and tetrahydrofuran (THF)/water and pure acetone compared with pure water, reflecting stronger CNF-solvent relative to inter-CNF interactions. Correspondingly, CNF-solvent suspensions in these solvent systems exhibit stronger inter-fibril network structures and enhanced recovery compared to water, indicating improved CNF-solvent affinity. Liquid cell AFM imaging in acetone–water mixtures and in pure acetone further confirm the presence of well-dispersed CNFs. By combining multiscale computation with targeted experiments, this study establishes a rational framework for solvent design to achieve stable nanocellulose dispersions for high-strength biobased materials and efficient bioenergy conversion.

cellulose

SENTRA: A Modular Computational Graph Framework for Critical Mineral and Materials Supply Chains: Part I: Network Construction Latent-Quantity Estimation, and Temporal Graph Forecasting

Global supply chains for critical minerals and materials are complex, evolving networks of countries, products, production stages, and trade relationships. Existing analytical approaches are limited by fragmented data and static network representations that do not capture the dynamic production dependencies linking raw materials, intermediate products, and final goods across multiple countries. Trade and production statistics provide only a partial view of domestic production, inventories, and material flows, making it difficult to identify indirect sourcing pathways, hidden dependencies, and embedded foreign exposures. This paper introduces the Supply Chain Exposure Network Tracking and Risk Assessment (SENTRA) framework, a modular graph-based computational framework for constructing, analyzing, and forecasting dynamic supply chain networks. As the first paper in a three-part methodological series, it establishes the computational foundation of SENTRA by constructing a temporal attributed multi-relational graph whose nodes represent product–country pairs and whose edges encode observed trade and within-country value-chain relationships. Statistical estimation and constrained optimization recover latent production, final demand, and product input dependency coefficients while enforcing economic accounting constraints. Graph-derived exposure measures quantify direct, transshipment, value-chain, and multi-hop supply chain dependencies independently of the forecasting model. A temporal graph forecasting architecture based on a relational graph neural network then forecasts the evolution of the graph under mass-balance constraints with distribution-free conformal uncertainty quantification. Validation on the global aluminum supply chain shows that the learned graph representations recover economically meaningful supply chain structure, accurately forecast out-of-sample trade relationships, and produce well-calibrated prediction intervals. Subsequent papers apply this computational foundation to exposure assessment, disruption analysis, and scenario-based policy analysis, and extend the framework to multimaterial supply chain modeling and decision support.

36 MATERIALS SCIENCE

Paleotribological models using preserved fossil tissue properties reveal functional significance of Eurasian mammoth dental evolution

Eurasian mammoths (Mammuthus) underwent substantial modifications in molar morphology as later-diverging species evolved progressively thinner enamel and increased enamel crest complexity. These features have been hypothesized to reduce whole-tooth wear and extend dental longevity as increasingly graze-dominated diets evolved within the lineage. This hypothesis has yet to be directly tested. Here, in this study, we developed an in-silico wear model using experimentally derived wear rates from fossil and extant proboscidean dental tissues. The models revealed that shifts in tissue topology do not affect whole-tooth wear rate, as inverse trends in lamellar frequency and enamel thickness preserve a consistent surface enamel area fraction; the determining factor of wear. Rather, topological shifts produce a wear-emergent secondary occlusal surface with greater numbers of triturating crests that create a regular, low-relief, file-like shearing pavement. These changes in occlusal architecture likely directly impacted the mastication capacity of Mammuthus dentitions, facilitating their dietary expansion to incorporate fibrous, lower-nutrient graze.

Dental wear

In Silico Chemical Experiments in the Age of AI: From Quantum Chemistry to Machine Learning and Back

Computational chemistry is an indispensable tool for understanding molecules and predicting chemical properties. However, traditional computational methods face significant challenges due to the difficulty of solving the Schrödinger equations and the increasing computational cost with the size of the molecular system. In response, there has been a surge of interest in leveraging artificial intelligence (AI) and machine learning (ML) techniques to in silico experiments. Integrating AI and ML into computational chemistry increases the scalability and speed of the exploration of chemical space. However, challenges remain, particularly regarding the reproducibility and transferability of ML models. This review highlights the evolution of ML in learning from, complementing, or replacing traditional computational chemistry for energy and property predictions. Starting from models trained entirely on numerical data, a journey set forth toward the ideal model incorporating or learning the physical laws of quantum mechanics. This paper also reviews existing computational methods and ML models and their intertwining, outlines a roadmap for future research, and identifies areas for improvement and innovation. Ultimately, the goal is to develop AI architectures capable of predicting accurate and transferable solutions to the Schrödinger equation, thereby revolutionizing in silico experiments within chemistry and materials science.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Toward Tunable Magnetic Dirac Semimetals: Mn Doping of Cd 3 As 2

Magnetic impurities provide a route toward increasing functionality in electronic materials, often enabling new device concepts and architectures. In the case of topological semimetals, dilute magnetic doping presents a particularly attractive approach for inducing a Dirac to Weyl phase change via time reversal symmetry breaking. However, efforts to realize changes in the electronic structure have been limited by challenges in incorporating magnetic impurities into crystals with sufficiently high electron mobilities to detect them via transport or spectroscopic techniques. Here, we demonstrate incorporation of Mn into Cd 3 ⁢As 2 Dirac semimetal thin films grown by molecular beam epitaxy (MBE). Using As-rich growth conditions and [001] oriented thin films, Mn compositions of >10% are achieved. Films contain uniform distributions of Mn with no evidence of secondary phases and exhibit electron mobilities greater than 10 000–30 000 cm 2 /Vs up to 5% Mn. An evolution in the magnetization behavior along with the emergence of a second quantum oscillation frequency at low Mn concentrations provide preliminary evidence of Mn-induced changes in the electronic structure that are consistent with a Weyl phase. This work demonstrates the potential of magnetically doping topological semimetal thin films and a pathway for synthesizing them.

36 MATERIALS SCIENCE

Characterization of lateral amorphous selenium photodetectors for low-photon and VUV detection at cryogenic temperatures

The performance of amorphous selenium (a-Se) as a cryogenic photodetector material is evaluated through a series of experiments using laterally structured devices operated in a custom optical test stand. These studies investigate the response of a-Se detectors to low-photon fluxes at high electric fields near avalanche conditions, the linearity of the photoconductive response over a wide dynamic range and the direct detection of narrowband 130 nm vacuum ultraviolet (VUV) illumination. At 87 K, matched-filter analysis shows reliable single-shot detection with efficiencies ≥80% and area under the curve (AUC) ≥ 0.85 using as few as ∼ 6800 incident 401 nm photons, corresponding to ∼ 3400 photons within field-active regions after accounting for geometric constraints. Measurements are performed at cryogenic temperatures using calibrated photon fluxes derived from a silicon photomultiplier reference and a characterized optical filter stack. Additional experiments using a tellurium-doped a-Se (a-SeTe) device explore the material's behavior under identical test conditions and demonstrate that avalanche is achievable in a-SeTe at cryogenic temperatures. The results demonstrate reproducible low-noise operation, VUV sensitivity and field-dependent gain behavior in a lateral a-Se architecture, representing the first reported observation of avalanche multiplication in laterally structured a-Se and a-SeTe devices at cryogenic temperatures. These findings support the potential integration of laterally structured a-Se devices into next-generation pixelated liquid-argon time projection chambers (TPCs) requiring scalable, high-field-compatible photon detection systems.

Amorphous selenium

Dynamic asymmetric strain imprinted into substrates by an oxide thin film

In film-substrate systems, the substrate role is often considered to be limited to providing static mechanical constraints. Dynamic film-substrate interactions when a structural change in the film modifies the substrate are generally disregarded. Here, using combined x-ray and electron microscopies, we observed that an electrically induced filament in a vanadium dioxide film created strong asymmetric strain in an underlying sapphire substrate. This asymmetric substrate strain fed back into the film and defined the filament expansion direction, revealing the importance of film-substrate dynamic interactions in determining film functionality. Furthermore, the strain imprint propagated at least tens of micrometers deep into the substrate, exceeding the film thickness by more than 200-fold, potentially enabling substrate functionalization as an active mechanical coupling media in three-dimensional integrated microelectronic architectures.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND

Deep Learning and Photogrammetric Reconstruction for Automated Crack Detection and Dimensional Measurement in Mining Operations

Surface crack detection and dimensional measurement at active mining sites present significant safety and operational challenges. Manual inspection methods are labor-intensive, spatially incomplete, and expose personnel to hazardous environments, while existing automated approaches have been developed primarily for concrete civil infrastructure and have not been validated on the complex, variable surfaces characteristic of mining environments. This dissertation presents an automated pipeline that integrates deep learning semantic segmentation with Structure-from-Motion photogrammetry to detect surface cracks and measure their aperture, length, and vertical displacement from standard RGB imagery acquired during routine Uncrewed Aerial Vehicle (UAV) survey operations, without requiring additional sensor hardware or manual measurement. The pipeline combines a U-Net architecture with an EfficientNet-B0 encoder, pretrained on the SDNET2018 concrete crack dataset and fine-tuned on a mining-specific dataset spanning laboratory concrete specimens, coal refuse impoundment embankments, and post-blast limestone quarry benches. Photogrammetric reconstruction is performed using COLMAP Structure-from-Motion and Multi-View Stereo, with crack segmentation masks projected into the reconstructed point cloud to enable three-dimensional vertical displacement measurement through local plane fitting and bimodal surface detection. The pipeline was validated across 36 controlled laboratory specimens at three imaging distances and four vertical displacement levels, achieving aperture measurement RMSE of 0.047 cm and R² of 0.954, and vertical displacement RMSE of 0.140 cm and R² of 0.966, against independent caliper measurements. Field application at a coal refuse impoundment in southwestern Pennsylvania detected 71 crack components across the embankment crest, with a dominant longitudinal crack exhibiting aperture values reaching 28 cm and a 95th percentile vertical displacement of 35.53 cm, consistent in magnitude and spatial distribution with simultaneously acquired LiDAR-derived estimates. Application across four post-blast limestone quarry bench datasets in California successfully characterized blast-induced fracture networks at ground sampling distances ranging from 0.59 to 1.23 cm/pixel, with detected crack geometries physically consistent with observable surface conditions at each site. The results demonstrate that deep learning-based crack detection and photogrammetric measurement can be integrated into routine UAV inspection workflows at mining sites, providing repeatable, scalable, and quantitative crack characterization across surface types, crack scales, and displacement magnitudes not previously addressed in the literature. The pipeline requires no dedicated surveying equipment beyond the UAV platforms already deployed at mine sites for survey and monitoring purposes, supporting practical adoption within existing operational workflows.

Crack detection, Dimensional Measurement

Wide-Bandgap Semiconductor Amplifiers for Fusion Plasma Heating and Control

This paper discusses power electronics developed under the ARPA-E GAMOW program to support nuclear fusion power production. The goal of this project was to develop and assess the potential for wide-bandgap (WBG) semiconductor devices in power electronics to enable high-efficiency and high-voltage solid-state systems for fusion plasma generation, heating, and control. The power electronics use an architecture in which multiple high-power boards can be combined to produce megawatt-level power, where using multiple boards provides high reliability. Two main areas of power electronics boards are developed in this project for fusion plasma heating and control applications: (1) pulse generation and control and (2) radiofrequency generation. The first area is for boards capable of driving high-voltage millisecond pulses at high duty cycles. The envisioned application of these pulses is in plasma control of magnetohydrodynamic instabilities, plasma position, and edge-localized modes. Pulse-width modulation allows for the implementation of a wide variety of linear and nonlinear control systems. The boards developed for this project could actuate control coils based on digital input signals and can be parallelized to provide megawatts of output power. The design of the pulse generator is a low-side load switch. A load switch was designed and constructed that utilized 2-kV-rated field-effect transistor (FET)-based cascodes developed by Qorvo under this project to perform initial testing of these cascodes. The second area is being implemented using class E amplifiers with WBG devices and a reactance steering network to handle inductive or capacitive plasma loads. Applications include ion cyclotron resonance heating (ICRH) and high-harmonic fast-wave (HHFW) heating. A class E reactance steering network is demonstrated in modeling and experiment with a resistive-inductive load that models an inductively-coupled plasma. Power combining of boards with class E reactance steering networks is also simulated and demonstrated experimentally, to enable scaling up to high power. Modeling of high-power-density cooling and remaining useful life is conducted to enable reliable, effectively cooled high-power electronics for fusion applications.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Laser–plasma amplification of an ultrabroadband laser pulse to 0.3 TW

Producing on-target laser intensities much greater than 10 23 W cm −2 with current laser technologies is a roadblock to accessing new regimes of physics such as strong-field quantum electrodynamics. Laser–plasma amplifiers show promise to realize these intensities by augmenting the final amplifier and compressor in traditional chirped-pulse-amplification architectures with a plasma-based amplification and compression stage that operates at a much higher damage threshold. Here we demonstrate amplification of an ultrabroadband (>60 nm) pulse in a laser–plasma Raman amplifier. We directly amplified seed intensities up to 3.7 × 10 15 W cm −2 and measured efficiencies up to 8.7%. Single-shot SPIDER measurements show a factor-of-2 reduction in the amplified pulse duration with final powers up to 0.3 TW, a 10× improvement over previous results. Final pulse durations of 64 fs are measured. Energy transfers greater than 220 mJ from the picosecond pump into the seed result in a 30× energy amplification of a 7.6 mJ seed. These results set the stage for a compact plasma afterburner based on Raman amplification that could extend the scientific capability of existing petawatt-class laser facilities to enable experiments at the intensity frontier.

Shaw, J. L. [Univ. of Rochester, NY (United States

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE