Search NASASearch

SEARCH · Search NASA

Results for “Space Operations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

371 records · Page 21

Fluidized Bed Gasification For Conversion of Biomass and Waste Materials to Renewable Hydrogen

This project was undertaken in order to study the potential for hydrogen production, at low cost, from mixtures of biomass and municipal solid waste (MSW). This approach allows for the production of hydrogen with a very low fossil carbon burden, while taking advantage of tipping fees (associated with MSW) to improved process economics. The team sourced three primary feedstocks (wood, MSW, and waste plastics) and characterized them comprehensively using established techniques with a long track record in the field of gasification. All three primary feedstocks were highly reactive and lost most of their mass during initial devolatilization. The production of tars, including heavy tars, was quite high, and was most problematic in the case of the MSW and Waste Plastics feedstocks. Little practical difference was identified between the MSW and Waste Plastics materials, and the addition of bed-forming materials (dolomite and brown alumina) to the feedstocks was found to reduce production of tars during devolatilization under thermogravimetric analysis and/or Fischer Assay conditions. A series of four tests in a lab-scale bubbling-fluidized-bed gasifier, at 50 psig of pressure and about 825 C, confirmed these findings. Pellet feedstocks, broken into fragments, were used for these tests, and pellets comprised of 50% MSW and 50% biomass were found to be the best option in terms of fossil carbon burden, economic potential, and gasification characteristics. Tests were then undertaken in a pilot-scale gasifier facility based on the GTI U-Gas fluidized-bed gasification technology. The feedstock handling train of the 20 TPD U-Gas pilot-scale gasifier located in Des Plaines, IL, was operated under simulated gasification conditions, and the 50/50 pellets were found to be very robust and unproblematic. An improved design for the forward end of the feedstock injection screw of the gasifier was developed and installed. The design approach was based on improved passive cooling of the front-most shroud at the end of the screw, since this approach was found in comprehensive modeling studies to be more than sufficient to accomplish the project objectives associated with this phase of the work, while also avoiding thermal gradients that could have caused heat-stress-induced damage to the refractory around the feedstock inlet if an active cooling approach had been applied. Careful technoeconomic analysis (TEA) of two possible 1000 TPD facilities was carried out. The TEA of conversion of MSW with corn stover in one case, and MSW with woody feedstock in the other case, showed that both had the potential to provide hydrogen at about $1/kg (minimum selling price, 2018 dollar basis). Of the two TEA cases, the one that was based on the conversion of wood in the southeastern USA was found to have slightly better economic potential. The other TEA case was based on a real location in Nebraska and called for corn stover feedstock conversion along with MSW. An underserved communities outreach program plan was developed in cooperation with personnel from the Nebraska Public Power District.

08 HYDROGEN

Formation of Bimetallic Nanoparticles via Exsolution Using a Reducible Metal Oxide Capping Layer

Bimetallic nanoparticles are promising catalysts that can improve performance in heterogeneous catalysis and solid-state electrochemistry. Exsolution is a useful method for forming such nanoparticles; however, it is limited by the elements present within the host oxide lattice. Here, in this work, we develop and demonstrate a strategy to form bimetallic particles from La 0.5 Sr 0.5 Ti 0.94 Ni 0.06 O 3 (LSTN) exsolution and using a reducible SnO 2 capping layer, expanding the range of elements available for bimetallic nanoparticle formation. Using this capping layer strategy, we formed nickel–tin (Ni 0 –Sn 0 ) bimetallic nanoparticles via exsolution. We used in situ near-ambient pressure X-ray photoelectron spectroscopy to monitor surface chemical changes during exsolution, showing that first, SnO 2 volatilized. This SnO 2 loss exposed the perovskite surface of LSTN to reducing conditions, which induced Ni exsolution, and compounded with SnO 2 reduction led to the formation of bimetallic Ni 0 –Sn 0 particles. To evaluate the associated microstructural evolution, we measured grazing incidence small-angle X-ray scattering (GISAXS), which confirmed the loss of the SnO 2 capping layer, and scattering simulations suggested the formation of bimetallic particles. We confirmed the bimetallic nanoparticle composition and morphology by Auger spectroscopy and scanning transmission electron microscopy. The resulting bimetallic nanoparticles were smaller and more thermally stable than the monometallic Ni counterparts on LSTN. This capping layer and exsolution approach allow synthesizing multimetallic nanoparticles and can be applied to other reducible metal oxides and perovskite hosts, broadening the compositional space for advanced catalytic materials.

36 MATERIALS SCIENCE

Laser–plasma amplification of an ultrabroadband laser pulse to 0.3 TW

Producing on-target laser intensities much greater than 10 23 W cm −2 with current laser technologies is a roadblock to accessing new regimes of physics such as strong-field quantum electrodynamics. Laser–plasma amplifiers show promise to realize these intensities by augmenting the final amplifier and compressor in traditional chirped-pulse-amplification architectures with a plasma-based amplification and compression stage that operates at a much higher damage threshold. Here we demonstrate amplification of an ultrabroadband (>60 nm) pulse in a laser–plasma Raman amplifier. We directly amplified seed intensities up to 3.7 × 10 15 W cm −2 and measured efficiencies up to 8.7%. Single-shot SPIDER measurements show a factor-of-2 reduction in the amplified pulse duration with final powers up to 0.3 TW, a 10× improvement over previous results. Final pulse durations of 64 fs are measured. Energy transfers greater than 220 mJ from the picosecond pump into the seed result in a 30× energy amplification of a 7.6 mJ seed. These results set the stage for a compact plasma afterburner based on Raman amplification that could extend the scientific capability of existing petawatt-class laser facilities to enable experiments at the intensity frontier.

Shaw, J. L. [Univ. of Rochester, NY (United States

Integration of Online Cross-Section Generation Capability with Depletion and Transient Solvers in Griffin

Griffin is a Multiphysics Object-Oriented Simulation Environment (MOOSE)-based reactor multiphysics analysis application jointly developed by Argonne and Idaho National Laboratories under the DOENE Nuclear Energy Advanced Modeling and Simulation (NEAMS) program. In FY25, an online crosssection generation capability based on the Self-Shielding Application Programming Interface (SSAPI) was demonstrated for TRISO-fueled reactor problems under steady-state conditions. This fiscal year, that capability was extended to support depletion and transient multiphysics calculations, enabling high-fidelity analyses that generate self-shielded cross sections on the fly from the actual evolving composition and temperature states rather than from pre-tabulated libraries. For depletion, a two-way coupling was established in which SSAPI computes compact-averaged self-shielded cross sections that the depletion solver then uses to advance the Bateman equations, with the updated compositions returned to SSAPI at each step; the depletion module was refactored to support both library-based and SSAPI-based cross sections, and additional logic was added to track daughter isotopes and to exclude minor isotopes for efficiency. For transient analysis, the SSAPI multigroup library was extended with the kinetics data required for time-dependent calculations, the Improved Quasi-Static (IQS) scheme was coupled with SSAPI, and several supporting capabilities were implemented, including a self-shielding treatment that lets control rods and drums move within a self-shielded model, which had previously been impossible and had ruled out rod- and drum-movement transients with on-the-fly cross sections altogether, a new mixing scheme for delayed-neutron precursor decay constants, a checkpoint-based restart workflow, and performance improvements such as pointwise cross-section interpolation and the bypassing of unnecessary Dancoff factor calculations. The implemented capabilities were verified against Serpent Monte Carlo solutions. For depletion, a prismatic pin-cell problem based on a Next Generation Nuclear Plant (NGNP) Very High Temperature Reactor benchmark showed excellent agreement, with eigenvalue differences within 200 pcm over the entire burnup range (up to 140 MWD/kgU) and fission-product and actinide inventories agreeing to within 0.8% and 2.5%, respectively; a heat-pipe microreactor assembly problem with a much higher fuel loading confirmed the same behavior and quantified the bias introduced when the multigroup equivalence effect is neglected. For transient analysis, a pin-cell problem with a step reactivity insertion and temperature feedback reproduced the analytically expected asymptotic power and showed close agreement between the direct and IQS solutions, and a two-dimensional microreactor core problem with control-drum rotation exercised the new moving-drum self-shielding treatment and demonstrated successful coupling of the online crosssection generation with both the direct and IQS transient methods. The capability was further exercised on a full-core pebble-bed problem, in which Griffin was coupled with the System Analysis Module (SAM) to simulate load-following operation of the gPBR with the Doppler feedback resolved at the TRISO fuel kernel temperature. These developments in Griffin provide a convenient, high-fidelity approach to cross-section generation for advanced thermal reactors with geometrically complex and highly heterogeneous configurations, including TRISO-fueled prismatic and pebble-bed systems, and support steady-state, depletion, and transient multiphysics calculations. They also enable self-shielded cross sections to be evaluated directly at the actual coupled state of the system, thereby establishing a foundation for high-fidelity, fully coupled multiphysics analysis of advanced reactors

Park, H.

Toward Trustworthy Autonomous Science: A Two-Year Community Roadmap

One year ago, the AISLE roadmap argued that autonomous laboratories operated as isolated islands and proposed a grassroots network organized around five critical dimensions. The field has since moved faster than that roadmap anticipated: multi-agent systems have produced experimentally validated hypotheses, self-driving laboratories have grown more interoperable and orchestrated, reasoning-trained and domain foundation models have raised the capability ceiling, and the Genesis Mission has placed autonomous experimentation at the center of U.S. federal science strategy, with industry emerging as a primary actor. Progress has met a sobering counter-current, including a corrected flagship discovery result, benchmarks showing that agents which rival experts on closed-ended questions still complete only a fraction of open-ended research, and fabricated citations surfacing at leading venues. We read this as the defining tension of the field: producing a candidate discovery is no longer the hard part, but verifying it is, and this asymmetry now limits autonomous science more than raw model capability. Accordingly, we update the roadmap around seven dimensions, revisiting the original five and elevating two former cross-cutting concerns, trust, verification, and reproducibility, and safety, security, and governance, to first-class status. We assess the original milestones (M1 through M14) as achieved, partially achieved, reframed, or open, add four new milestones (M15 through M18) for the elevated dimensions, and scope the path forward to a two-year horizon, with the first year concentrating on interfaces, protocol adoption, and the scaffolding of verification, and the second targeting federation, zero-trust coordination, and governance. Throughout, we position the grassroots network as the interoperability fabric that lets national programs, international initiatives, and commercial platforms connect rather than re-silo.

99 GENERAL AND MISCELLANEOUS

Plasma phosphorylated tau217 strongly associates with memory deficits in the Alzheimer’s disease spectrum

Abstract Plasma phosphorylated tau (p-tau) biomarkers open unprecedented opportunities for identifying carriers of Alzheimer’s disease pathophysiology in early disease stages using minimally invasive techniques. Plasma p-tau biomarkers are believed to reflect tau phosphorylation and secretion. However, it remains unclear to what extent the magnitude of plasma p-tau abnormalities reflects neuronal network disturbance in the form of cognitive impairment. To address this question, we included 103 cognitively unimpaired elderly and 40 cognitively impaired, amyloid-β-positive individuals from the TRIAD cohort, in addition to 336 cognitively unimpaired and 216 cognitively impaired, amyloid-β-positive older adults from the BioFINDER-2 cohort. Participants had tau PET scans, amyloid PET scans or amyloid CSF, p-tau217, p-tau181 and p-tau231 blood measures, structural T1-MRI and cognitive assessments. In this cross-sectional study, we used regression models and correlation analyses to assess the relationship between plasma biomarkers and cognitive scores. Furthermore, we applied receiver operating characteristic curves to assess cognitive impairment across plasma biomarkers. Finally, we categorized participants into amyloid (A), p-tau (T1) and tau PET (T2) positive (+) or negative (−) profiles and ran non-parametric comparisons to assess differences across cognitive domains. We found that plasma p-tau217 was more associated with cognitive performance than p-tau181 and p-tau231 and that this relationship was particularly strong for memory scores (TRIAD: βp-tau217 = −0.53, βp-tau181 = −0.35 and βp-tau231 = −0.24; BioFINDER-2: βp-tau217 = −0.52, βp-tau181 = −0.24 and βp-tau231 = −0.29). Associations in amyloid-β-positive participants resembled these results, but other cognitive scores also showed strong associations in cognitively impaired individuals. Moreover, plasma p-tau217 outperformed plasma p-tau181 and plasma p-tau231 in identifying memory impairment (area under the curve values for TRIAD: p-tau217 = 0.86, p-tau181 = 0.77 and p-tau231 = 0.75; and for BioFINDER-2: p-tau217 = 0.86, p-tau181 = 0.76 and p-tau231 = 0.81) and in identifying executive function impairment only in the BioFINDER-2 cohort (p-tau217 = 0.82, p-tau181 = 0.76 and p-tau231 = 0.76). Lastly, we showed that subtle memory deficits were present in A+T1+T2− participants for plasma p-tau217 (P = 0.007) and plasma p-tau181 (P = 0.01) in the TRIAD cohort and for all biomarkers across cognitive domains in A+T1+T2− and A+T1+T2− individuals (P < 0.001 in all) in the BioFINDER-2 cohort. The A+T1+T2− individuals showed cognitive deficits in both cohorts (P < 0.001 in all). Together, our results suggest that plasma p-tau217 stands out as a biomarker capable of identifying memory deficits attributable to Alzheimer’s disease and that memory impairment certainly occurs in amyloid-β- and plasma p-tau-positive individuals who have no significant amounts of tau in the neocortex.

Neurosciences & Neurology

First multi-institutional systematic comparison of the neutron ambient dose equivalent produced by proton therapy systems

Objective. Isochronous cyclotrons, synchrocyclotrons, and synchrotrons are used to accelerate protons for proton therapy. An accurate measurement of neutron doses generated by these accelerators and associated delivery systems and its clinical relevance requires systematic protocols and proper neutron dosimetry for a meaningful assessment. We present the first comprehensive comparison of neutron ambient dose equivalent (H*(10)) produced by clinically operational proton therapy systems. Approach. Treatment plans with 10 cm modulation-depth and ranges of 10 cm (R10M10) and 25 cm (R25M10) were created to cover a 10 × 10 × 10 cm 3 water target. The pencil beam scanning proton therapy machines studied were: two gantry-mounted synchrocyclotrons (Hyperscan, Mevion, half-gantry), two isochronous cyclotrons (ProBeam, Varian, full-gantry), one isochronous cyclotron (Proteus, IBA, full-gantry), and two synchrotrons (PROBEAT, Hitachi, full- and half-gantry). Proton beams were delivered to 30 × 30 × 40 cm 3 plastic water phantoms. WENDI-II and LUPIN-BF3-NP neutron rem-meters were positioned at three angles (0°, 45°, 90°) relative to the beam direction to measure the neutron H*(10) at distances between 50–300 cm from the isocenter. Main results. H*(10) showed dependence on beam energy, machine type, and measurement location. The highest reading was for the gantry-mounted synchrocyclotron, whereas other systems produced approximately comparable neutron doses. In all cases, the H*(10) reduced with distance from the isocenter. The H*(10) drop at 2 m distance compared to that at 0.5 m was a factor of ∼5 for the gantry-mounted synchrocyclotron whereas in other systems the decrease was a factor of 10. The WENDI-II device suffered from dead-time-associated under-estimation of the dose by a factor of ∼2–3 under the synchrocyclotron beam due to its high dose-per-pulse. However, WENDI-II and LUPIN-BF3-NP results were within reasonable agreement in isochronous cyclotron and synchrotron beams, indicating that both devices are suitable for those systems. Significance. Neutron H*(10) is dependent on various parameters including beam energy, measurement location, as well as machine design. Caution must be exercised in choosing the appropriate neutron-dose-measurement device to be used for low-duty-factor, particularly in high-instantaneous-rate proton delivery systems. By delivering the same volumetric proton dose across different machines, this work provides a benchmark for inter-system comparisons and serves as a foundation for future studies.

LUPIN

Combined dark matter search towards dwarf spheroidal galaxies with Fermi -LAT, HAWC, H.E.S.S., MAGIC, and VERITAS

Dwarf spheroidal galaxies (dSphs) are excellent targets for indirect dark matter (DM) searches using gamma-ray telescopes because they are thought to have high DM content and a low astrophysical background. The sensitivity of these searches is improved by combining the observations of dSphs made by different gamma-ray telescopes. We present the results of a combined search by the most sensitive currently operating gamma-ray telescopes, namely: the satellite-borne Fermi -LAT telescope; the ground-based imaging atmospheric Cherenkov telescope arrays H.E.S.S., MAGIC, and VERITAS; and the HAWC water Cherenkov detector. Individual datasets were analyzed using a common statistical approach. Results were subsequently combined via a global joint likelihood analysis. We obtain constraints on the velocity-weighted cross section 〈σv〉 for DM self-annihilation as a function of the DM particle mass. This five-instrument combination allows the derivation of up to 2-3 times more constraining upper limits on 〈σv〉 than the individual results over a wide mass range spanning from 5 GeV to 100 TeV. Depending on the DM content modeling, the 95% confidence level observed limits reach 1.5×10 -24 cm 3 s -1 and 3.2×10 -25 cm 3 s -1 , respectively, in the τ + τ - annihilation channel for a DM mass of 2 TeV.

79 ASTRONOMY AND ASTROPHYSICS

Low-energy calibration of SuperCDMS HVeV cryogenic silicon calorimeters using Compton steps

Cryogenic calorimeters for low-mass dark matter searches have achieved sub-eV energy resolutions, driving advances in both low-energy calibration techniques and our understanding of detector physics. The energy deposition spectrum of gamma rays scattering off target materials exhibits step-like features, known as Compton steps, near the binding energies of atomic electrons. Here, we demonstrate a successful use of Compton steps for sub-keV calibration of cryogenic silicon calorimeters, utilizing four SuperCDMS High-Voltage eV-resolution detectors operated with 0 V bias across the crystal. This new calibration at 0 V is compared with the established high-voltage calibration using optical photons. The comparison indicates that the detector response at 0 V is about 30% weaker than expected, highlighting challenges in detector response modeling for low-mass dark matter searches.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

The SPT-deep Cluster Catalog: Sunyaev–Zel’dovich Selected Clusters from Combined SPT-3G and SPTpol Measurements over 100 Square Degrees

We present a catalog of 500 galaxy cluster candidates in the SPT-Deep field: a 100 deg$^{2}$ field that combines data from the SPT-3G and SPTpol surveys to reach noise levels of 3.0, 2.2, and 9.0 μK-arcmin at 95, 150, and 220 GHz, respectively. Candidates are selected via the thermal Sunyaev–Zel’dovich (SZ) effect with a minimum significance of ξ = 4.0, resulting in a catalog of purity ∼89%. Optical data from the Dark Energy Survey and infrared data from the Spitzer Space Telescope are used to confirm 442 cluster candidates. The clusters span 0.12 < z ≲ 1.8 and 1.0 × 10$^{14}$M$_{⊙}$/h$_{70}$ < M$_{500c}$ < 8.7 × 10$^{14}$M$_{⊙}$/h$_{70}$. The sample’s median redshift is 0.74, and the median mass is 1.7 × 10$^{14}$M$_{⊙}$/h$_{70}$; these are the lowest median mass and highest median redshift of any SZ-selected sample to date. We assess the effect of infrared emission from cluster member galaxies on cluster selection by performing a joint fit to the infrared dust and tSZ signals by combining measurements from SPT and overlapping submillimeter data from Herschel/SPIRE. We find that at high redshift (z > 1), the tSZ signal is reduced by $17.9_{−3.2}^{+3.8}$%$(3.8_{−0.7}^{+0.9}$%$)$ at 150 GHz (95 GHz) due to dust contamination. We repeat our cluster finding method on dust-nulled SPT maps and find the resulting catalog is consistent with the nominal SPT-Deep catalog, suggesting dust contamination does not significantly impact the SPT-Deep selection function; we attribute this lack of bias to the inclusion of the SPT 220 GHz band.

79 ASTRONOMY AND ASTROPHYSICS

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE