Search NASASearch

SEARCH · Search NASA

Results for “distribution networks”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

291 records · Page 3

AutoSourceID-Classifier: Star-galaxy classification using a convolutional neural network with spatial information

Aims.Traditional star-galaxy classification techniques often rely on feature estimation from catalogs, a process susceptible to introducing inaccuracies, thereby potentially jeopardizing the classification’s reliability. Certain galaxies, especially those not manifesting as extended sources, can be misclassified when their shape parameters and flux solely drive the inference. We aim to create a robust and accurate classification network for identifying stars and galaxies directly from astronomical images. Methods.The AutoSourceID-Classifier (ASID-C) algorithm developed for this work uses 32x32 pixel single filter band source cutouts generated by the previously developed AutoSourceID-Light (ASID-L) code. By leveraging convolutional neural networks (CNN) and additional information about the source position within the full-field image, ASID-C aims to accurately classify all stars and galaxies within a survey. Subsequently, we employed a modified Platt scaling calibration for the output of the CNN, ensuring that the derived probabilities were effectively calibrated, delivering precise and reliable results. Results.We show that ASID-C, trained on MeerLICHT telescope images and using the Dark Energy Camera Legacy Survey (DECaLS) morphological classification, is a robust classifier and outperforms similar codes such as SourceExtractor. To facilitate a rigorous comparison, we also trained an eXtreme Gradient Boosting (XGBoost) model on tabular features extracted by SourceExtractor. While this XGBoost model approaches ASID-C in performance metrics, it does not offer the computational efficiency and reduced error propagation inherent in ASID-C’s direct image-based classification approach. ASID-C excels in low signal-to-noise ratio and crowded scenarios, potentially aiding in transient host identification and advancing deep-sky astronomy.

Astronomy & Astrophysics

SODAs: sparse optimization for the discovery of differential and algebraic equations

Differential-algebraic equations (DAEs) integrate ordinary differential equations (ODEs) with algebraic constraints, providing a fundamental framework for developing models of dynamical systems characterized by time-scale separation, conservation laws and physical constraints. While sparse optimization has revolutionized model development by allowing data-driven discovery of parsimonious models from a library of possible equations, existing approaches for dynamical systems assume DAEs can be reduced to ODEs by eliminating variables before model discovery. This assumption limits the applicability of such methods for DAE systems with unknown constraints and time scales. We introduce sparse optimization for differential-algebraic systems (SODAs), a data-driven method for the identification of DAEs in their explicit form. By discovering the algebraic and dynamic components sequentially without prior identification of the algebraic variables, this approach leads to a sequence of convex optimization problems. It has the advantage of discovering interpretable models that preserve the structure of the underlying physical system. To this end, SODAs improves since SODAs is singular numerical stability when handling high correlations between library terms, caused by near-perfect algebraic relationships, by iteratively refining the conditioning of the candidate library. We demonstrate the performance of our method on biological, mechanical and electrical systems, showcasing its robustness to noise in both simulated time series and real-time experimental data.

DAE

Dynamic Diketoenamine Crosslinking Unlocks Vinyl Polymer Vitrimers From β ‐Triketone Chemistry

Covalent adaptable networks (CANs) offer a compelling strategy to unite the mechanical robustness of thermosets with the reprocessability of thermoplastics, yet achieving simultaneous durability, processability, and true recyclability remains challenging. We introduce diketoenamine (DKE) vitrimers derived from β-triketone methacrylate monomers and demonstrate how rational monomer design dictates network processability, viscoelasticity, and recyclability. By systematically varying the spacer length between the β-triketone (TK) moiety and the polymer backbone, we identify a key structure–property relationship that dictates vitrimer behavior. Networks bearing TK pendants minimally displaced from the backbone suppress creep but exhibit limited stress relaxation, whereas extended spacers yield lower glass transition temperatures, higher effective crosslink densities, and efficient stress dissipation, enabling optical transparency and reprocessability. Extending this platform to ultra-high molecular-weight prepolymers introduces physical entanglements as secondary crosslinks, further enhancing dimensional stability without compromising processability. Both mechanical and chemical recycling validate the closed-loop circularity of these materials. Furthermore, these results establish TK methacrylates as a versatile platform for designing high-performance vitrimers that integrate durability, reprocessability, and true closed-loop recyclability.

closed-loop recyclability

Coalition for Community-Supported Affordable Geothermal Energy Systems (C2SAGES)

The C2SAGES project evaluated the feasibility of a community geothermal system for the planned Windy Ridge affordable housing development in Hinesburg, Vermont. Led by GTI Energy with Vermont Gas Systems, LN Consulting, NREL, and Frontier Energy, the work assessed technical design, energy performance, costs, business models, community engagement, maintenance, workforce development, and permitting. The proposed system was designed to serve 100% of the development’s heating, cooling, and domestic hot water loads. Compared with a baseline using air-source heat pumps and natural gas water heating, the geothermal system was estimated to reduce HVAC and domestic hot water energy use by about 45% to 48%, lower operating and maintenance costs, and reduce 30-year life-cycle costs by 37% for Phase 1 and 10% for Phase 2. Technical testing and modeling indicated that the Windy Ridge site is suitable for a community-scale geothermal system. The project also developed borehole field layouts, piping concepts, pump house designs, controls, maintenance plans, and supporting engineering drawings. The business model analysis found that first cost, ownership structure, and customer affordability remain major deployment challenges. Utility-led maintenance and operation were viewed favorably, but traditional utility cost-recovery models may require subsidy or revised financing structures to be practical for affordable housing. Community engagement highlighted the need for clear public education, transparent financing, reliable long-term maintenance, trained technicians, and the potential to pair geothermal systems with weatherization. Overall, the report concludes that community geothermal is technically feasible and offers meaningful energy, emissions, and life-cycle cost benefits, but broader deployment will depend on workable financing models and workforce readiness.

15 GEOTHERMAL ENERGY

Frontier Job-Centric Telemetry Dataset

Comprehensive analysis of high-performance computing (HPC) systems requires linking workload execution to system behavior. This kind of analysis is vital for diagnosing performance issues, managing capacity, detecting anomalous workloads, and understanding how applications interact with system hardware. This job-centric telemetry dataset unifies scheduler job records with node-level measurements, enabling direct association between workloads and their corresponding power, thermal, and performance characteristics. It contains sanitized, scheduler related metadata for 152,400 individual jobs that ran on the Frontier supercomputer and ended on selected days throughout 2024 and 2025, a subpopulation of ~6.8% of the total number of allocated jobs with non-zero run time on the system over that same period. Each is linked with files that contain telemetry time series records of the power utilization and temperature behavior of its allocated nodes and their processors during the run time of the job. Where available, a portion of the job files also contain network performance time series. Jobs are sampled from select days that reflect normal levels of user activity and possess job size distributions with large numbers of leadership class jobs (>20% of Frontier nodes). Jobs in this dataset attempt to best represent successful user workflows.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

DESI Strong Lens Foundry. III. Keck Spectroscopy for Strong Lenses Discovered Using Residual Neural Networks

We present spectroscopic data of strong lenses and their source galaxies using the Keck Near-Infrared Echellette Spectrometer (NIRES) and the Dark Energy Spectroscopic Instrument (DESI), providing redshifts necessary for nearly all strong-lensing applications with these systems, especially the extraction of physical parameters from lensing modeling. These strong lenses were found in the DESI Legacy Imaging Surveys using residual neural networks and followed up by our Hubble Space Telescope program, with all systems displaying unambiguous lensed arcs. With NIRES, we target eight lensed sources at redshifts difficult to measure in the optical range and determine the source redshifts for six, between z s = 1.675 and 3.332. DESI observed one of the remaining source redshifts, as well as an additional source redshift within the six systems. The two systems with nondetections by NIRES were observed for a considerably shorter 600 s at high airmass. Combining NIRES infrared spectroscopy with optical spectroscopy from our DESI Strong Lensing Secondary Target Program, these results provide the complete lens and source redshifts for six systems, a resource for refining automated strong lens searches in future deep- and wide-field imaging surveys and addressing a range of questions in astrophysics and cosmology.

Agarwal, Shrihan [University of Chicago, IL (Unite

Crossover from quantum to classical transport

Understanding the crossover from quantum to classical transport phenomena has become of fundamental importance not only for technological applications due to the creation of sub-10nm transistors – an important building block of our modern life – but also for elucidating the role played by quantum mechanics in the evolutionary fitness of biological complexes. This article provides a basic introduction into the nature of charge and energy transport in the quantum and classical regimes. Here, it will identify the characteristic transport properties in both limits, and demonstrate how they can be connected through the loss of quantum mechanical coherence. I will identify the salient features of the crossover physics, and demonstrate their importance in opening new transport regimes and for understanding efficient and robust energy transport in biological complexes.

charge

An attention-based neural ordinary differential equation framework for modeling inelastic processes

To preserve strictly conservative behavior as well as model the variety of dissipative behavior displayed by solid materials, we propose a significant enhancement to the internal state variable-neural ordinary differential equation (ISV-NODE) framework. In this data-driven, physics-constrained modeling framework internal states are inferred rather than prescribed. The ISV-NODE consists of: (a) a stress model dependent on observable deformation and inferred internal state, and (b) a model of the evolution of the internal states. The enhancements to ISV-NODE proposed in this work are multifold: (a) a partially input convex neural network stress potential provides polyconvexity in terms of observed strain while leaving the inferred state unconstrained, and (b) an internal state flow model uses common latent features to inform novel attention-based gating and drives the flow of internal state only in dissipative regimes. We demonstrated that this architecture can accurately model dissipative and conservative behavior across an isotropic, isothermal elastic-viscoelastic-elastoplastic spectrum with three exemplars, while maintaining fundamental principles by design.

97 MATHEMATICS AND COMPUTING

Thermal Reservoir Networks for Modularly Expandable Thermal Microgrids

The Department of Defense (DoD) faces the substantial challenge of cost-effectively retrofitting one to two installations per month, each comprising approximately 1,000 buildings, to improve resilience, reduce energy consumption, and enhance energy supply security. Achieving these objectives requires optimal system selection and effective risk mitigation during system integration. To address this need, we introduce Platform-Based Design (PBD), a structured, hierarchical methodology adapted from other industrial sectors to the domain of energy system retrofits. We demonstrate the effectiveness of PBD through a techno-economic feasibility study comparing geothermal-coupled thermal energy networks (TENs) with conventional energy systems for heating, cooling, and powering 17 buildings at Joint Base Andrews (JBA) in Maryland. Our analysis illustrates that the PBD approach enables rigorous, data-driven, sequential decision making, resulting in a family of Pareto-optimal systems, among which the TEN emerged as the most promising solution. The selected TEN design integrates geothermal borefields, heat recovery heat pumps, photovoltaic (PV) arrays, and battery storage. Compared to the baseline system – gas heating combined with air-source chillers – the proposed TEN reduces annual imported energy by 74% and peak electricity demand by 45%, achieves a levelized cost of energy of $\$0.210$/kWh, and substantially enhances resilience. Life-cycle costs increase by approximately 6%, and initial investment costs are about 2.5 times higher than the baseline. However, if central plant infrastructure, district loops, and utility-scale PV and battery systems are privately funded and operated, the initial investment would fall below the baseline system cost. Critical to achieving these significant performance improvements were detailed nonlinear dynamic simulations coupling geothermal heat transfer, energy system operation, and realistic feedback control logic. These simulations identified essential design modifications and control strategy refinements that substantially reduced energy use, peak demand, and compressor shortcycling, thereby improving durability and reliability—issues that would have been significantly more expensive to resolve during operation. Additionally, the verification step highlighted sensitivities to key design parameters that could reduce initial investment by approximately $\$2$ million and reduce annual life-cycle costs more than $\$300,000$. We recommend adopting the PBD methodology for future feasibility studies and TEN pilot projects to gain valuable operational experience. Furthermore, we recommend that DoD invest in transferring and scaling the PBD methodology to other installations. This entails developing standardized computational frameworks and component libraries as well as training industry in conducting PBD. Such investments would enable rapid, robust, reliable, and cost-effective retrofits, supporting DoD’s ambitious energy system modernization goals.

Wetter, Michael [Lawrence Berkeley National Labora

Lunar accelerometer network gravitational observatory (LANGO)

With ground-based interferometers detecting hundreds of gravitational-wave (GW) events, GW astronomy has continued to blossom. U.S. and European scientists are developing plans to construct third-generation ground-based interferometers. ESA has proceeded through the mission formulation phase of space-based interferometer in a lower-frequency band, 10 –4 –0.1 Hz. Despite all these exciting developments, there is still a missing frequency band, 0.1–10 Hz. This mid-frequency band is rich with interesting astrophysical events. Coalescence and merger of intermediate-mass black holes (IMBHs) will occur in this frequency band. Coalescing stellar-mass BHs will pass through this frequency band days before they reach the frequency band of LIGO and Virgo. Detection of such signals would enable a mid-frequency detector to issue an advance notice to the high-frequency GW detectors, as well as to optical, x-ray and γ-ray telescopes. We propose Lunar Accelerometer Network Gravitational Observatory (LANGO) to detect GWs in this frequency band. In the first phase (LANGO 1), we propose to deploy four ambient-temperature (250 K) accelerometers in the tetrahedral or in a square configuration on the hemisphere facing the Earth. After successful operation at 250 K, LANGO would be upgraded to a cryogenic (4 K) version with over two orders of magnitude increased sensitivity (LANGO 2), with coherent rejection of seismic noise implemented. LANGO is a full-tensor detector, capable of determining the source direction and wave polarization. Each test mass (TM) is suspended as a pendulum with resonance frequency ∼ 0.01 Hz, thus is only weakly coupled to the lunar surface horizontally. LANGO is designed to detect the relative motion of globally separated, nearly free, TMs by using the Moon as a large quiet platform. LANGO 1 and 2 accelerometers aim at sensitivities ⩽ 10 –11 m s –2 Hz –1/2 and ⩽ 10 –13 m s –2 Hz –1/2 in the horizontal axes over the frequency band of 1 mHz–10 Hz, which yield GW sensitivities 1.1 x 10 -21 Hz -1/2 and 3.8 x 10 -24 Hz -1/2 at 1 Hz, respectively. The LANGO accelerometers will be 10 3 –10 5 times more sensitive than Apollo seismometers. With such sensitivity, LANGO will also make great contribution to the advancement of lunar geophysics.

79 ASTRONOMY AND ASTROPHYSICS

Correlated Ion Transport Governed by Dynamic Local Structure in High Concentration and Localized High Concentration Electrolytes

Understanding the dynamics of cluster formation and network percolation provides the mechanistic link between microscopic solvation structure and transport in concentrated electrolytes, including localized high-concentration electrolytes (LHCEs). Although recent studies have shown LHCEs to form micelle-like aggregates at specific compositions, a quantitative understanding of how solvation structures and transport properties depend on salt–solvent–diluent ratios remains limited. Here, we integrate molecular dynamics with Onsager transport analyses to chart the evolution of solvation microstructure and associated ionic transport in LiFSI/DMC/TTE, achieving good agreement with experimental conductivities across composition. We show that micelle formation arises from a dynamic instability of the cation–anion network, and that network percolation is the primary determinant of conductivity. Two compositional thresholds emerge: A critical network concentration (CNC) and a critical micelle concentration (CMC) that delineate transitions from extended percolating networks to micelle-like clusters and then to fragments. These structural transitions rationalize the nonintuitive conductivity decrease at intermediate dilution despite monotonically decreasing viscosity and provide composition-level design rules for LHCEs.

Mohanakrishnan, Rohith Srinivaas [University of Ca

Microstructural Engineering of Cu-Rich Nanoprecipitate formation in NiCoFeCrCu0.12 High-Entropy Alloy via Severe Plastic Deformation for Enhanced Irradiation Tolerance

This study demonstrates a defect-engineering approach for controlling Cu-rich precipitates in FeNiCrCoCu0.2 high-entropy alloys (Cu-HEAs), delivering a novel pathway for next-generation nuclear reactor materials with superior irradiation resistance. This work establishes that severe plastic deformation (SPD) processing via Shear Assisted Processing and Extrusion (ShAPE) and Friction Stir Layer Deposition (FSLD) creates dense dislocation networks and subgrain boundaries that fundamentally alter precipitation behavior under identical thermal treatments. Atom probe tomography (APT) indicates that SPD produces a metastable, atomically homogeneous solid solution that, upon moderate heat treatment (500°C/10 hour), develops remarkedly stronger Cu clustering than the as-cast counterpart. High-temperature exposure (800°C/100 h) produces near-pure Cu precipitates (~90 at% Cu) with significantly enhanced defect-sink efficacy in SPD-processed alloys: precipitate sizes of 50-60 nm and number densities of 2.7-3.8 × 10¹7 m?³, compared to 89 nm and 0.44 × 10¹7 m?³ in as-cast materials. Collectively, the findings establish defect-mediated precipitation control as a scalable, high-impact route to tailor sink density and distribution in HEAs, enabling microstructures optimized for irradiation tolerance and mechanical robustness in nuclear reactor environments.

Meher, Subhashish

Adaptation of virtual synchronous generators to dynamic conditions in power grids

Virtual synchronous generators (VSGs) are widely adopted as grid-forming controls for inverter-based resources. However, when grid conditions vary significantly as characterized by changes in short-circuit ratio (SCR) and the reactance-to-resistance (X/R) ratio, fixed-gain designs and the commonly used P–Q decoupling assumption can become inaccurate. Such conditions can degrade transient power performance, leading to oscillations, prolonged settling, and overshoot, particularly in stiff-grid operating points. This paper quantifies how grid strength and impedance-dependent coupling affect the active–reactive power dynamics of a conventional VSG over a broad range of SCR and X/R values. An adaptive VSG tuning framework is then developed by combining (i) a coupling-explicit, impedance-parameterized state-space model to enable systematic controller synthesis, (ii) a full-state-feedback law designed via pole placement to meet prescribed damping and settling-time specifications, and (iii) a physics-informed neural network (PINN)–based online grid-impedance estimator that updates controller gains in real time as grid conditions vary. Offline simulations in MATLAB/Simulink and real-time validation on an OPAL-RT platform show that the proposed method preserves consistent damping and settling behavior with reduced overshoot across wide SCR and X/R ranges, compared with fixed-gain VSG baselines.

Adaptive control

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE