Search NASASearch

SEARCH · Search NASA

Results for “model-predictive HVAC control (MPC)”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,432 records · Page 4

Impact of Crystalline Phases on Low-Activity Waste Glass Durability: Insights from PCT and VHT

During vitrification of nuclear wastes, slow cooling along the container centerline promotes crystalline phase formation, which can alter residual glass composition and reduce chemical durability. This study investigates the effects of crystalline phases on the chemical durability of low-activity waste (LAW) borosilicate glasses using the product consistency test (PCT) and vapor hydration test (VHT) on container centerline cooled (CCC) samples. A preliminary model (R2 = 0.88) was developed to predict CCC PCT responses based on glass composition, PCT data from quenched glasses, and measured crystal fractions. Using the latest LAW glass dataset, the feasibility of predictive modeling is evaluated, limitations in current data and methods are identified, and challenges for improving model accuracy are discussed to guide future data collection and model development.

borosilicate glass

Introduction to the Special Issue on Advanced Air Mobility Noise: Predictions, Measurements and Perception

This Special Issue focuses on noise associated with Advanced Air Mobility (AAM), an emerging class of predominantly electric distributed-propulsion aircraft designed for urban and regional transportation. As these vehicles move toward certification and deployment, noise has become a central challenge for regulatory approval and public acceptance, particularly due to operations in densely populated areas and at low altitudes. The 24 contributions in this issue address three key aspects of AAM noise: prediction, measurement, and human perception. Prediction studies span a wide range of modeling fidelities, from high-resolution simulations to improved semi-analytical approaches, and examine complex aeroacoustic mechanisms including rotor interactions, turbulence ingestion, and broadband noise generation. Measurement studies, largely at model scale, provide new insights into tonal and broadband noise characteristics across configurations and operating conditions, while supporting model validation. Perception-focused contributions investigate annoyance, sound quality metrics, and auralization, emphasizing the role of context and operational factors in shaping human response. Together, these works highlight the interdisciplinary nature of AAM noise research and the need for integrated approaches to enable quieter vehicle design.

Perception

Summary of Research Report

Ten papers, published in various publications, on buckling, and the effects of imperfections on various structures are presented. These papers are: (1) Buckling mode localization in elastic plates due to misplacement in the stiffner location; (2) On vibrational imperfection sensitivity on Augusti's model structure in the vicinity of a non-linear static state; (3) Imperfection sensitivity due to elastic moduli in the Roorda Koiter frame; (4) Buckling mode localization in a multi-span periodic structure with a disorder in a single span; (5) Prediction of natural frequency and buckling load variability due to uncertainty in material properties by convex modeling; (6) Derivation of multi-dimensional ellipsoidal convex model for experimental data; (7) Passive control of buckling deformation via Anderson localization phenomenon; (8)Effect of the thickness and initial im perfection on buckling on composite cylindrical shells: asymptotic analysis and numerical results by BOSOR4 and PANDA2; (9) Worst case estimation of homology design by convex analysis; (10) Buckling of structures with uncertain imperfections - Personal perspective.

Isaac Elishakoff

The Role of Nuclear Data Sensitivities in Prompt α-Eigenvalue Predictions of Delayed Critical Benchmarks

Alpha (α) eigenvalues, which describe the logarithmic time derivative of the neutron population in a multiplying system, are integral to time-dependent behavior and diagnostic applications. However, uncertainties in the evaluated nuclear data can significantly impact the accuracy of transport simulations for such quantities. This work explores the use of machine learning models to predict two key outputs, α-eigenvalues and keff bias, using input features derived from α-eigenvalue sensitivities to nuclear data. The criticality safety benchmark models used in this study come from the International Handbook of Evaluated Criticality Safety Benchmark Experiments. Three models, random forest, XGBoost, and NGBoost, are trained on both energy-resolved and energy-summed α sensitivities. For the α-eigenvalue bias prediction, NGBoost achieved the highest R 2 (0.9476) using energy-resolved features, while XGBoost performed best using summed sensitivities. In contrast, when predicting the keff bias, all the models showed moderate predictive capability (best R 2 ≈ 0.72), as the mapping from the static α-sensitivities to the static keff bias was less direct. SHAP (SHapley Additive exPlanations) analysis was used to interpret the model predictions. Across both prediction tasks, the features associated with neutron capture [H-1 (n, γ)], uranium scattering reactions (such as 235 U elastic/inelastic), and actinide capture/fission reactions (such as 239 Pu and 234 U) were consistently identified as the most impactful. This highlights the key role of specific nuclear reactions and energy ranges in shaping both time-dependent and steady-state criticality behavior. These results demonstrated that α-sensitivities, despite being computed for time-dependent metrics, can provide valuable insights for predicting both α-eigenvalues and the keff bias. Moreover, machine learning models offer a promising pathway for uncovering important nuclear data dependencies and guiding future data evaluation efforts.

Nuclear data

System Identification for Integrated Aircraft Development and Flight Testing [l'Identification Des Systemes Pour le Developpement Integre des Aeronefs et les Essais en Vol]

Over the last decades flight vehicles such as aircraft and helicopters entering service and requiring increased operational effectiveness have with few exceptions experienced prolonged flight test development to achieve full certification. In many cases the original requirements had later to be reduced to enable release to service. The impact on the customer, and manufacturer has been considerable leading to increased costs and or reduced operational capabilities. These costly experiences are largely a result of the flight vehicle not behaving as modelled and designed. The evaluation of flight test data can be used as a tool for validating windtunnel results and mathematical models describing the flight dynamical behaviour. In this sense the uncertainty of important aerodynamic stability and control parameters can be reduced and the confidence of aircraft mathematical models improved. An additional important factor comes from the implementation of active control systems offering the promise of significantly increased flight vehicle performance and operational capability. This approach extends the traditional trade-offs between aerodynamics, structures and propulsion systems to include full- time, full-authority fly-by-wire/light systems. It is imperative that the aerodynamic stability and control parameters of such integrated flight and propulsion control systems have to turn out inflight as predicted, since inherent stability margins will be lower and the flight control system must correct these deficiencies to provide flight critical redundancy and safety. With the methodology of system identification from flight tests it is possible to sense the control inputs and the flight vehicle reactions Such as accelerations, rates and attitudes. The mathematical model, e.g. the model structure and parameters, has to be determined from the relationship of the measured control inputs and the system's responses. The aim of this symposium was to review the present state of the art of flight vehicle system and parameter identification techniques, and to provide a critical appraisal of current methods developed and applied to flight test data in a number of NATO nations. Particular emphasis was placed on practical aspects and lessons learned in order to generate information useful to the flight test community in industry and government agencies. The technical papers share invaluable experience and emphasize the advances of flight vehicle system identification over the last years to the point where confidence and robustness level is now reasonably high. The symposium covered overviews of identification methodologies, flight test techniques, recent aircraft and helicopter application programs, and a session of short papers covering up-to-the-minute flight test results. A final discussion included prepared comments from experts and concluded with key issues learned in the application of system identification and future research needs. The essential benefits to NATO nations can be condensed as follows: More accurate mathematical models for high bandwidth flight control systems, Improved assessment and evaluation of flying qualities, High fidelity mathematical models for flight vehicle development and mission training simulators, and generally, Reduced flight test time and costs.

Advisory Group for Aerospace Research and Developm

Machine Learning for Predicting Team Functioning in HERA Missions

Team functioning is integral to success in future long term space exploration missions. Proactively detecting declines in team functioning can mitigate conflict and ensure mission success. This project developed a speech-based artificial intelligence (AI) system that unobtrusively predicts degradation in team functioning, including performance and cohesion, in the Human Exploration Research Analog (HERA) Campaigns 4 and 5. The AI system conducted automated analysis of the prosodic (tone of voice) and linguistic (language content) components of speech, modeling interpersonal dynamics at both the turn-taking and day-wide levels. We investigated team functioning via observing structured interactions (i.e., multi-mission space exploration vehicle-extra vehicular activity [MMSEV-EVA], team interaction battery [TIB]) and unstructured interactions before the MMSEV-EVA task. We developed machine learning models to predict team functioning (objective task accuracy, self reported team efficacy and self reported team cohesion) by analyzing OpenSmile acoustic features, linguistic descriptors extracted via the linguistic inquiry and word count (LIWC) dictionary, and semantic embeddings. In the TIB, static models using logistic regression and random forests were not able to predict task accuracy, but predicted team efficacy and cohesion during both the decision making and relational tasks to a moderate level (60-70%). Majority voting on the individual turns to predict day long team efficacy further increased accuracies (70-80%). Finally, long short-term memory (LSTM) models showed the best performance across all variables (80-91%), including task performance. In the MMSEV-EVA, static models achieved an accuracy of 60% with majority voting, which increased to 80% through the incorporation of mission day as a variable, accounting for the learning effect. A key finding across both tasks was the "team-dependent" nature of these interactions; models achieved much higher accuracy when trained on prior days of the same team's data rather than attempting to generalize across entirely different teams, with even 1-2 days of prior data per team achieving 5-15% improvement over team-independent models. In addition, the incorporation of pre-task data from the same team also improves model performance, e.g., incorporating data from the decision-making task of the TIB, which preceded the relational task, improved the prediction of team efficacy and cohesion during the latter. We compared model performance when trained on machine-generated data compared to data that had been further corrected by human annotators. Overall, models trained on human-corrected data exhibited a modest improvement in performance, particularly when acoustic features were used. We found no significant correlation between word error rate (WER) and model accuracy (r(55) = -0.08, p = 0.51), but model’s accuracy was significantly higher for medium/high quality transcription (0.74 (SD = 0.48)) compared to the low-quality group (0.64 (SD = 0.36)) (t(63)=2.82, p = 0.006). Based on these, several design recommendation emerge, that could inform Standards at NASA. Models predicting team functioning should incorporate at least one to two days of historical interaction data, include brief pre-task discussions, and explicitly model temporal learning effects, especially for longer operational tasks. Minimum quality standards for automated speech-processing pipelines are needed, given the performance gains observed with manually corrected acoustic data. Finally, systems should leverage both acoustic features and language embeddings in complementary ways, with modality choices and fusion strategies tailored to mission context, task demands, and data quality requirements.

Shrivatsa Mishra

MHONGOOSE: A MeerKAT nearby galaxy H I survey

The MHONGOOSE (MeerKAT H IObservations of Nearby Galactic Objects: Observing Southern Emitters) survey maps the distribution and kinematics of the neutral atomic hydrogen (H I) gas in and around 30 nearby star-forming spiral and dwarf galaxies to extremely low H Icolumn densities. The H Icolumn density sensitivity (3σover 16 km s −1 ) ranges from ∼5 × 10 17 cm −2 at 90″ resolution to ∼4 × 10 19 cm −2 at the highest resolution of 7″. The H Imass sensitivity (3σover 50 km s −1 ) is ∼5.5 × 10 5 M ⊙ at a distance of 10 Mpc (the median distance of the sample galaxies). The velocity resolution of the data is 1.4 km s −1 . One of the main science goals of the survey is the detection of cold accreting gas in the outskirts of the sample galaxies. The sample was selected to cover a range in H Imasses from 10 7 M ⊙ to almost 10 11 M ⊙ in order to optimally sample possible accretion scenarios and environments. The distance to the sample galaxies ranges from 3 to 23 Mpc. In this paper, we present the sample selection, survey design, and observation and reduction procedures. We compared the integrated H Ifluxes based on the MeerKAT data with those derived from single-dish measurement and find good agreement, indicating that our MeerKAT observations are recovering all flux. We present H Imoment maps of the entire sample based on the first ten percent of the survey data, and find that a comparison of the zeroth- and second-moment values shows a clear separation in the physical properties of the H Ibetween areas with star formation and areas without related to the formation of a cold neutral medium. Finally, we give an overview of the H I-detected companion and satellite galaxies in the 30 fields, five of which have not previously been cataloged. We find a clear relation between the number of companion galaxies and the mass of the main target galaxy.

Astronomy & Astrophysics

Understanding and Predicting the Spatially Resolved Adsorption Properties of Nanoporous Materials

Using knowledge from statistical thermodynamics and crystallography, we develop an image–image translation model, called SorbIIT, that uses three-dimensional grids of adsorbate–adsorbent interaction energies as input to predict the spatially resolved loading surface of nanoporous materials over a broad range of temperatures and pressures. SorbIIT consists of a closed-form differential model for loading-surface prediction and a U-Net to generate spatial differential distributions from the energy grids. SorbIIT is trained using the energy grids and adsorbate distributions (obtained from high-throughput simulations) of 50 synthesized and 70 hypothetical zeolites and applied for predicting the adsorption of carbon dioxide, hydrogen sulfide, n-butane, 2-methylpropane, krypton, and xenon in other zeolites from 256 to 400 K. In conclusion, employing a quadratic isotherm model for the local differentiation, SorbIIT yields mean R 2 values of 0.998 for total adsorption and 0.6904 for local adsorption with a resolution of 0.2 Å, and a value of 0.721 for the structural similarity of the local loading distribution.

Sun, Yangzesheng [Univ. of Minnesota, Minneapolis,

First principles nickel-cadmium and nickel hydrogen spacecraft battery models

The principles of Nickel-Cadmium and Nickel-Hydrogen spacecraft battery models are discussed. The Ni-Cd battery model includes two phase positive electrode and its predictions are very close to actual data. But the Ni-H2 battery model predictions (without the two phase positive electrode) are unacceptable even though the model is operational. Both models run on UNIX and Macintosh computers.

Timmerman, P.

High-resolution modeling of indoor radon exposure with uncertainty quantification in Utah

Indoor radon accounts for 37% of population-level exposure to ionizing radiation in the United States. However, radon metrics are typically reported at coarse spatial scales, potentially obscuring meaningful local variation. We developed a high-resolution modeling framework to estimate indoor radon concentrations across Utah while explicitly quantifying predictive uncertainty. A total of 19,497 residential radon measurements collected between 2006 and 2017 were combined with environmental and housing characteristics and analyzed using a geospatial neural network that accommodates spatial dependence and nonlinear associations. Predictions were generated on a uniform hexagonal grid at 0.73 km2 resolution (H3 level 8). Out-of-sample predictions aggregated to the H3 level 8 grid showed good agreement with observed concentrations (Pearson r=0.64), while household-level predictions exhibited more moderate agreement (r=0.45). The model produced well-calibrated uncertainty estimates, with 24.1% of held-out observations exceeding the predicted 75th-percentile threshold. Maps of predicted radon concentrations and the probability of exceeding the U.S. EPA action level of 148 Bq/m3 (4 pCi/L) revealed substantial fine-scale spatial heterogeneity that was not apparent in conventional coarse-resolution summaries, with greater local variability observed in densely monitored urban counties than in sparsely sampled regions. High-resolution radon models that explicitly quantify uncertainty provide a useful framework for characterizing the spatial distribution of indoor radon and identifying areas of elevated exceedance risk. These findings highlight the value of fine-scale monitoring data and uncertainty-aware modeling approaches for radon exposure assessment, environmental risk characterization, and radon-related health research.

Wu, Yunhan [ORNL] (ORCID:0000000178842994)

Performance Evaluation of Intravehicular Activity Spacesuit Without the Use of a Liquid Cooling Garment

The Orion Crew Survival Systems (OCSS) suit is equipped with safety technology that protects the crew during launch and re-entry and Intravehicular Activities (IVA). It was designed to be used with a Liquid Cooling Garment (LCG), an undergarment with tubes that circulate water to remove excess heat from a crew member. The objective of this study is to determine if the Modified Advanced Crew Escape Suit (MACES), which has very similar material, assembly, and pressure characteristics as the OCSS suit, can be used without the LCG to maintain the crew within their heat storage requirements where exceedances can lead to cognitive and physiological impairment. Testing was completed for both hot and cold environments where test subjects wore varying configurations of the MACES suit ensemble without the LCG. The tests were conducted in a temperature and humidity-controlled chamber with four test subjects ranging from extra small to large. Target metabolic rates were reached using an arm ergometer. To represent the different suit configurations seen in the concept of operations, the suit was tested with visor-down, visor-up, and without helmet and gloves. The test data was analyzed to determine how the environment, test subject size, metabolic rate, and suit configuration affected heat storage. Furthermore, the test data was used to correlate the METMAN model, a transient metabolic man program that simulates the heat transfer within a crew member’s body and from a crew member to the surrounding environment. The METMAN correlation showed good agreement between predicted heat storage and test data at the end of the test duration, with a coefficient of determination (R 2 ) value of 0.8999 for visor-down cases and R 2 of 0.8902 for visor-up and helmet off cases. A correlated METMAN model allows for the analysis and prediction of IVA suit performance without an LCG in additional scenarios.

Spacesuit

Performance Evaluation of Intravehicular Activity Spacesuit Without the Use of a Liquid Cooling Garment

The Orion Crew Survival Systems (OCSS) suit is equipped with safety technology that protects the crew during launch and re-entry and Intravehicular Activities (IVA). It was designed to be used with a Liquid Cooling Garment (LCG), an undergarment with tubes that circulate water to remove excess heat from a crew member. The objective of this study is to determine if the Modified Advanced Crew Escape Suit (MACES), which has very similar material, assembly, and pressure characteristics as the OCSS suit, can be used without the LCG to maintain the crew within their heat storage requirements where exceedances can lead to cognitive and physiological impairment. Testing was completed for both hot and cold environments where test subjects wore varying configurations of the MACES suit ensemble without the LCG. The tests were conducted in a temperature and humidity-controlled chamber with four test subjects ranging from extra small to large. Target metabolic rates were reached using an arm ergometer. To represent the different suit configurations seen in the concept of operations, the suit was tested with visor-down, visor-up, and without helmet and gloves. The test data was analyzed to determine how the environment, test subject size, metabolic rate, and suit configuration affected heat storage. Furthermore, the test data was used to correlate the METMAN model, a transient metabolic man program that simulates the heat transfer within a crew member’s body and from a crew member to the surrounding environment. The METMAN correlation showed good agreement between predicted heat storage and test data at the end of the test duration, with a coefficient of determination (R 2 ) value of 0.8999 for visor-down cases and R 2 of 0.8902 for visor-up and helmet off cases. A correlated METMAN model allows for the analysis and prediction of IVA suit performance without an LCG in additional scenarios.

Human Thermal Modeling

Divergent carbon use efficiency-growth rate tradeoff in popular biological growth models

Carbon use efficiency (CUE) is an important trait emerging from processes regulating biological growth. CUE can be computed either based on the growth of structural biomass or total biomass divided by substrate uptake rate. Nonequilibrium thermodynamics and observations suggest that, for an exponentially growing population of cells, structural biomass CUE should first increase, then peak, and finally decrease with specific growth rate; meanwhile, total biomass CUE increases asymptotically with specific growth rate. We compared predictions from six popular models that are often used for plant and microbial growth in existing ecosystem models. We found that, for an exponentially growing population of biological cells, (1) the source-driven Pirt and Compromise models predict that structural biomass CUE increase asymptotically with growth rate; (2) the apparent sink-driven modified Droop model predicts that structural biomass CUE decreases with growth rate; and (3) the sink-driven variable internal storage model and two dynamic energy budget models predict that structural biomass CUE first increases, then peaks, and finally decreases with growth rate. Moreover, the modified Droop model predicts that total biomass CUE is constant with growth rate, while all other five models predict that total biomass CUE increases with growth rate asymptotically. For non-exponential biological growth, we show that there is no static relationship between total biomass CUE or structural biomass CUE with respect to either growth rate or temperature. Therefore, we contend that biological growth models should explicitly represent interactions between substrate acquisition, substate transformation, and maintenance respiration to better capture observed CUE dynamics, and the sink-driven model should be preferred for general ecosystem biogeochemistry modeling.

Tang, Jinyun [Lawrence Berkeley National Laborator

Machine Learning for Predicting Multipactor Susceptibility in Planar RF Structures

Multipactor discharge is a persistent challenge in high-power microwave (HPM) and accelerator systems, where secondary electron avalanches can cause heating, vacuum degradation, and failure. This work presents the first supervised machine learning (ML) framework for multipactor prediction, trained on high-fidelity 3D Particle-in-Cell (PIC) simulation data in planar geometries. The model maps operational, geometric, and material-dependent secondary electron yield (SEY) parameters to the time-averaged electron growth rate, enabling rapid reconstruction of susceptibility charts. Among the models evaluated, tree-based ensemble methods such as Random Forest and Extra Trees demonstrate superior generalization to unseen materials compared to neural networks such as multilayer perceptron (MLP). Performance metrics, including Intersection over Union (IoU), Structural Similarity Index Measure (SSIM), and Pearson correlation, show close agreement with simulation benchmarks. Principal Component Analysis attributes generalization limits to material feature-space disjointedness.

43 PARTICLE ACCELERATORS

Micrometer: Micromechanics transformer for predicting full field mechanical responses of heterogeneous materials

Predicting mechanical responses of heterogeneous materials across scales remains a significant challenge. Traditional computational methods often struggle with complex and multiscale nature of these materials, limiting their effectiveness in real-world applications. Here, in this paper, we introduce Micrometer, a vision transformer based deep learning model designed to predict full field mechanical responses of heterogeneous materials, bridging the gap between computer vision and solid mechanics problems. We show that Micrometer, trained on a large-scale high-resolution dataset of 2D fiber-reinforced composites, can achieve state-of-the-art performance in predicting microscale strain fields across a wide range of material properties and loading conditions. Our model demonstrates accuracy and computational efficiency in applications such as computational homogenization and multiscale modeling, reducing computational time by up to two orders of magnitude compared to conventional numerical solvers while maintaining less than 1 % errors in predicting macroscale stress fields. Furthermore, we showcase Micrometer’s adaptability through transfer learning experiments on new materials with limited data, highlighting its potential to tackle diverse scenarios in computational solid mechanics. These results represent a significant step towards AI-driven innovation in materials science, addressing the limitations of traditional numerical methods and paving the way for more efficient simulations of heterogeneous materials across various industrial applications.

Composite materials

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE

Sparse non-Markovian Noise Modeling of Transmon-Based Multi-Qubit Operations

The influence of noise on quantum dynamics is one of the main factors preventing current quantum processors from performing accurate quantum computations. Sufficient noise characterization and modeling can provide key insights into the effect of noise on quantum algorithms and inform the design of targeted error protection protocols. However, constructing effective noise models that are sparse in model parameters, yet predictive can be challenging. In this work, we present an approach for effective noise modeling of multi-qubit operations on transmon-based devices. Through a comprehensive characterization of seven devices offered by the IBM Quantum Platform, we show that the model can capture and predict a wide range of single- and two-qubit behaviors, including non-Markovian effects resulting from spatiotemporally correlated noise sources. The model’s predictive power is further highlighted through multi-qubit dynamical decoupling demonstrations and an implementation of the variational quantum eigensolver. As a training proxy for the hardware, we show that the model can predict expectation values within a relative error of 0.5%; this is a sevenfold improvement over default hardware noise models. Through these demonstrations, we highlight key error sources in superconducting qubits and illustrate the utility of reduced noise models for predicting hardware dynamics.

open quantum systems & decoherence