Search NASA⌕ Search

SEARCH · Search NASA

Results for “ensemble modeling system”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Systematic and objective evaluation of Earth system models: PCMDI Metrics Package (PMP) version 3

Systematic, routine, and comprehensive evaluation of Earth system models (ESMs) facilitates benchmarking improvement across model generations and identifying the strengths and weaknesses of different model configurations. By gauging the consistency between models and observations, this endeavor is becoming increasingly necessary to objectively synthesize the thousands of simulations contributed to the Coupled Model Intercomparison Project (CMIP) to date. The Program for Climate Model Diagnosis and Intercomparison (PCMDI) Metrics Package (PMP) is an open-source Python software package that provides quick-look objective comparisons of ESMs with one another and with observations. The comparisons include metrics of large- to global-scale climatologies, tropical inter-annual and intra-seasonal variability modes such as the El Niño–Southern Oscillation (ENSO) and Madden–Julian Oscillation (MJO), extratropical modes of variability, regional monsoons, cloud radiative feedbacks, and high-frequency characteristics of simulated precipitation, including its extremes. The PMP comparison results are produced using all model simulations contributed to CMIP6 and earlier CMIP phases. An important objective of the PMP is to document the performance of ESMs participating in the recent phases of CMIP, together with providing version-controlled information for all datasets, software packages, and analysis codes being used in the evaluation process. Among other purposes, this also enables modeling groups to assess performance changes during the ESM development cycle in the context of the error distribution of the multi-model ensemble. Quantitative model evaluation provided by the PMP can assist modelers in their development priorities. In this paper, we provide an overview of the PMP, including its latest capabilities, and discuss its future direction.

54 ENVIRONMENTAL SCIENCES↗

Navigating the Noise: Bringing Clarity to ML Parameterization Design With O $\boldsymbol{\mathcal{O}}$(100) Ensembles

Abstract Machine‐learning (ML) parameterizations of subgrid processes (here of turbulence, convection, and radiation) may one day replace conventional parameterizations by emulating high‐resolution physics without the cost of explicit simulation. However, uncertainty about the relationship between offline and online performance (i.e., when integrated with a large‐scale general circulation model) hinders their development. Much of this uncertainty stems from limited sampling of the noisy, emergent effects of upstream ML design decisions on downstream online hybrid simulation. Our work rectifies the sampling issue via the construction of a semi‐automated, end‐to‐end pipeline for size ensembles of hybrid simulations, revealing important nuances in how systematic reductions in offline error manifest in changes to online error and online stability. For example, removing dropout and switching from a Mean Squared Error to a Mean Absolute Error loss both reduce offline error, but they have opposite effects on online error and online stability. Other design decisions, like incorporating memory, converting moisture input from specific humidity to relative humidity, using batch normalization, and training on multiple climates do not come with any such compromises. Finally, we show that ensemble sizes of may be necessary to reliably detect causally relevant differences online. By enabling rapid online experimentation at scale, we can empirically settle debates regarding subgrid ML parameterization design that would have otherwise remained unresolved in the noise.

Lin, Jerry [Department of Earth System Sciences Un↗

Python Library for Monte Carlo Simulations with Ab Initio and Machine-Learned Interatomic Potentials

There is a growing need in the simulation community for software that provides a transparent, reproducible, usable, and extensible (TRUE) Monte Carlo (MC) simulation framework employing energies from ab initio methods and machine-learning interatomic potentials (MLIPs). We introduce a Python library (ASE-MC) that adds Monte Carlo functionality to the Atomic Simulation Environment (ASE) package. Now, we can combine the powerful tools used to build systems and perform ab initio and MLIP in ASE with MC simulation algorithms to sample the configurational space with a concise Python script. After presenting the design philosophy, we demonstrate the flexibility of our approach using selected examples. These example simulations include liquid water described with a message-passing MLIP in the canonical and isothermal–isobaric ensembles, sampling the characteristic dihedral angle of biphenyl and comparing an MLIP to first-principles calculations, and a grand canonical Monte Carlo simulation of ammonia adsorption on Pt(111). These examples showcase the main features of the software, which include flexibility in the choice of ab initio or MLIP engine, ab initio or MLIP grand canonical MC with cavity bias insertions and deletions, the ability to add custom MC moves to the move set, and how users can condense complex MC workflows into a single Python script. Finally, this library serves as a framework for reproducible Monte Carlo simulations, facilitating easy reproduction of the work and application to new systems.

97 MATHEMATICS AND COMPUTING↗

Daily evapotranspiration changes during heatwaves at 32 NEON sites, 2019-2021

This dataset provides partitioned evapotranspiration (ET, the combined loss of water from soil and plant surfaces) anomalies during heatwave events—soil evaporation (E) and transpiration (T)—for 268 heatwave events across 32 National Ecological Observatory Network (NEON) flux sites in the contiguous United States from 2019–2021. Using an ensemble of four high-frequency turbulence methods (Flux-variance Similarity, Conditional Eddy Covariance [CEC], CEC with Water-Use Efficiency, and Conditional Eddy Accumulation; see Zahn and Bou-Zeid 2024), half-hourly transpiration-to-evapotranspiration (T/ET) ratios were derived from 20 hertz (Hz, cycles per second) eddy covariance measurements of carbon dioxide (CO₂) and water vapor (H₂O) concentrations. The dataset spans six vegetation types including evergreen and deciduous forests, grasslands, cultivated crops, shrublands, and emergent herbaceous wetlands. Data Package Contents: The dataset includes a single CSV (comma-separated values) file containing daily anomalies (deviations from baseline conditions) for transpiration (Delta_T), evaporation (Delta_E), total evapotranspiration (Delta_ET), and T/ET ratio (Delta_T_ET) during each day of identified heatwave events. The file also includes site codes, dates, heatwave event identifiers, and day-of-heatwave indicators. The CSV file can be opened with spreadsheet software (Microsoft Excel, Google Sheets) or programming environments (Python, R, MATLAB). This resource enables researchers to investigate ecosystem-specific responses to thermal extremes, validate land surface model partitioning of ET fluxes, and examine feedbacks between water cycling and surface energy balance during heatwaves. The dataset is particularly valuable for studies linking vegetation hydraulic strategies to climate resilience, as it captures the divergent responses of shallow-rooted versus deep-rooted ecosystems. Potential applications include improving drought early warning systems, informing irrigation management strategies, and advancing our mechanistic understanding of land-atmosphere interactions under extreme heat conditions.

Day of Heatwave↗

Global burned area increasingly explained by climate change

Fire behaviour is changing in many regions worldwide. However, nonlinear interactions between fire weather, fuel, land use, management and ignitions have impeded formal attribution of global burned area changes. Here, in this work, we demonstrate that climate change increasingly explains regional burned area patterns, using an ensemble of global fire models. The simulations show that climate change increased global burned area by 15.8% (95% confidence interval (CI) [13.1–18.7]) for 2003–2019 and increased the probability of experiencing months with above-average global burned area by 22% (95% CI [18–26]). In contrast, other human forcings contributed to lowering burned area by 19.1% (95% CI [21.9–15.8]) over the same period. Moreover, the contribution of climate change to burned area increased by 0.22% (95% CI [0.22–0.24]) per year globally, with the largest increase in central Australia. Our results highlight the importance of immediate, drastic and sustained GHG emission reductions along with landscape and fire management strategies to stabilize fire impacts on lives, livelihoods and ecosystems.

54 ENVIRONMENTAL SCIENCES↗

High-Resolution Model Intercomparison Project phase 2 (HighResMIP2) towards CMIP7

Abstract. Robust projections and predictions of climate variability and change, particularly at regional scales, rely on the driving processes being represented with fidelity in model simulations. Consequently, the role of enhanced horizontal resolution in improved process representation in all components of the climate system continues to be of great interest. Recent simulations suggest the possibility of significant changes in both large-scale aspects of the ocean and atmospheric circulations and in the regional responses to climate change, as well as improvements in representations of small-scale processes and extremes, when resolution is enhanced. The first phase of the High-Resolution Model Intercomparison Project (HighResMIP1) was successful at producing a baseline multi-model assessment of global simulations with model grid spacings of 25–50 km in the atmosphere and 10–25 km in the ocean, a significant increase when compared to models with standard resolutions on the order of 1° that are typically used as part of the Coupled Model Intercomparison Project (CMIP) experiments. In addition to over 250 peer-reviewed manuscripts using the published HighResMIP1 datasets, the results were widely cited in the Intergovernmental Panel on Climate Change report and were the basis of a variety of derived datasets, including tracked cyclones (both tropical and extratropical), river discharge, storm surge, and impact studies. There were also suggestions from the few ocean eddy-rich coupled simulations that aspects of climate variability and change might be significantly influenced by improved process representation in such models. The compromises that HighResMIP1 made should now be revisited, given the recent major advances in modelling and computing resources. Aspects that will be reconsidered include experimental design and simulation length, complexity, and resolution. In addition, larger ensemble sizes and a wider range of future scenarios would enhance the applicability of HighResMIP. Therefore, we propose the High-Resolution Model Intercomparison Project phase 2 (HighResMIP2) to improve and extend the previous work, to address new science questions, and to further advance our understanding of the role of horizontal resolution (and hence process representation) in state-of-the-art climate simulations. With further increases in high-performance computing resources and modelling advances, along with the ability to take full advantage of these computational resources, an enhanced investigation of the drivers and consequences of variability and change in both large- and synoptic-scale weather and climate is now possible. With the arrival of global cloud-resolving models (currently run for relatively short timescales), there is also an opportunity to improve links between such models and more traditional CMIP models, with HighResMIP providing a bridge to link understanding between these domains. HighResMIP also aims to link to other CMIP projects and international efforts such as the World Climate Research Program lighthouse activities and various digital twin initiatives. It also has the potential to be used as training and validation data for the fast-evolving machine learning climate models.

54 ENVIRONMENTAL SCIENCES↗

Improving Trustworthiness of Data-Driven Power Grid Contingency Analysis With Bayesian Residual Graph Neural Networks

The evolving energy landscape requires novel tools to efficiently perform contingency analysis and reliability assessment of power grids, potentially in real-time. The high computational cost of traditional power flow solvers limits their applicability in practice. Machine learning (ML) surrogates such as deep neural networks (NNs) accelerate power flow solvers computations, enabling high-order contingency analysis and real-time decision-making by learning highly nonlinear functions and integrating grid topology via graph architectures. However, (graph) NNs lack predictive power away from training data and do not provide predictive confidence estimates. Here, we present a Bayesian residual graph NN that integrates knowledge from low-fidelity data via residual training and embeds granular quantification of uncertainties, improving trustworthiness critical for high-consequence decision-making. Applying Bayesian concepts to NNs is challenging due to the high-dimensionality of both the parameter space, complicating derivation of a meaningful prior, and the output space in large grid systems, requiring enhanced techniques to assess the predicted high-dimensional uncertainties. Our contributions include: (1) Deriving a prior for fully connected and graph NNs that leverages low-fidelity data to guide mean predictions and appropriately control prior predictive uncertainty. (2) Integrating this prior within an ensembling with anchoring scheme for efficient approximate posterior inference. (3) Deriving enhanced metrics to assess accuracy of both the mean and uncertainty predictions in high dimensions, appropriately accounting for correlations propagated through graph layers. The resulting Bayesian residual graph NN is tested on a contingency analysis task for 14-bus and 118-bus grids.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

Separate Surface and Bulk Topological Anderson Localization Transitions in Disordered Axion Insulators

In topological phases of matter for which the bulk and boundary support distinct electronic gaps, there exists the possibility of decoupled mobility gaps in the presence of disorder. This is in analogy with the well-studied problem of realizing separate or concomitant bulk-boundary criticality in conventional Landau theory. Using a three-dimensional axion insulator having clean, gapped surfaces with 𝑒 2 /2⁢ℎ quantized Hall conductance, we show that the bulk and surface mobility gap evolve differently in the presence of disorder. The decoupling of the bulk and surface topology yields a regime that realizes a two-dimensional, unquantized anomalous Hall metal in the Gaussian unitary ensemble on each surface, which shares some spectral and response properties akin to the surface states of a conventional 3D topological insulator. The generality of these results, as well as extensions to other insulators and superconductors, is discussed.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Mountain Basin Controls on the Snow-to-Streamflow Signal: An AIC-Weighted Multiple Linear Regression Framework

A regression-based analysis quantifies how basin characteristics modulate the snow-to-streamflow signal. First, we use the ERA5-Land reanalysis gridded product (European Centre for Medium Range Weather Forecasts reanalysis 5 -Land component) for 4,655 hydrologic unit code - 10 (HUC10) mountain basins across the western United States (US) for water years 1987–2024. Linear regressions are performed for peak snow water equivalent (SWE) and annual streamflow for each mountain basin. Models use ordinary least squares in Python’s statsmodels package. After which, an Akaike Information Criterion (AIC)–weighted ensemble multiple linear regression (MLR) framework with 47 watershed traits is used to predict the linear regression coefficient of determination (r-squared) defining the ability of peak SWE to predict annual streamflow across all mountain basin. Predictor sets are constrained to avoid multicollinearity by excluding models with variance inflation factors (VIF) greater than 5. Mountain basin traits included in the MLR include seasonal climate, topography, vegetation type and structure, and bedrock geology. Accepted models are considered if their AIC is within 2.0 of the model with the minimum AIC, or best model. To compare predictor influence across acceptable models, we computed standardized regression coefficients. To evaluate structural redundancy among models, we constructed binary inclusion vectors for each acceptable model, denoting whether a predictor was present (1) or absent (0). Core predictor variables are defined as occurring in at least 67% of the acceptable models. For this regional analysis, only one model was found acceptable, with higher snow-to-streamflow translation (higher r-squared) occurring in colder mountain basins with higher relative winter precipitation, more snow accumulation and a lower fraction of annual precipitation that falls in the spring and summer. The second component of the data package uses previously published, high-resolution output from an integrated hydrological model of the East River watershed using the U.S. Geological Survey Groundwater and Surface water Flow model (GSFLOW, doi:10.15485/1998576). East River MLR expands upon the approach described above to explore the response of five streamflow metrics—annual streamflow, runoff efficiency, 7-day minimum flow, low-flow duration, and non-perennial stream fraction to snow system indicators including peak SWE, snow-covered area, snow disappearance date, and the fraction of basin area characterized by low-to-no snow, as well as seasonal precipitation and temperature, and annual hydrologic variables representing soil moisture, evapotranspiration (ET), the partitioning of incoming precipitation to evapotranspiration (ET/P), groundwater storage, and groundwater inflow to streams. MLR was done on all water years (P0: 1987-2024) and for each period as determined in the split analysis using pooled regression techniques (P1: 1987-2011 and P2: 2012-2024) to evaluate shifting predictor variable emphasis on streamflow generation. Results indicate that since 2012, peak SWE has lost statistical strength in its prediction of annual streamflow and runoff efficiency, and the indirect influence of spring temperature has emerged as critically important. Low-flow metrics remain largely influenced by soil moisture, vegetation water use and groundwater inflows with summer precipitation becoming a direct influence on minimum summer flow. Together, these data and Python-based analysis tools provide a framework for identifying the key watershed characteristics that control how streamflow responds to snow from year to year. The package also helps quantify uncertainty in statistical models and assess how snow–streamflow relationships vary across regions and over time. This dataset contains comma-separated values files (.csv), text files (.txt), python code files (.py), figure files (.png), and shapefiles (.cpg, .dbf, .prj, .sbn, .sbx, .shp, .xml). Further details on file contents and MLR execution can be found in the readme file and the FLMD files. Work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

54 ENVIRONMENTAL SCIENCES↗

Determination of confinement regime boundaries via separatrix parameters on Alcator C-Mod based on a model for interchange-drift-Alfvén turbulence

The separatrix operational space (SepOS) model (Eich and Manz 2021 Nucl. Fusion 61 086017) is shown to predict the L–H transition, the L-mode density limit, and the ideal magnetohydrodynamic ballooning limit in terms of separatrix parameters for a wide range of Alcator C-Mod plasmas. The model is tested using Thomson scattering measurements across a wide range of operating conditions on C-Mod, spanning $\overline{n}_{\mathrm{e}}$ = 0.3–5.5 $\times 10^{20}\,$m −3 , $B_{\mathrm{t}} = 2.5$–8.0 T, and $B_{\mathrm{p}}$= 0.1–1.2 T. An empirical regression for the electron pressure gradient scale length, $\lambda_{{p}_{\mathrm{e}}}$, against a turbulence control parameter, $\alpha_{\mathrm{t}}$, and the poloidal fluid gyroradius, $\rho_{\mathrm{s,p}}$, is constructed for H-modes and found to require positive exponents for both regression parameters, indicating turbulence widening of near-scrape-off layer widths at high $\alpha_{\mathrm{t}}$ and an inverse scaling with $B_{\mathrm{p}}$, consistent with results on ASDEX Upgrade. The SepOS model is also tested in the unfavorable drift direction and found to apply well to all three boundaries, including the L–H transition as long as a correction to the Reynolds energy transfer term, $\alpha_\mathrm{RS} \lt 1$ is applied. I-modes typically exist in the unfavorable drift direction for values of $\alpha_{\mathrm{t}} \lesssim 0.3$. Finally, an experiment studying the transition between the Type-I ELMy and EDA H-mode is analyzed using the same framework. It is found that a recently identified boundary $\alpha_{\mathrm{t}} = 0.55$ at excludes most EDA H-modes but that the balance of wavenumbers responsible for the L-mode density limit, namely $k_\mathrm{EM} = k_\mathrm{RBM}$, may better describe the transition on C-Mod. The ensemble of boundaries validated and explored is then applied to project regime access and limit avoidance for the SPARC primary reference discharge parameters.

ELM suppression↗

Off-Equilibrium Reactivity of Boron-Enriched Metal Diboride Surfaces in Electroreduction Conditions

Boron-based materials, featuring B-dependent reactivity and diverse phases, are emerging as promising catalyst systems. However, the catalytic mechanism on many borides remains poorly understood due to complex surface reconstructions under reaction conditions. Here, we investigate the MoB 2 surface in conditions of hydrogen evolution reaction in acidic media, using grand canonical global optimization, grand canonical density functional theory, ab initio molecular dynamics, free energy surface sampling, and an analytical model for electrochemical barrier evaluation. We propose a boron-enrichment strategy to tune the surface reactivity of the hexagonal face of MoB 2 . We reveal the dynamic nature of the B-enriched surface under H coverage and kinetic trapping of the system in the metastable regime with an extensive examination of the deactivation pathways. The metastable center B site on B-enriched surfaces, featuring buckled-up configuration and a usual relaxation effect, is found to be highly active toward HER via the Volmer–Heyrovsky mechanism. In conclusion, this work demonstrates how off-equilibrium behaviors can arise from the interplay between adsorbate coverage and surface reconstruction on a seemingly simple surface, and we present a theoretical framework and computational workflows to address these behaviors, along with other realistic complexities, in kinetics simulations.

Adsorption↗

Maximum Entropy Principle in Deep Thermalization and in Hilbert-Space Ergodicity

We report universal statistical properties displayed by ensembles of pure states that naturally emerge in quantum many-body systems. Specifically, two classes of state ensembles are considered: those formed by (i) the temporal trajectory of a quantum state under unitary evolution or (ii) the quantum states of small subsystems obtained by partial, local projective measurements performed on their complements. These cases, respectively, exemplify the phenomena of “Hilbert-space ergodicity” and “deep thermalization.” In both cases, the resultant ensembles are defined by a simple principle: The distributions of pure states have maximum entropy, subject to constraints such as energy conservation, and effective constraints imposed by thermalization. We present and numerically verify quantifiable signatures of this principle by deriving explicit formulas for all statistical moments of the ensembles, proving the necessary and sufficient conditions for such universality under widely accepted assumptions, and describing their measurable consequences in experiments. We further discuss information-theoretic implications of the universality: Our ensembles have maximal information content while being maximally difficult to interrogate, establishing that generic quantum state ensembles that occur in nature hide (scramble) information as strongly as possible. Our results generalize the notions of Hilbert-space ergodicity to time-independent Hamiltonian dynamics and deep thermalization from infinite to finite effective temperature. Our work presents new perspectives to characterize and understand universal behaviors of quantum dynamics using statistical and information-theoretic tools.

Eigenstate thermalization↗

Advancing Organized Convection Representation in the Unified Model: Implementing and Enhancing Multiscale Coherent Structure Parameterization

To address the effect of stratiform latent heating on meso- to large-scale circulations, an enhanced implementation of the Multiscale Coherent Structure Parameterization (MCSP) is developed for the Met Office Unified Model. MCSP represents the top-heavy stratiform latent heating from under-resolved organized convection in general circulation models. We couple the MCSP with a mass-flux convection scheme (CoMorph-A) to improve storm lifecycle continuity. The improved MCSP trigger is specifically designed for mixed-phase deep convective cloud, combined with a background vertical wind shear, both known to be crucial for stratiform development. We also test a cloud top temperature dependent convective-stratiform heating partitioning, in contrast to the earlier fixed partitioning. Assessments from ensemble weather forecasts and decadal simulations demonstrate that MCSP directly reduces cloud deepening and precipitation areas by moderating mesoscale circulations. Indirectly, it amends tropical precipitation biases, notably correcting dry and wet biases over India and the Indian Ocean, respectively. Remarkably, the scheme outperforms a climate model ensemble by improving seasonal precipitation cycle predictions in these regions. The scheme also improves Madden-Julian Oscillation (MJO) spectra, achieving better alignment with observational and reanalysis data by intensifying the simulated MJO over the Indian Ocean during phases 4 to 5. However, the scheme increases precipitation overestimation over the Western Pacific. Shifting from fixed to temperature-dependent convective-stratiform partitioning reduces the Pacific precipitation overestimation and further improves the seasonal cycle in India. Spatially correlated biases highlight the necessity for advances beyond deterministic approaches to align MCSP with environmental conditions.

54 ENVIRONMENTAL SCIENCES↗

A protocol for model intercomparison of impacts of marine cloud brightening climate intervention

A modeling protocol (defined by a series of climate model simulations with specified model output) is introduced. Studies using these simulations are designed to improve the understanding of climate impacts using a strategy for climate intervention (CI) known as marine cloud brightening (MCB) in specific regions; therefore, the protocol is called MCB-REG (where REG stands for region). The model simulations are not intended to assess consequences of a realistic MCB deployment intended to achieve specific climate targets but instead to expose responses to interventions in six regions with pervasive cloud systems that are often considered candidates for such a deployment. A calibration step involving simulations with fixed sea surface temperatures (SSTs) is first used to identify a common forcing, and then coupled simulations with forcing in individual regions and combinations of regions are used to examine climate impacts. Synthetic estimates constructed by superposing responses from simulations with forcing in individual regions are considered a means of approximating the climate impacts produced when MCB interventions are introduced in multiple regions. A few results comparing simulations from three modern climate models (CESM2, E3SMv2, and UKESM1) are used to illustrate the similarities and differences between model behavior and the utility of estimates of MCB climate responses that were synthesized by summing responses introduced in individual regions. Cloud responses to aerosol injections differ substantially between models (CESM2 clouds appear much more susceptible to aerosol emissions than the other models), but patterns in precipitation and surface temperature responses were similar when forcing is imposed with similar amplitudes in the same regions. A previously identified La Niña-like response to forcing introduced in the Southeast Pacific is evident in this study, but the amplitude of the response was shown to markedly differ across the three models. Other common response patterns were also found and are discussed. Forcing in the Southeast Atlantic consistently (across all three models) produces weaker global cooling than that in other regions, and the Southeast Pacific and South Pacific show the strongest cooling. This indicates that the efficiency of a given intervention depends on not only the susceptibility of the clouds to aerosol perturbations, but also the strength of the underlying radiative feedbacks and ocean responses operating within each region. These responses were generally robust across models, but more studies and an examination of responses with ensembles would be beneficial.

54 ENVIRONMENTAL SCIENCES↗

flat10MIP: an emissions-driven experiment to diagnose the climate response to positive, zero and negative CO2 emissions

Abstract. The proportionality between global mean temperature and cumulative emissions of CO2 predicted in Earth system models (ESMs) is the foundation of carbon budgeting frameworks. Deviations from this behavior could impact estimates of required net-zero timings and negative emissions requirements to meet the Paris Agreement climate targets. However, existing ESM diagnostic experiments do not allow for direct estimation of these deviations as a function of defined emissions pathways. Here, we perform a set of climate model diagnostic experiments for the assessment of transient climate response to cumulative CO2 emissions (TCRE), the Zero Emissions Commitment (ZEC), and climate reversibility metrics in an emissions-driven framework. The emissions-driven experiments provide consistent independent variables simplifying simulation, analysis and interpretation, with emissions rates more comparable to recent levels than existing protocols using model-specific compatible emissions from the CMIP DECK 1pctCO2 experiment, where emissions rates tend to increase during the experiment, such that at the time of CO2 doubling in year 70, emissions are much greater than present-day values. A base experiment, “esm-flat10”, has constant emissions of CO2 of 10 GtC per year (near-present-day values), and initial results show that the TCRE estimated in this experiment is about 0.1 K less than that obtained using 1pctCO2. A subset of ESMs exhibit land carbon sinks that saturate during this experiment. A branch experiment, esm-flat10-zec, illustrates that both positive and negative ZEC effects are less pronounced under esm-flat10 than under 1pctCO2 – the magnitude of ZEC50 in ESMs is, on average, reduced by 30 % compared with 1pctCO2 branch experiments. A final experiment, esm-flat10-cdr, assesses climate reversibility under negative emissions, where we find that peak warming may occur before or after net zero and that the asymmetry in temperature at a given level of cumulative emissions between the positive and negative emissions phases is well described by ZEC in most models. Further, we find that existing probabilistic simple climate model (SCM) ensembles tend to overestimate temperature reversibility compared with ESMs, highlighting the need for additional constraints. We propose a set of climate diagnostic indicators to quantify various aspects of climate reversibility. These experiments were suggested as potential candidates in CMIP7 and have since been adopted as “fast track” simulations.

Sanderson, Benjamin M↗

Comparing multi-model ensemble simulations with observations and decadal projections of upper atmospheric variations following the Hunga eruption

The Hunga Tonga–Hunga Ha'apai Model–Observation Comparison (HTHH–MOC) project aims to comprehensively investigate the evolution of volcanic water vapor and sulfur emissions and their subsequent atmospheric impacts and underlying response mechanisms using state-of-the-art global climate models. This study evaluates multi-model ensemble simulations participating in the HTHH–MOC free-run experiment with climate projections for 10 years (2022–2032). Model results are evaluated against satellite observations to assess their ability to reproduce the observed evolution of stratospheric water vapor, aerosols, temperature, and ozone from 2022 to 2024. The participating models accurately capture the observed distribution patterns and associated upper atmospheric responses, providing confidence for their future projections. Model simulations suggest that the Hunga eruption-induced stratospheric water vapor anomaly lasts 4–7 years, with a water vapor e-folding time of 31–43 months. This prolonged water vapor perturbation leads to significant stratospheric and mesospheric cooling, resulting in significant ozone loss in the upper stratosphere and lower mesosphere for 7–10 years. Comparisons between simulations with both SO 2 and H 2 O emissions and those with H 2 O-only emissions indicate that the pronounced dipole response with upper-stratospheric cooling and lower-stratospheric warming is driven by the combined effects of SO 2 and H 2 O injections. These results highlight the prolonged atmospheric impacts of the Hunga eruption and the potential critical role of stratospheric water vapor in modulating long-term atmospheric chemistry and dynamics.

Zhuo, Zhihong [Univ. of Quebec, Montreal, QC (Cana↗

Radiative, Hydrologic, and Circulation Responses to Warming in Cess‐Potter Simulations Using the Global 3.25‐km SCREAM

Using the global 3.25-km Simple Cloud Resolving E3SM Atmosphere Model (SCREAM 3 km), a pair of 13-month Cess-Potter simulations are performed to quantify the radiative feedbacks and the hydrologic and circulation responses to warming. Large-scale aspects of SCREAM 3 km's top-of-atmosphere radiative fluxes, precipitation rates, and circulations are in good agreement with observations and reanalysis, with notable differences, including a drier lower free-troposphere in the Tropics, reduced precipitation and humidity over the Tropical West Pacific, and poleward shifted Southern Hemisphere midlatitude jet. In response to warming, SCREAM 3 km predicts a total radiative feedback within the top 15% of the CMIP5 and CMIP6 models, which puts it substantially higher than the feedback reported by other kilometer-scale models. SCREAM 3 km's high radiative feedback stems from a strongly positive shortwave cloud feedback, most prominent over the mid- and high-latitudes. SCREAM 3 km's high precipitation response also puts it among the highest of CMIP models, whereas its circulation response is within the spread of CMIP models. An ensemble of five perturbed initial condition Cess-Potter simulations with a 12 km version of SCREAM (SCREAM 12 km) is performed to characterize uncertainty and resolution sensitivity. It suggests that the uncertainty from analyzing a pair of 1-year simulations is small compared to the inter-model spread in feedbacks and precipitation response. SCREAM 12 km also produces a strong precipitation response to warming but a much lower cloud feedback and total radiative feedback. The results from these experiments suggest that the spread in climate feedbacks will likely persist in the next generation of kilometer-scale models.

54 ENVIRONMENTAL SCIENCES↗

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE↗