Search NASA⌕ Search

SEARCH · Search NASA

Results for “Exascale”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Implicit Thermochemical Nonequilibrium Compressible Flow Simulations on Unstructured Grids Using GPUs

As next-generation exascale-class systems arrive, existing software must be updated accordingly to effectively utilize these systems. For high concurrency and energy efficiency, many of these systems utilize GPU architectures. In this work, we present a CUDA C++ implementation of FUN3D's thermochemical nonequilibrium capability for turbulent flows. Efficiency is demonstrated at scale using the Summit system at the Oak Ridge Leadership Computing Facility which is representative of future exascale systems. This work enables faster, higher fidelity, and scale-resolving simulations of thermochemical nonequilibrium flows including reentry, hypersonics, and combustion.

CFD, GPU, HPC, Hypersonics, Chemistry↗

E3SM‐Arctic: Regionally Refined Coupled Model for Advanced Understanding of Arctic Systems Interactions

Earth system models are essential tools for climate projections, but coarse resolutions limit regional accuracy, especially in the Arctic. Regionally refined meshes (RRMs) enhance resolution in key areas while maintaining computational efficiency. This paper provides an overview of the United States (U.S.) Department of Energy's (DOE's) Energy Exascale Earth System Model version 2.1 with an Arctic RRM, hereafter referred to as E3SMv2.1-Arctic, for the atmosphere (25 km), land (25 km), and ocean/ice (10 km) components. We evaluate the atmospheric component and its interactions with land, ocean, and cryosphere by comparing the RRM (E3SM2.1-Arctic) historical simulations (1950–2014) with the uniform low-resolution (LR) counterpart, reanalysis products, and observational data sets. The RRM generally reduces biases in the LR model, improving simulations of Arctic large-scale mean fields, such as precipitation, atmospheric circulation, clouds, atmospheric river frequency, and sea ice thickness. However, it introduces a seasonally dependent surface air temperature bias, reducing the LR cold bias in summer but enhancing the LR warm bias in winter, which contributes to the underestimated winter sea ice area and volume. Radiative feedback analysis shows similar climate feedback strengths in both model configurations, with the RRM exhibiting a more positive surface albedo feedback and contributing to a stronger surface warming than LR. These findings underscore the importance of high-resolution modeling for advancing our understanding of Arctic climate changes and their broader global impacts, although some persistent biases appear to be independent of model resolution at 10–100 km scales.

Energy Exascale Earth System Model (E3SM)↗

Large-scale Multiphysics Simulations of Small Modular Reactors Operating in Natural Circulation

Thanks to the advancements in high-performance computing, advanced modeling and simulation have become crucial in driving the development and deployment of next-generation nuclear reactors, such as small modular reactors (SMRs). SMRs offer the promise of cost-effective baseload electricity production and improved safety, while addressing some of the challenges associated with large reactor designs, such as high capital costs and extended construction timelines. As part of the Exascale Computing Project, the large-scale multiphysics simulation of an entire SMR primary system has been achieved by combining computational fluid dynamics and neutronics. In addition to the successful demonstration of full-core SMR simulations, the current study integrated the impact of natural circulation into the system. Natural circulation is the primary mechanism driving coolant circulation in SMRs. The mass flow rate in the core depends on the core power, and a numerical model has been developed to predict it. The pressure drop caused by the helical coil steam generator was also accounted for by developing a pressure drop correlation based on high-fidelity large eddy simulation results, further improving prediction accuracy. In conclusion, the results of the study demonstrate that the implemented natural circulation model is effective in predicting the responses of SMR full-core multiphysics simulations.

ECP↗

Providing a Flexible and Comprehensive Software Stack Via Spack, an Extreme-Scale Scientific Software Stack, and Software Development Kits

To manage the complex demands of modern high-performance computing (HPC), software applications increasingly depend on software developed by other teams, often at other institutions. An HPC software ecosystem approach is required to support dependencies on third-party scientific software. An ecosystem approach provides layers of activity above the individual software product level that promote interoperability, quality improvement, porting, testing, and deployment. The U.S. Exascale Computing Project (ECP) developed its HPC software ecosystem using a three-pronged approach. First, the ECP adopted and invested in Spack, a package manager designed to handle complex HPC package dependencies. Second, the ECP created the Extreme Scale Scientific Software Stack, an effort that supports developing, deploying, and running scientific applications on HPC platforms. Third, the ECP supported software product communities, or software development kits, to develop and promote best practices, improve software interoperability, and other collaborative efforts. This article describes ECP contributions to HPC software ecosystem challenges.

97 MATHEMATICS AND COMPUTING↗

Data and scripts associated with a manuscript analyzing ELM-FATES parameter sensitivity under pre-fire and postfire scenarios using machine learning

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript “Fire Severity-Dependent Shifts in Vegetation Parameter Sensitivity: A Pre- and Post-Fire Analysis Using ELM-FATES and Explainable AI” submitted to Journal of Advances in Modeling Earth Systems (Zahura et al. 2026). The study examines vegetation physiological parameters controlling pre-fire and post-fire vegetation dynamics. To support this analysis, 73 vegetation parameters in Functionally Assembled Terrestrial Ecosystem Simulator (FATES) (Fisher et al., 2018) , which is coupled with E3SM (Energy Exascale Earth System Model) land model (ELM, ELM-FATES), were perturbed using a Sobol sequence to generate 1,024 ensemble members for two plant functional types: needleleaf evergreen extratropical trees (NEET) and C3 grass. Simulations were conducted for the pre-fire period (2016) and post-fire period (2018–2023). Burn severity was represented by modifying the Nesterov index in FATES to 75,000, 150,000, and 300,000 for low, moderate, and high severity, respectively. A no-fire scenario was also included. Simulations were performed for 16 grid cells in the American River Watershed across different burn severities and plant functional types. XGBoost (eXtreme Gradient Boosting) models were trained using the parameter ensembles and ELM-FATES-simulated outputs, including leaf area index (LAI), gross primary productivity (GPP), aboveground biomass, vegetation evaporation, transpiration, and soil evaporation. Models were trained separately for each year and burn severity, followed by SHAP (SHapley Additive exPlanations) analysis to identify changes in dominant parameters after fire disturbance. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. The data package contains the ELM-FATES simulation data. The scripts and data related to the analysis will be added later. The inputs and outputs from ELM-FATES are inside the “FATES” folder. “FATES_domain_surface” contains the domain and surface netcdfs that were used to run ELM-FATES in the study area. “FATES_parameters” contains the 1024 ensembles that were generated using Sobol sequence. “FATES_outputs” folder contains ELM-FATES simulated variables. All files are .csv and .nc (NetCDF).

Aboveground biomass↗

Profiling and Improving I/O Performance of a Large-Scale Climate Scientific Application

Exascale computing systems are soon to emerge, which will pose great challenges on the huge gap between computing and I/O performance. Many large-scale scientific applications play an important role in our daily life. The huge amounts of data generated by such applications require highly parallel and efficient I/O management policies. In this paper, we adopt a mission-critical scientific application, GEOS-5, as a case to profile and analyze the communication and I/O issues that are preventing applications from fully utilizing the underlying parallel storage systems. Through in-detail architectural and experimental characterization, we observe that current legacy I/O schemes incur significant network communication overheads and are unable to fully parallelize the data access, thus degrading applications' I/O performance and scalability. To address these inefficiencies, we redesign its I/O framework along with a set of parallel I/O techniques to achieve high scalability and performance. Evaluation results on the NASA discover cluster show that our optimization of GEOS-5 with ADIOS has led to significant performance improvements compared to the original GEOS-5 implementation.

Exascale↗

ytopt: Autotuning Scientific Applications for Energy Efficiency at Large Scales

As we enter the exascale computing era, efficiently utilizing power and optimizing the performance of scientific applications under power and energy constraints has become critical and challenging. We propose a low-overhead autotuning framework to autotune performance and energy for various hybrid MPI/OpenMP scientific applications at large scales and to explore the tradeoffs between application runtime and power/energy for energy efficient application execution, then use this framework to autotune four ECP proxy applications—XSBench, AMG, SWFFT, and SW4lite. Our approach uses Bayesian optimization with a Random Forest surrogate model to effectively search parameter spaces with up to 6 million different configurations on two large-scale HPC production systems, Theta at Argonne National Laboratory and Summit at Oak Ridge National Laboratory. The experimental results show that our autotuning framework at large scales has low overhead and achieves good scalability. Using the proposed autotuning framework to identify the best configurations, we achieve up to 91.59% performance improvement, up to 21.2% energy savings, and up to 37.84% EDP (energy delay product) improvement on up to 4096 nodes.

Autotuning↗

CI/CD Efforts for Validation, Verification and Benchmarking OpenMP Implementations

Software developers must adapt to keep up with the changing capabilities of platforms so that they can utilize the power of High-Performance Computers (HPC), including exascale systems. OpenMP, a directive-based parallel programming model, allows developers to include directives to existing C, C++, or Fortran code to allow node level parallelism without compromising performance. This paper describes our CI/CD efforts to provide easy evaluation of the support of OpenMP across different compilers using existing testsuites and benchmark suites on HPC platforms. Our main contributions include (1) the set of a Continuous Integration (CI) and Continuous Development (CD) workflow that captures bugs and provides faster feedback to compiler developers, (2) an evaluation of OpenMP (offloading) implementations supported by AMD, HPE, GNU, LLVM, and Intel, and (3) evaluation of the quality of compilers across different heterogeneous HPC platforms. With the comprehensive testing through the CI/CD workflow, we aim to provide a comprehensive understanding of the current state of OpenMP (offloading) support in different compilers and heterogeneous platforms consisting of CPUs and GPUs from NVIDIA, AMD, and Intel.

Jarmusch, Aaron↗

Flexible User-Defined Domain Decomposition in Kilometer-Scale E3SM Land Model Simulation

The Energy Exascale Earth System Model (E3SM) Land Model (ELM) has been extended to kilometer-scale (km-ELM) resolutions, enabling high-fidelity simulations of terrestrial processes at 1 km x 1 km grid spacing. In ELM, domain decomposition partitions the computational domain across processors, ensuring efficient parallel execution. Currently, round-robin decomposition is applied, providing a straightforward way to distribute computational workload. As ELM continues evolving at the kilometer-scale (km-scale), particularly with integrating lateral flow modeling, decomposition strategies must also account for the increased workload and data movement. This paper introduces a flexible user-defined domain decomposition framework, allowing users to customize domain partitioning based on application requirements. The impact of different decomposition strategies is evaluated across various applications concerning computation, communication, and I/O. Results demonstrate that while 1D partitioning yields superior I/O performance, k-nearest neighbors (KNN) clustering effectively reduces inter-process communication overhead. This study lays the groundwork for scalable partitioning in large-scale land surface simulations, enhancing next-generation Earth system modeling.

Wang, Dali [ORNL] (ORCID:0000000168065108)↗

Anticipating how rain-on-snow events will change through the 21st century: lessons from the 1997 new year’s flood event

The California-Nevada 1997 New Year’s flood was an atmospheric river (AR)-driven rain-on-snow (RoS) event and remains the costliest in their history. The joint occurrence of saturated soils, rainfall, and snowmelt generated inundation throughout northern California-Nevada. Although AR RoS events are projected to occur more frequently with climate change, the warming sensitivity of their flood drivers across scales remains understudied. We leverage the regionally refined mesh capabilities of the Energy Exascale Earth System Model (RRM-E3SM) to recreate the 1997 New Year’s flood with horizontal grid spacings of 3.5 km across California, with forecast lead times of up to 4 days, and across six warming levels ranging from pre-industrial conditions to +3.5° C. We describe the sensitivity of the flood drivers to warming including AR duration and intensity, precipitation phase, intensity and efficiency, snowpack mass and energy changes, and runoff efficiency. Our findings indicate current levels of climate change negligibly influence the flood drivers. At warming levels ≥ 1.7° C, AR hazard potential increases, snowpack nonlinearly decreases, antecedent soil moisture decreases (except where the snowline retreats), and runoff decreases (except in the southern Sierra Nevada where antecedent snowpack persists). Storm total precipitation increases, but at rates below warming-induced increases in saturation-specific humidity. Warming intensifies short-duration, high-intensity rainfall, particularly where snowfall-to-rainfall transitions occur. This study highlights the nonlinear tradeoffs in 21st-century RoS flood hazards with warming and provides water management and infrastructure investment adaptation considerations.

54 ENVIRONMENTAL SCIENCES↗

Scalable training of trustworthy and energy-efficient predictive graph foundation models for atomistic materials modeling: a case study with HydraGNN

We present our work on developing and training scalable, trustworthy, and energy-efficient predictive graph foundation models (GFMs) using HydraGNN, a multi-headed graph convolutional neural network architecture. HydraGNN expands the boundaries of graph neural network (GNN) computations in both training scale and data diversity. It abstracts over message passing algorithms, allowing both reproduction of and comparison across algorithmic innovations that define nearest-neighbor convolution in GNNs. This work discusses a series of optimizations that have allowed scaling up the GFMs training to tens of thousands of GPUs on datasets consisting of hundreds of millions of graphs. Our GFMs use multitask learning (MTL) to simultaneously learn graph-level and node-level properties of atomistic structures, such as energy and atomic forces. Using over 154 million atomistic structures for training, we illustrate the performance of our approach along with the lessons learned on two state-of-the-art US Department of Energy (US-DOE) supercomputers, namely the Perlmutter petascale system at the National Energy Research Scientific Computing Center and the Frontier exascale system at Oak Ridge Leadership Computing Facility. The HydraGNN architecture enables the GFM to achieve near-linear strong scaling performance using more than 2000 GPUs on Perlmutter and 16,000 GPUs on Frontier.

97 MATHEMATICS AND COMPUTING↗

ESM data downscaling: a comparison of super-resolution deep learning models

Abstract Climate projections at fine spatial resolutions are required to conduct accurate risk assessment for critical infrastructure and design adaptation planning. Generating these projections using advanced Earth system models (ESM) requires significant computational resources. To address this issue, various statistical downscaling techniques have been introduced to generate fine-resolution data from coarse-resolution simulations. In this study, we evaluate and compare five deep learning-based downscaling techniques, namely, super-resolution convolutional neural networks, fast super-resolution convolutional neural network ESM, efficient sub-pixel convolutional neural network, enhanced deep residual network (EDRN), and super-resolution generative adversarial network (SRGAN). These techniques are applied to a dataset generated by the Energy Exascale Earth System Model (E3SM), focusing on key surface variables such as surface temperature, shortwave heat flux, and longwave heat flux. Models are trained and validated using paired fine-resolution (0.25 $$^{\circ }$$ ∘ ) and coarse-resolution (1 $$^{\circ }$$ ∘ ) monthly data obtained from a 9-year simulation. Next, blind testing is performed using monthly data obtained from two different years outside of the training and validation set. To evaluate the efficiency of each technique, different statistical metrics are used, including mean squared error (MSE), peak signal-to-noise ratio (PSNR), structural similarity index measure (SSIM), and learned perceptual image patch similarity (LPIPS). The results show that EDRN outperforms other algorithms in terms of PSNR, SSIM, and MSE, but struggles to capture fine-scale features in the data. In contrast, SRGAN, a generative model that uses perceptual loss, excels in capturing fine details at boundaries and internal structures, resulting in lower LPIPS than other methods.

Pawar, Nikhil M. (ORCID:0000000211613289)↗

Evaluating ecosystem water use efficiency and recovery dynamics during flash droughts: insights from observations and model simulations

Flash droughts (FD), rapidly emerging in a warming future, disrupt ecosystems, agriculture, and water security. Ecosystem water use efficiency (WUE), the ratio of gross primary production (GPP) to actual evapotranspiration (AET), balances carbon assimilation and water loss. FD rapidly disrupts this balance, making WUE critical for assessing plant stress and recovery. Here, this study investigates the dynamics of landscape-scale WUE, and the components of GPP and AET under FD utilizing both observed data from the Missouri Ozark AmeriFlux site (US-MOz) and version 2 of the U.S. Department of Energy’s Earth, Energy, Exascale System Model (E3SM) Land Model (ELMv2). Observations and simulations reveal GPP as dominant for WUE during earlier FD events (2005, 2007, 2012), shifting to AET in recent events (2014, 2018). This agreement indicates that the ELM can capture the shifting dynamics of GPP and AET in regulating WUE under FD conditions. However, the ELM systematically underestimates both GPP and AET and does so in a manner that does not preserve their ratio. As a result, WUE is also underestimated, suggesting that GPP is more strongly underestimated than AET. Furthermore, the ELM also underestimates the speed of GPP recovery, producing an artificially prolonged GPP recovery time following FD events. Observed environmental drivers such as vapor pressure deficit (VPD), soil moisture (SM), and predawn leaf water potential (PLWP) effectively predict WUE, but ELM primarily highlights SM, underestimating VPD’s role. This study demonstrates that relying solely on soil moisture fails to capture the rapid hydraulic recovery observed in PLWP, underscoring the necessity of integrating plant hydraulics into land surface models to improve flash drought predictability.

Evapotranspiration↗

In situ multi-tier auto-ignition detection applied to dual-fuel combustion simulations

Here we use an anomaly detection methodology that is centered on analyzing fourth-order joint moments (co-kurtosis), particularly focusing on its application in auto-ignition of combustion problems with large numbers of species. Unsupervised anomaly detection is challenging to generalize across problem types and domains. A recent technique, centered on analyzing information in the fourth-order joint moment co-kurtosis, has shown promise, especially for high-dimensional scientific data. In this work we present developments to the co-kurtosis based anomaly detection method needed to make it effective and scalable for large-scale distributed scientific data, such as those generated by massively parallel simulations. An in situ co-kurtosis algorithm is employed as the anomaly detection method for identifying ignition kernels in simulations of turbulent combustion. Here, we extend an existing methodology which identifies regions of the domain where anomalies are present, and add another tier of anomaly detection where the individual samples contributing to the anomaly are identified. We apply this algorithm on-the-fly to a variety of turbulent reacting flow problems and compare it to the widely used (but significantly more expensive) chemical explosive mode analysis (CEMA). We demonstrate the ability of the method to detect and identify the onset of low and high temperature ignition which can be used for computational steering, as chemical and combustion anomalies occur intermittently at spatio-temporal locations unknown a priori. Finally, we apply our lightweight in situ algorithm to an exascale high-fidelity simulation with a total of 2.4 Trillion degrees of freedom, performed using an adaptive mesh refinement solver. Furthermore, through a scalability analysis, we show that the relative computational cost of this in-situ anomaly detection algorithm compared to an iteration of the reacting flow solver is negligible.

97 MATHEMATICS AND COMPUTING↗

A Comprehensive Chemistry Evaluation and Diagnostics Package for E3SM – ChemDyg Version 1.1.0

The Chemistry Evaluation and Diagnostics Package (ChemDyg) is an open-source tool designed for the Energy Exascale Earth System Model (E3SM) developed by the U.S. Department of Energy. ChemDyg facilitates routine evaluation, tailored development, and in-depth analysis of atmospheric chemistry through its modular architecture, allowing users to compare model outputs with observational data. Version 1.1.0 introduces a robust set of diagnostic capabilities, including climatology, time evolution of key tracers, diurnal and annual cycle analyses, and extensive budget diagnostics. These features help identify model discrepancies and enhance the representation of atmospheric chemistry in E3SM. Each self-contained diagnostic set includes dedicated scripts and documentation for ease of use. The interactive HTML output improves data accessibility, accelerating chemistry model development. Additionally, ChemDyg's flexible framework allows for customization, enabling users to create unique diagnostic sets for specific scientific contributions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Software stewardship and advancement of a high-performance computing scientific application: QMCPACK

Here, we provide an overview of the software engineering efforts and their impact in QMCPACK, a production-level ab-initio Quantum Monte Carlo open-source code targeting high-performance computing (HPC) systems. Aspects included are: (i) strategic expansion of continuous integration (CI) targeting CPUs, using GitHub Actions own runners, and NVIDIA and AMD GPUs used in pre-exascale systems, (ii) incremental reduction of memory leaks using sanitizers, (iii) incorporation of Docker containers for CI and reproducibility, and (iv) refactoring efforts to improve maintainability, testing coverage, and memory lifetime management. We quantify the value of these improvements by providing metrics to illustrate the shift towards a predictive, rather than reactive, maintenance approach. Our goal, in documenting the impact of these efforts on QMCPACK, is to contribute to the body of knowledge on the importance of research software engineering (RSE) for the stewardship and advancement of community HPC codes to enable scientific discovery at scale.

97 MATHEMATICS AND COMPUTING↗

Quantifying the long-term changes of terrestrial water storage and their driving factors

Global warming is expected to cause changes in terrestrial water storage (TWS) across the land surface, with widespread impacts on ecosystems and society. Although extensive research has been performed to analyze TWS changes and possible drivers during the post-2000 period, longer-term evolution of TWS and associated environmental forcings remain relatively unexplored. In this study, we evaluated the performance of the Energy Exascale Earth System model (E3SM) land model ELM version 1 (ELM v1) in simulating global TWS, and used factorial simulations of ELMv1 to quantify global TWS changes and their drivers during 1948–2012. We found that ELM’s agreed best with existing satellites and reconstruction datasets in temperate regions unaffected by irrigation. Biome- and climate zone-averaged TWS mainly increased at rates between 0 and 10 mm/year over 1948–2012, but the second half of that period saw smaller positive trends than the first half or even negative trends. Climate change explained >80 % of the TWS trends across most biomes and climate zones, followed by land use and land cover change. The physiological and phenological effects of CO2 primarily induced noticeable TWS trends in the more humid biomes and climate zones across different latitudes. In contrast, nitrogen deposition and aerosol deposition generally had smaller and negative impacts across the biomes and climate regions. Among the meteorological drivers analyzed, the long-term average imbalance between precipitation (P), evapotranspiration (E), and runoff (Q) contributed >50 % of the TWS trends in most biomes and climate zones, with nonlinearity being induced by spatially heterogenous changes in E/P and Q/P ratios. The accumulated detrended anomalies in P, E, and Q also often contributed substantially, while the trends difference between P, E, and Q contributed little. Together, these findings unveiled an intensification of the global TWS and its diverse patterns of climate change and different non-withdrawal human-induced alterations, contributing to a more comprehensive understanding and projection of the global water cycle.

54 ENVIRONMENTAL SCIENCES↗

Storylines for the 1997 New Year’s Flood: The role of watershed antecedent conditions and future warming in shaping discharge in the Truckee River watershed

The 1997 New Year’s flood was among the most devastating floods in the Truckee River watershed located in western Nevada. This event resulted from complex interactions of flood drivers, such as extreme precipitation, wet antecedent watershed conditions, warm temperatures and rapid snowmelt. We leveraged simulated forcings from the regionally refined mesh capabilities of the Energy Exascale Earth System Model (RRM-E3SM) and a process-based hydrological model to recreate the 1997 New Year’s flood for the Truckee River watershed across four climate warming levels ranging from the current temperatures to + 4° C. For each scenario, we conducted ensemble simulations with the same forcing but with 100 different seasonal watershed antecedent conditions, which were randomly sampled from long-term hydrological simulations. The results show that the 1997 New Year’s flood can be reproduced or exceeded consistently only when the antecedent watershed conditions are wet, specifically when streamflows are above the 75th percentile of the climatological value. There is negligible change in ensemble mean peakflows for Truckee River near Reno; however, there are increases of 18% and 14% under the warming levels of + 3° C and + 4° C, respectively. The increases in peakflows under future climate warming are attributed to wetter antecedent watershed conditions and enhanced snowmelt. Furthermore, the largest increases in peakflows occur at small, high-elevation headwater basins along the Sierra Nevada crest. This study highlights that changes in extreme flood events will result from the complex interplay of multiple flood drivers. It also demonstrates the potential of storyline approaches to analyze future realizations of these extreme events under different climate scenarios.

Climate change impact study↗