Search NASA⌕ Search

SEARCH · Search NASA

Results for “execution”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Parallelized real-time physics codes for plasma control on DIII-D

A real-time safe multi-threading library was developed on the DIII-D plasma control system to optimize the real-time TORBEAM and real-time STRIDE physics codes. These physics codes are crucial for future fusion power plant operation as they provide information about electron cyclotron wave propagation and heating as well as inform about ideal plasma stability limits. The real-time TORBEAM code executed consistently in under 20 ms while the real-time STRIDE code computes in 100 ms. The multi-threading library developed in this work can be applied to other real-time physics-based codes that will be crucial for the next generation of fusion devices.

DIII-D↗

A survey on checkpointing strategies: Should we always checkpoint à la Young/Daly?

The Young/Daly formula provides an approximation of the optimal checkpointing period for a parallel application executing on a supercomputing platform. It was originally designed to handle fail-stop errors for preemptible tightly-coupled applications, but has been extended to other application and resilience frameworks. Here, we provide some background and survey various scenarios to assess the usefulness and limitations of the formula, both for preemptible applications and workflow applications represented as a graph of tasks. We also discuss scenarios with uncertainties, and extend the study to silent errors. We exhibit cases where the optimal period is of a different order than that dictated by the Young/Daly formula, and finally we explain how checkpointing can be further combined with replication.

97 MATHEMATICS AND COMPUTING↗

Analyzing inference workloads for spatiotemporal modeling

Ensuring power grid resiliency, forecasting climate conditions, and optimization of transportation infrastructure are some of the many application areas where data is collected in both space and time. Spatiotemporal modeling is about modeling those patterns for forecasting future trends and carrying out critical decision-making by leveraging machine learning/deep learning. Once trained offline, field deployment of trained models for near real-time inference could be challenging because performance can vary significantly depending on the environment, available compute resources and tolerance to ambiguity in results. Users deploying spatiotemporal models for solving complex problems can benefit from analytical studies considering a plethora of system adaptations to understand the associated performance-quality trade-offs. To facilitate the co-design of next-generation hardware architectures for field deployment of trained models, it is critical to characterize the workloads of these deep learning (DL) applications during inference and assess their computational patterns at different levels of the execution stack. In this paper, we develop several variants of deep learning applications that use spatiotemporal data from dynamical systems. We study the associated computational patterns for inference workloads at different levels, considering relevant models (Long short-term Memory, Convolutional Neural Network and Spatio-Temporal Graph Convolution Network), DL frameworks (Tensorflow and PyTorch), precision (FP16, FP32, AMP, INT16 and INT8), inference runtime (ONNX and AI Template), post-training quantization (TensorRT) and platforms (Nvidia DGX A100 and Sambanova SN10 RDU). Overall, our findings indicate that although there is potential in mixed-precision models and post-training quantization for spatiotemporal modeling, extracting efficiency from contemporary GPU systems might be challenging. Instead, co-designing custom accelerators by leveraging optimized High Level Synthesis frameworks (such as SODA High-Level Synthesizer for customized FPGA/ASIC targets) can make workload-specific adjustments to enhance the efficiency.

97 MATHEMATICS AND COMPUTING↗

A terminology for scientific workflow systems

The term “scientific workflow” has evolved over the last two decades to encompass a broad range of compositions of interdependent compute tasks and data movements. It has also become an umbrella term for processing in modern scientific applications. Today, many scientific applications can be considered as workflows made of multiple dependent steps, and hundreds of workflow systems have been developed to manage and run these scientific workflows. However, no turnkey solution has emerged from the field to address the diversity of scientific processes and the infrastructure on which they are supposed to be implemented. Instead, new research problems requiring the execution of scientific workflows with some novel feature often lead to the development of an entirely new workflow system. A direct consequence of this situation is that many existing workflow management systems (WMSs) share some salient features, offer similar functionalities, and can manage the same categories of workflows but at the same time also have some distinct capabilities that can be important for specific applications. This situation makes researchers who develop workflows face the complex question of selecting a WMS. This selection can be driven by technical considerations, to find the system that is the most appropriate for their application and for the computing and storage resources available to them, or other factors such as reputation, adoption, strong community support, or long-term sustainability. To address this problem, a group of WMS developers and practitioners joined their efforts to produce a community-based terminology of WMSs. This paper summarizes their findings and introduces this new terminology to characterize WMSs. Furthermore, this terminology is composed of fives axes: workflow structure and characteristics, composition, orchestration, data management, and metadata capture. Each axis comprises several concepts that capture the prominent features of WMSs. Based on this terminology, this paper also presents a classification of 23 existing WMSs according to the proposed axes and terms.

Community-based terminology↗

Bridging paradigms: Designing for HPC-Quantum convergence

Here, this paper presents a comprehensive software stack architecture for integrating quantum computing (QC) capabilities with High-Performance Computing (HPC) environments. While quantum computers show promise as specialized accelerators for scientific computing, their effective integration with classical HPC systems presents significant technical challenges. We propose a hardware-agnostic software framework that supports both current noisy intermediate-scale quantum devices and future fault-tolerant quantum computers, while maintaining compatibility with existing HPC workflows. The architecture includes a quantum gateway interface, standardized APIs for resource management, and robust scheduling mechanisms to handle both simultaneous and interleaved quantum–classical workloads. Key innovations include: (1) a unified resource management system that efficiently coordinates quantum and classical resources, (2) a flexible quantum programming interface that abstracts hardware-specific details, (3) A Quantum Platform Manager API that simplifies the integration of various quantum hardware systems, and (4) a comprehensive tool chain for quantum circuit optimization and execution. We demonstrate our architecture through implementation of quantum–classical algorithms, including the variational quantum linear solver, showcasing the framework’s ability to handle complex hybrid workflows while maximizing resource utilization. This work provides a foundational blueprint for integrating QC capabilities into existing HPC infrastructures, addressing critical challenges in resource management, job scheduling, and efficient data movement between classical and quantum resources.

97 MATHEMATICS AND COMPUTING↗

Approximation of refrigerant thermophysical properties using neural networks to speed up transient thermofluid simulations

Accurate and efficient evaluations of refrigerant thermophysical properties and their partial derivatives are essential for transient simulations of thermofluid systems, where several computations need to be executed at each integration time step. Since the utilization of an Equation of State for retrieving properties based on a pair of independent inputs typically involves numerical iterations in solution procedures, when the input variables differ from the refrigerant state variables employed in dynamic models, a variety of approaches including lookup table interpolation and curve fitting have been developed to explicitly approximate these properties based on the state variables, and consequently eliminate internal iterations. This paper presents an alternative method that exploits derivative-informed neural networks to model refrigerant properties explicitly from inputs of pressure and enthalpy, while ensuring consistent partial derivatives generated by differentiating the neural networks. Computational speed and accuracy of the proposed approach are demonstrated via transient simulations of a discretized heat exchanger model in Modelica, and comparisons against other property evaluation routines. Simulation results indicate that the proposed approach can realize a significant speedup with negligible discrepancies in predicted transients. The method is implemented in an open-source Modelica library.

Ma, Jiacheng↗

Metal hydrides: a historical perspective

Metal hydrides are known for their outstanding performance as materials for hydrogen storage and processing. These materials find applications for short- and long-term energy storage, compression and supply of hydrogen gas, thermal energy storage, as electrodes and electrolytes in rechargeable batteries, for the microstructural optimisation of functional materials, in thin film technologies, as catalysts, getters and in many other uses. After the discovery of the first binary metal hydrides back in the 19th century, their studies covered all possible binary M-H systems and expanded rapidly into the field of ternary hydrides following the recognition of the excellent hydrogen storage performance of LaNi 5 - and TiFe-based materials, which operate efficiently at room temperature and at near-ambient H 2 pressures. This review aims to provide an overview of the early works, as well as selected recent results on various classes of metal hydrides. It also covers the recent activities from the major contributing countries and continents, including USA, Europe, Japan, China and Australia. These studies relate to achieving the hydrogen storage systems goals set by the Department of Energy in the United States which inspired the research activities at the national and international level, through execution of the tasks on hydrogen-based energy storage managed by the International Energy Agency. The review is prepared by international experts in the field and covers the most important past developments and also presents the recent achievements in the field.

08 HYDROGEN↗

Hardware acceleration for HPS algorithms in two and three dimensions

We provide a flexible, open-source framework for hardware acceleration, namely massively-parallel execution on general-purpose graphics processing units (GPUs), applied to the hierarchical Poincaré–Steklov (HPS) family of algorithms for building fast direct solvers for linear elliptic partial differential equations. To take full advantage of the power of hardware acceleration, we propose two variants of HPS algorithms to improve performance on two- and three-dimensional problems. In the two-dimensional setting, we introduce a novel recomputation strategy that minimizes costly data transfers to and from the GPU; in three dimensions, we modify and extend the adaptive discretization technique of Geldermans and Gillman [1] to greatly reduce peak memory usage. We provide an open-source implementation of these methods written in JAX, a high-level accelerated linear algebra package, which allows for the first integration of a high-order fast direct solver with automatic differentiation tools. We conclude with extensive numerical examples showing our methods are fast and accurate on two- and three-dimensional problems.

Fast direct solvers↗

An uncertainty visualization framework for large-scale cardiovascular flow simulations: A case study on aortic stenosis

We present a generalizable uncertainty quantification (UQ) and visualization framework for lattice Boltzmann method simulations of high Reynolds number vascular flows, demonstrated on a patient-specific stenosed aorta. The framework combines EasyVVUQ for parameter sampling with large-eddy simulation turbulence modeling in HemeLB, and executes ensembles on the Frontier exascale supercomputer. Spatially resolved metrics, including entropy and isosurface-crossing probability, are used to map uncertainty in pressure and wall shear stress fields directly onto vascular geometries. Two sources of model variability are examined: inlet peak velocity and the Smagorinsky constant. Inlet velocity variation produces high uncertainty downstream of the stenosis where turbulence develops, while upstream regions remain stable. Smagorinsky constant variation has little effect on the large-scale pressure field but increases WSS uncertainty in localized high-shear regions. In both cases, the stenotic throat manifests low entropy, indicative of robust identification of elevated WSS. By linking quantitative UQ measures to three-dimensional anatomy, the framework improves interpretability over conventional 1D UQ plots and supports clinically relevant decision-making, with broad applicability to vascular flow problems requiring both accuracy and spatial insight.

Hemodynamics↗

Development and validation of a software for simulating γ-γ coincidence emission and detection probabilities

Gamma-gamma coincidence spectrometers have the potential to significantly enhance detection sensitivity for ultra-trace radionuclide measurements. The implementation of these spectrometers, however, is limited by the complexity of acquisition hardware, data processing and quantification. This work reports development of a novel radionuclide quantification software for γ-γ coincidence measurements. For any radionuclide, the software parses the Evaluated Nuclear Structure Data File (ENSDF) database, recursively simulating all possible γ-γ coincidence signatures and their respective emission and detection probabilities. Implemented using Python programming language, the software employs several strategies to boost overall computational performance. Since coincidence-based spectrometers are of notable interest in monitoring compliance for the Comprehensive Nuclear-Test-Ban Treaty (CTBT), the software’s execution was tested for 84 CTBT-relevant radionuclides. To date, the software has been experimentally validated for 15 radionuclides using the Advanced Radionuclide Gamma spectrOmeter (ARGO) at Pacific Northwest National Laboratory, USA (PNNL). Notably, the software can be operated in convergence mode, whereby coincidence detection efficiency’s convergence behavior can help avoid unreliable radionuclide activity estimates. With growing number of coincidence spectrometers worldwide, this paper aims to assist the radiation metrology community in developing similar software for their system.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

RELAP5-3D validation studies based on the High Temperature Test facility

In the spring and summer of 2019, experiments were conducted at the High Temperature Test Facility (HTTF) that form the basis of an upcoming high-temperature gas-cooled reactor (HTGR) thermal hydraulics (T/H) benchmark. HTTF is an integral effects test facility for HTGR T/H modeling validation. This paper presents RELAP5-3D models of two of those experiments: PG-27, a pressurized conduction cooldown (PCC); and PG-29, a depressurized conduction cooldown (DCC). These models used the RELAP5-3D model of HTTF originally developed by Paul Bayless as a starting point. The sensitivity analysis and uncertainty quantification code, RAVEN was used to perform calibration studies for the steady-state portion of PG-27. Here we developed four PG-27 calibrations based on steady-state conditions. These calibrations all used an effective thermal conductivity equal to 36 % of the measured thermal conductivity, but they differed with respect to the frictional pressure drops and radial conduction models. These models all captured the trends in steady-state temperature distributions and transient temperature behavior well. All four calibrations show room for improvement in predicting the transient temperature rise. The smallest error in temperature rise during the transient was a 21 % underprediction, and the largest was a 48 % underprediction. The errors in transient temperature rise are largely a result of a mismatch in power density between the RELAP5-3D model and the experiment due to the location of active heater rods along the boundary between heat structures in the model. The best of these calibrations was applied to PG-29 to model the DCC. Once again, temperatures during the transient were underpredicted but trends in temperature were captured. The RELAP5-3D model captured trends in the data but could not reproduce measured temperatures exactly. This result is not attributed to deficiencies in the experimental data or to RELAP5–3D itself. Rather, this result likely arises due to the some of the assumptions and decisions made when the RELAP5-3D model was first developed, prior to the execution of HTTF experiments. An agreement in prediction of temperature trends but challenges reproducing HTTF temperatures within measurement uncertainty is consistent with previous analyses of HTTF in the literature. Future RELAP5-3D validation activities centered around HTTF may be able to provide greater insight into the code’s capabilities for HTGR modeling with a more finely nodalized model.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Design of a separate effects MiniFuel irradiation experiment investigating microstructure evolution in high burnup UO 2

The microstructural evolution of UO 2 fuel pellets during commercial operation in light water reactors (LWRs) is known to vary significantly across the pellet radius due to spatial variations in local temperature and burnup. The primary obstacle to extending LWR refueling cycles to 24-month intervals is the susceptibility of certain high burnup fuel microstructures to fuel fragmentation, relocation, and dispersal (FFRD) during a loss of coolant accident (LOCA). Although FFRD of the high burnup structure in the rim region of a pellet is well studied, the fine fragmentation that has been observed in a second region, near the midradius of the pellet (termed the “dark zone”) following mock LOCA testing of high burnup commercial fuel rods is less understood. This paper describes the design, analysis, and execution of a separate effects MiniFuel irradiation experiment that aims to identify the specific temperature and burnup regimes under which FFRD-susceptible dark zone microstructures form. The small disc specimens (3 mm diameter by ∼0.3 mm thick) enable more precise control of the relatively uniform temperature and burnup conditions. A total of 42 specimens were fabricated with typical LWR fuel densities (∼96%–98% of theoretical density) and grain sizes (∼12 μm) and are being irradiated over a range of temperatures (600°C–1000°C) and discharge burnups (50–72 MWd/kg-U) that bound the midradius region of high burnup LWR fuel. Fuel specimens with identical 235 U enrichments were inserted in two irradiation locations in the High Flux Isotope Reactor and are currently undergoing irradiation to further evaluate the impact of rate effects (fission rate, time at temperature) on the microstructural evolution. The fuel fabrication and the thermal and neutronic simulations used for designing the experiment are detailed in this paper. A secondary objective of the experiment is to observe fission gas release (FGR) under the various irradiation conditions, and this work provides first-order predictions of FGR from all fuel specimens. The insights gained from these experiments will inform future high burnup core designs that could minimize the formation of susceptible microstructures and ultimately enable 24-month refueling cycles while minimizing the fraction of the fuel susceptible to FFRD.

FFRD↗

The state of the art for neutron irradiation experiments from the perspective of the High Flux Isotope Reactor (HFIR)

Irradiation experiment campaigns are critical to advancing nuclear energy technologies by providing data on material performance under relevant radiation conditions. Successful irradiation experiments require integrated design efforts that balance technical goals with facility constraints. Here, this paper presents an expert-informed overview of irradiation experiment design at the High Flux Isotope Reactor. It addresses the nuclear materials research and irradiation experiment communities to guide them toward developing technically sound, facility-compatible campaigns. The High Flux Isotope Reactor is a multipurpose reactor supporting isotope production, neutron scattering, and materials testing. Its high, steady-state neutron flux is ideal for irradiation experiments, but successful execution demands coordinated thermal, structural, and reactor physics analyses. The paper outlines the complete development workflow from concept definition and design optimization to safety qualification and post-irradiation examination. Standardized capsule platforms are also discussed in terms of flexibility, specimen capacity, and thermal performance. Common failure modes such as unanticipated geometric variations, can impact temperature-dose profiles and compromise data reliability. Therefore, detailed thermal modeling and accurate as-built characterization are essential for meaningful post-irradiation data interpretation. Key recommendations include early engagement all stakeholders, clearly defined design expectations, and alignment of specimen geometries with post-irradiation examination capabilities. This approach reduces design iterations, enhances data quality, and supports more efficient use of irradiation resources. Strategic and well-planned irradiation testing not only improves individual campaign success but also accelerates the deployment of advanced nuclear technologies. By closing critical data gaps and reducing development risks, the nuclear materials community can more effectively contribute to the future of clean, resilient energy systems.

Experiments↗

Comparative life cycle assessment of remote potable water supply for the Department of Defense

The Department of Defense (DOD) and other agencies, including relief organizations, require potable water for remote missions around the globe. As part of recent initiative by the U.S. Federal government through Executive Order 14057, the DOD has been instructed to investigate the sustainability of operations and practices within the context of climate change. One such practice that needs to be addressed is the procurement of potable water, an essential requirement of any remote mission or location. Currently, there are three primary means of procuring potable water at remote locations: bottled water, on-site purification, or tie-in to existing, local infrastructure. The first two operations are often considered the most secure options, but have sustainability concerns. The purpose of this study is to compare the environmental impacts of bottled water procurement versus on-site treatment via a mobile Reverse Osmosis Water Purification Unit (ROWPU), which uses multiple levels of filtration to make potable water from a local source. A cradle-to-gate assessment was developed for both systems to compare different options for potable water supply. An in person inventory was paired with data taken from the Ecoinvent 3.8 database to directly compare the two systems. The two systems are compared on a 5-year timeline to analyze the environmental impact of repeated bottled water transport versus diesel generator-fueled on-site treatment. Across all impact categories, the results indicate that high energy costs of the reverse osmosis process have significantly less impact on the environment than the repetitive transport and procurement of bottled water. The results of the study have important implications for advancing sustainable operations for remote communities or temporary settlements.

54 ENVIRONMENTAL SCIENCES↗

NEML2: An efficient and modular multiphysics constitutive modeling library for hybrid computing environments

This paper presents NEML2, an open-source, high-performance library developed for constitutive material modeling, designed to support the flexible and modular development of models for complex material behavior. Building on the foundational structure of its predecessor, NEML, the NEML2 library introduces significant improvements, including enhanced vectorization, automatic differentiation, and seamless integration with PyTorch, facilitating the application of machine learning techniques in material simulations. NEML2 provides a C++ backend with Python bindings, enabling users to create custom material models that can be executed efficiently on both CPU and GPU platforms. The library also supports coupling with Multiphysics simulation frameworks like MOOSE, making it suitable for realistic simulations involving coupled physical processes. Rigorous quality assurance through unit and regression testing ensures the reliability of results, while the extensible, user-friendly design encourages collaboration and reproducibility across the scientific community. This paper provides an overview of NEML2’s architecture, core features, and applications, highlighting its impact on accelerating material qualification and advancing computational methods in materials science.

GPU↗

EnergyPlus-MCP: A model-context-protocol server for ai-driven building energy modeling

Traditional building energy modeling with the EnergyPlus building performance simulation engine requires domain expertise, programming skills, and intensive manual efforts limiting its effective adoption. This paper introduces EnergyPlus-MCP, the first open-source Model Context Protocol (MCP) server specifically designed for EnergyPlus simulation workflows, establishing a new foundational infrastructure for AI-driven building energy modeling. The MCP server implements a layered architecture with 35 specialized tools spanning model management, editing and analysis, HVAC and other systems configuration inspection, and simulation execution, enabling Large Language Models to interact with EnergyPlus through conversational interfaces. The server addresses critical workflow barriers by automating model validation, streamlining energy efficiency measures modification, and providing intelligent output management with interactive visualization. Through practical demonstrations using a multi-zone building retrofit analysis, we show how the EnergyPlus-MCP server significantly reduces manual efforts while maintaining full simulation rigor. By providing accessible natural language interfaces to sophisticated building energy analysis, this approach enables scalable deployment of simulation expertise across public and private organizations, educational institutions, and research teams, fundamentally transforming traditional building energy modeling practices.

AI↗

HARD: A performance portable radiation hydrodynamics code based on FleCSI framework

Hydrodynamics And Radiation Diffusion (HARD) is an open-source application for high-performance simulations of compressible hydrodynamics with radiation-diffusion coupling. Built on the FleCSI (Bergen et al., 2021 [1]) (Flexible Computational Science Infrastructure) framework, HARD expresses its computational units as tasks whose execution can be orchestrated by multiple back-end runtimes, including Legion (Bauer et al., 2012 [2]), MPI (Forum, 1994 [3]), and HPX (Kaiser et al., 2020 [4]). Node-level parallelism is handled through Kokkos (Edwards et al., 2014 [5]), providing a single-source, portable code base that runs efficiently on laptops, small homogeneous clusters, and the largest heterogeneous supercomputers currently available. To ensure scientific reliability, HARD includes a regression test suite that automatically reproduces canonical verification problems such as the Sod and LeBlanc shock tubes, and the Sedov blast wave, comparing numerical solutions against known analytical results. The project is distributed under an OSI-approved license, hosted on GitHub, and accompanied by reproducible build scripts and continuous integration workflows. This combination of performance portability, verification infrastructure, and community-focused development makes HARD a sustainable platform for advancing radiation hydrodynamics research across multiple domains.

97 MATHEMATICS AND COMPUTING↗

Computational capacity in hydrodynamic real-time hybrid simulation applied to simulate the dynamic response of floating offshore wind turbines

Real-time hybrid simulation (RTHS) mitigates similitude distortions in model-scale tests of floating offshore wind turbines (FOWTs) by coupling physical experiments with numerical models in real time. The coupling requires faster-than-real-time numerical computations to satisfy temporal similitude with the physical experiment, presenting a bottleneck for using more complex numerical models in RTHS. This paper presents a hydrodynamic-RTHS (hydro-RTHS) framework for FOWTs that simulates the hydrodynamics physically and the aerodynamics numerically with sensor feedback from the physical testing. The framework adapts the three-loop hardware architecture to leverage greater computational resources and mitigate strict temporal requirements, enabling more computationally demanding numerical analyses in hydro-RTHS. The three-loop hardware architecture integrates multiple machines, each dedicated to either numerical analysis or RTHS controls, with a rate-transition algorithm to synchronize the tasks executed across the different machine processors. Virtual and physical tests verified and validated the hydro-RTHS framework, respectively. The ”virtual” tests, which approximates the physical domain numerically, verified the RTHS framework with respect to a numerical full-scale complete FOWT model simulated in the open-source software, OpenFAST. The virtual tests were able to maintain comparable control signals while enabling greater computational resources for the numerical calculations. Real-world physical tests demonstrated that the hydro-RTHS framework computes aerodynamic forces similar to the complete OpenFAST model, validating the hydro-RTHS framework using the three-loop hardware architecture. Findings show that the hydro-RTHS framework with the three-loop hardware architecture is computationally efficient, with reserve capacity to simulate more complex problems due to the customized software, hardware, and rate-transition algorithm.

17 WIND ENERGY↗