Search NASASearch

SEARCH · Search NASA

Results for “Software evolution”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Massively parallel phase-field simulations targeting exascale

The interface thickness in the phase-field (PF) method limits its simulation scales. Consequently, large-scale PF simulations become prohibitively expensive for resolving the extremely fine microstructures that typically form during rapid solidification processing. This challenge is significant in predicting microstructure evolution in metal additive manufacturing and has been identified by the United States Department of Energy’s Exascale Computing Project. Here, to address this, we develop a multi-GPU and MPI-based massively parallel simulation code, utilizing state-of-the-art algorithms, software, and libraries, for large-scale three-dimensional (3D) PF simulations. We report the first GPU-parallel PF simulations on Frontier (currently the second TOP500 exascale cluster) and Summit machines, taking dendritic growth as an example problem. We evaluate the parallel performance of our implementation using scaling studies with more than 24 000 GPUs (among the largest known computations to date) and the acceleration performance using large-scale simulations of dendritic growth in 3D. Finally, massively parallel GPUs in these supercomputers enabled the first coupled multiscale simulations of laser melting and subsequent dendritic solidification on the scale of a full melt-pool, demonstrating the feasibility of performing PF simulations with a point total over 2 billion grid points within an acceptable time.

Exascale

Spectral proper orthogonal decomposition of active wake mixing dynamics in a stable atmospheric boundary layer

Recent advancements in the use of active wake mixing (AWM) to reduce wake effects on downstream turbines open new avenues for increasing power generation in wind farms. However, a better understanding of the fluid dynamics underlying AWM is still needed to make wake mixing a reliable strategy for wind farm flow control. In this work, a spectral proper orthogonal decomposition (SPOD) is used to analyze the dynamics of coherent flow structures that are induced in the wake through blade pitch actuation. The data are generated using the ExaWind software suite to perform large eddy simulations of an NREL 2.8 MW turbine operating in a stable atmospheric boundary layer. SPOD tracks the modal behavior of flow structures from their generation in the turbine induction field through their growth in the near-wake region and to their subsequent evolution and energy transfers in the far wake. SPOD is shown to be a useful tool in the context of AWM because it translates the wavenumber and frequency inputs to the turbine controller to structures in the wake. A decomposition of the radial shear stress flux in the wake is also developed using SPOD to measure the contribution of coherent flow structures to mean flow turbulent entrainment and wake recovery. The effectiveness of AWM is connected to its ability to excite inherent structures in the wake of the turbine that arise using baseline controls. The effects of AWM on blade loading are also analyzed by connecting the axial force along the blade to the SPOD analysis of the turbine induction field. Lastly, the performance of different AWM strategies is demonstrated in a two-turbine array.

17 WIND ENERGY

TChem-atm v1.0

SAND2024-11300O TChem-atm is a software library that was developed to solve complex kinetic models for atmospheric chemistry applications. TChem-atm interface employs a hierarchical parallelism design to exploit the massive parallelism available from modern computing platforms. It also supports gas atmospheric chemistry applications, e.g., the energy exascale earth system model. TChem can be used as a box model or coupled with a climate model to compute the time evolution of gas tracer species. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Safta, Cosmin

A Hands-On Curriculum for Training in HPC Cluster Deployment and Management

This paper presents the design, methodology, and outcomes of the High-Performance Computing Technologies (HPCT) course, a hands-on training program focused on the system-side of HPC cluster deployment and administration. Delivered as part of the Master in High Performance Computing (MHPC) program, the course introduces students to key concepts in cluster configuration, including networking, software stack provisioning, job scheduling, and monitoring. Initially taught in person, the course was transitioned to an online format during the COVID-19 pandemic. This shift led to the development of openly available instructional material and a flipped-classroom approach that continues to support both in-person and hybrid delivery. All course materials are publicly available at www.hpc.temple.edu/mhpc/hpc-technology/index.html. By documenting the structure, infrastructure, and evolution of HPCT, this paper offers a model for accessible HPC system training that supports workforce development in computational science.

Posada Correa, Fernando [ORNL] (ORCID:000000022565

Nanopolysaccharide Builder: A User-Friendly Tool for Atomistic Models of Polysaccharide-Based Nanostructures

Here, we introduce Nanopolysaccharide Builder (NPB), a user-friendly software tool designed to construct polysaccharide nanostructures─mainly those based on cellulose, chitin, and chitosan─using experimental data or user-defined parameters. NPB enables the generation of cellulose and chitin allomorphs with customizable biochemical topologies and also facilitates the construction of large bundles that replicate nanostructures found in biological support systems, including plant cell walls and arthropod cuticles. The software outputs atomic Cartesian coordinates in Protein Data Bank (PDB) format and also provides atom connectivity files in PSF and PARM formats, ensuring seamless integration with major molecular dynamics (MD) engines such as NAMD, CHARMM, GROMACS, AMBER, OpenMM, and LAMMPS. Built on an interactive visualization framework, NPB features a graphical user interface (GUI) and supports both macOS and Linux operating systems. By enabling detailed atomic-scale studies of polysaccharide evolution in extracellular matrices and cell walls of algae, bacteria, fungi, and plants, NPB is poised to advance AI-guided research in sustainable chemical development and biomass utilization.

Wan, Zhangmin [Univ. of British Columbia, Vancouve

3D Convective Urca Process in a Simmering White Dwarf

Abstract A proposed setting for thermonuclear (Type Ia) supernovae is a white dwarf that has gained mass from a companion to the point of carbon ignition in the core. In the early stages of carbon burning, called the simmering phase, energy released by the reactions in the core drive the formation and growth of a core convection zone. One aspect of this phase is the convective Urca process, a linking of weak nuclear reactions to convection, which may alter the composition and structure of the white dwarf. The convective Urca process is not well understood and requires 3D fluid simulations to properly model the turbulent convection, an inherently 3D process. Because the neutron excess of the fluid both sets and is set by the extent of the convection zone, the realistic steady state can only be determined in simulations with real 3D mixing processes. Additionally, the convection is relatively slow (Mach number less than 0.005) and thus a low Mach number method is needed to model the flow over many convective turnovers. Using the MAESTROeX low Mach number hydrodynamic software, we present the first full-star 3D simulations of the A = 23 convective Urca process, spanning hundreds of convective turnover times. Our findings on the extent of mixing across the Urca shell, the characteristic velocities of the flow, the energy-loss rates due to neutrino emission, and the structure of the convective boundary can be used to inform 1D stellar models that track the longer-timescale evolution.

Boyd, Brendan (ORCID:0000000254199751)

Improving Additive Manufactured Component Performance through Multi-Scale Microstructure Simulation and Process Optimization

The purpose of this project was to utilize computational tools to understand the relationships between processing, microstructure, and properties for additively manufactured (AM) aluminum alloys for automotive applications, and to provide an engineering solution for helping to optimize process conditions. The project leverages ORNL developments in computational modeling, including AM process modeling, phase-field based microstructure evolution predictions, and data analytics techniques for mapping process conditions to material outcomes. The project utilized an Al-Cu-Mn-Zr alloy as a model material for studying formation of defects and microstructural features in response to variations in process conditions. Based on both pre-existing experimental data and simulation results, statistical process maps were constructed to identify regions of process space with minimal defect formation and advantageous microstructures and properties. The software tools used for this purpose were successful disseminated to GM, who were able to successful compile the relevant HPC codes within their own computing ecosystem and perform initial calculations to reproduce ORNL results.

36 MATERIALS SCIENCE

Image processing workflow yielding high contrast synchrotron nanoscale computed tomography data from Ni-YSZ electrodes

The operating lifetime of Ni-YSZ fuel electrodes used in solid oxide electrolysis cells and fuel cells (SOECs and SOFCs) is limited by Ni redistribution, one of the primary degradation mechanisms that must be overcome to extend the longevity and maximize the performance of SOECs and SOFCs. To achieve this, 3D microstructural data is needed to relate both initial performance and performance loss over time to microstructural properties and their evolution throughout operation under various conditions. However, 3D microstructure data remains relatively scarce within the literature due to multiple challenges in acquiring and analyzing such data reliably. This work presents a workflow for acquiring and processing synchrotron X-ray nanoscale computed tomography (nano-CT) data from Ni-YSZ electrodes. Parameters for each step in the nano-CT workflow are described up to the final result (a 3D reconstruction), with particular emphasis on image alignment using freely available software. Following the results of a parametric sweep of the image alignment step, high contrast, low signal-to-noise 3D nano-CT data is obtained with relatively short compute times. While the exact methods best suited to samples with different microstructural qualities, or similar Ni-YSZ nano-CT data obtained from other sources may deviate from the solution found herein, this work also generalizes the decision points and evaluation of each step to provide a starting point to adapt this workflow to other datasets.

08 HYDROGEN

Resilient Entanglement Distribution in a Multihop Quantum Network

The evolution of quantum networking requires architectures capable of dynamically reconfigurable entanglement distribution to meet diverse user needs and ensure tolerance against transmission disruptions. We introduce multihop quantum networks to improve network reach and resilience by enabling quantum communications across intermediate nodes, thus broadening network connectivity and increasing scalability. We present multihop two-qubit polarization-entanglement distribution within a quantum network at the Oak Ridge National Laboratory campus. Our system uses wavelength-selective switches for adaptive bandwidth management on a software-defined quantum network that integrates a quantum data plane with classical data and control planes, creating a flexible, reconfigurable mesh. Our network distributes entanglement across six nodes within three subnetworks, each located in a separate building, optimizing quantum state fidelity and transmission rate through adaptive resource management. Additionally, we demonstrate the network's resilience by implementing a link recovery approach that monitors and reroutes quantum resources to maintain service continuity despite link failures—paving the way for scalable and reliable quantum networking infrastructures.

Alshowkan, Muneer [Oak Ridge National Laboratory (

Improving the Capabilities and Computational Efficiency of the RTE+RRTMGP Radiation Code (Final Report)

This report details progress on the RTE+RRTMGP radiation codes made during the period of performance. RTE+RRTMGP is a set of codes for computing radiative fluxes in planetary atmospheres. RRTMGP uses a k-distribution to provide an optical description (absorption and possibly Rayleigh optical depth) of the gaseous atmosphere, along with the relevant source functions, on a pre-determined spectral grid given temperatures, pressures, and gas concentration. RTE computes fluxes given spectrally-resolved optical descriptions and source functions. Spectrally-resolved fluxes are summarized (“reduced”) via a user extensible class. The initial release of the code and the design choices are described in Pincus et al. 2019; the codes are available on Github. Although RRTMGP was based on current (at the time) empirical spectroscopic data, RTE and RRTMGP were developed in large part to modernize software practices. The design focused on flexibility broadly interpreted: by separating code from data and allowing data to drive computation; in coupling to the host model (e.g. the coupling of clouds to radiative fluxes is user-controlled); with respect to programming languages (computational tasks are accessed via widely-compatible C interfaces); and with respect to hardware (the codes run on a range of CPU and GPU architectures). The code also puts an emphasis on modularity and clarity. RTE+RRTMGP v1.0 was released in September 20219. This award supported the evolution of the RTE+RRTMGP code base to support greater flexibility, accuracy, and efficiency.

54 ENVIRONMENTAL SCIENCES

Modeling of a Four-Stage Linear Ionization Cooling Channel for a Muon Collider in g4Beamline

A previous study of an eight-stage rectilinear ionization cooling channel in the ICOOL software demonstrated a five-order-of-magnitude reduction in a muon beam’s 6D emittance. In this study, we look to compare the ways ICOOL and Muons, Inc.’s g4Beamline software model ionization cooling by comparing their modeling of the first four stages of this optimized cooling channel constructed in ICOOL. We begin by identifying the parameters used to construct the optimized ionization cooling channel in ICOOL. We then reconstruct this beam in g4Beamline with identical parameter specifications and simulate the cooling of an identical input beam. Finally, we compare the two simulations based on their beam transmission, longitudinal emittance, and transverse emittance along the channel length. Through this process, we demonstrate that G4Beamline accurately reproduces transverse cooling results but predicts systematically different longitudinal emittance evolution while maintaining similar overall cooling performance, reproducing a 97.9% reduction in 6D emittance over four stages.

Keeler, Dominic [Purdue U., West Lafayette] (ORCID

Modeling of a Four-Stage Linear Ionization Cooling Channel for a Muon Collider in G4Beamline

A previous study of an eight-stage rectilinear ionization cooling channel in the ICOOL software demonstrated a five-order-of-magnitude reduction in a muon beam’s 6D emittance. In this study, we look to compare the ways ICOOL and Muons, Inc.’s g4Beamline software model ionization cooling by comparing their modeling of the first four stages to this optimized cooling channel constructed in ICOOL. We begin by identifying the parameters used to construct the optimized ionization cooling channel in ICOOL. We then reconstruct this beam in g4Beamline with identical parameter specifications and simulate the cooling of an identical input beam. Finally, we compare the two simulations based on their beam transmission, longitudinal emittance, and transverse emittance along the channel length. Through this process, we demonstrate that G4Beamline accurately reproduces transverse cooling results but predicts systematically different longitudinal emittance evolution while maintaining similar overall cooling performance, reproducing a 97.9% reduction in 6D emittance over four stages.

Keeler, Dominic [Purdue U., West Lafayette] (ORCID

Modeling Offshore Wind Farm Performance in Coastal Low-Level Jets Using Coupled Mesoscale-Microscale Large Eddy Simulations

Accurately predicting wind farm reliability under complex offshore atmospheric conditions remains a key challenge, particularly during noncanonical meteorological events such as coastal low-level jets (LLJs). LLJs, characterized by strong nonmonotonic vertical shear and directional veer, depart significantly from the simplified inflow assumptions embedded in conventional design standards, low-fidelity engineering models, and microscale large eddy simulations of the atmospheric boundary layer. In this work, we use the virtual wind farm framework—an exascale, graphics processing unit–accelerated large eddy simulation platform coupled with high-fidelity aeroservoelastic turbine models and advanced mesoscale-microscale coupling via the ExaWind software stack—to investigate turbine responses under realistic LLJ forcing. Simulations are performed over the U.S. North Atlantic offshore domain with the use of meteorological inputs from New York State Energy Research and Development Authority buoy data, focusing on a representative LLJ case impacting the International Energy Agency 15 MW reference turbine. Our results show that LLJs can cause up to 50% power deficits in downstream turbine rows and significantly amplify low-speed shaft and tower loads through nonlinear coupling between complex inflow characteristics and turbine structural dynamics. Two primary mechanisms drive these load amplifications: (1) unique LLJ inflow features—including veer and vertical/lateral shear—and (2) the downstream evolution of the flow under stable thermal stratification, which suppresses turbulence mixing and alters wake recovery. These mechanisms produce streamwise variations in turbine loading not captured by standard hub height–based metrics or existing design load case (DLC) definitions. This study highlights the critical role of rotor-scale flow gradients in driving fatigue and system-level aeroelastic responses, challenging current DLC and control strategies. We advocate the integration of full-flow field, environment-aware wind inputs into load modeling and control algorithms. By leveraging exascale computing to resolve mesoscale-microscale coupling, this work lays the groundwork for next-generation offshore wind turbine design and operation in meteorologically complex marine environments.

17 WIND ENERGY

PQML: Enabling the Predictive Reproducibility on NISQ Machines for Quantum ML Applications

Quantum computing represents a groundbreaking approach to high-performance computing. In recent years, quantum computers have progressed from single-qubit processors to systems boasting over 400 qubits. The presence of such a large number of qubits offers significant advantages, including enhanced computational speed—a capability beyond classical computing methods. However, the current stage of quantum computing is referred to as the noisy intermediate-scale quantum (NISQ) era. The existence of noise in this era presents challenges in testing quantum computing applications, leading to considerable variance in application results. Furthermore, the diverse noise characteristics observed across different machines exacerbate this issue, complicating the selection of the appropriate machine for application execution. In response to these challenges, we introduce our Predictive Quantum Machine Learning (PQML) tool. This tool is designed to predict outcomes when executing identical quantum machine learning applications—specifically, a critical suite of variational quantum algorithms—across various quantum computers during the NISQ era. This effort relies on data collected over a 12-month period. To the best of our knowledge, this study represents the first attempt to ensure reproducibility across quantum computers for complex circuits. Additionally, we have developed a model capable of forecasting the accuracy of quantum computers for variational quantum algorithms, with a particular emphasis on quantum machine learning as a case study.

Senapati, Priyabrata [Kent State University]

Three-Dimensional Heat Flux and Thermal Analysis of Angled Tungsten Samples on DIII-D

ITER-grade tungsten and dispersoid-strengthened tungsten samples with the top surface angled at ~15° towards the incident plasma flux were exposed to 9 H-mode discharges with edge-localized modes (ELMs) in the lower divertor of DIII-D tokamak using the Divertor Material Evaluation System (DiMES). Surface damage included cracking and flaking of material on the two samples farthest away from the plasma strike point, and significant melting of the two samples closest to the strike point. Heat flux and thermal analysis tools new to DIII-D have been applied to better understand this material response and to help optimize the exposure conditions for future experiments. SMITER field-line tracing simulations based on IRTV data and EFIT equilibria estimate an average inter-ELM perpendicular heat flux, 𝑞⊥,𝑖nter−𝐸LM , on the angled surfaces of 10.1 – 19.6 MW/m² for a majority of the 9 discharges, increasing to 15.6 – 24.5 MW/m² for the single, higher-power shot where samples melted. Fast camera data showed shallow intra-ELM melting and re-solidification, which transitioned to bulk inter-ELM melting with melt motion in the 𝐽⃗ 𝑥 𝐵⃗ direction. About 50% of the protruding volume of the most affected sample was displaced via melt-motion. SIERRA thermal modeling software was able to reproduce an onset time of melting consistent with fast camera data and final sample conditions, within < 200 ms. Maximum surface temperatures of 3122 K and 2787 K are estimated for the samples farthest away from the strike point, while the closest samples achieve melting at 4067 ms and 4750 ms into the ~5000 ms plasma exposure. A +10% increase in both the SMITER 𝑞⊥,𝑖nter−𝐸LM calculations and the estimated ELM heat loads 𝑞⊥, 𝐸LM was required to achieve this result, which is within the uncertainty of the diagnostic data but likely accounts for non-ideal geometry effects plus other physics uncertainties not included in this first iteration of modeling. This work provided valuable estimates of the 3D temperature evolution to help better understand the observed surface morphology and internal recrystallization of samples, which are discussed in detail in a complementary manuscript [1]. Benchmarking efforts with more diagnosed DIII-D experiments are underway to further refine the SMITER and SIERRA models for DiMES. Future use of these tools will enable researchers to precisely target heat flux exposure conditions in DIII-D to test, but not exceed, the thermomechanical limitations of novel plasma-facing materials.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Automation of Nanoparticle Synthesis Processes in a Plasma Environment Using LabVIEW

This work presents an automated control system for the synthesis of nanomaterials by plasma-enhanced chemical vapor deposition (PECVD), implemented using the LabVIEW software environment. The main objective of the study is to develop an integrated hardware-software platform that enables sequential control of the key stages of the PECVD process, including vacuum chamber preparation, pressure monitoring, working gas supply, plasma ignition, power matching, cyclic nanomaterial growth, and optical monitoring of nanoparticles in the plasma environment. The use of LabVIEW made it possible to integrate actuator control, experimental parameter acquisition, and realtime process visualization within a single automated system. The automated cycle begins with evacuation of the reaction chamber to a predefined base pressure. Transition to the next stage is permitted only after the specified pressure threshold has been reached, ensuring reproducible initial conditions for each experiment. The program then controls the supply of the working gas through mass flow controllers (MFCs). In this work, two gas-flow control modes were considered: analog control using a 0-5 V voltage signal and digital communication via RS-232 interface. It was shown that the analog approach requires accurate scaling of the control voltage, since applying 5 V corresponds to full-scale opening of the controller and results in the maximum gas flow. In contrast, the RS232 interface enables the gas flow rate to be specified directly in sccm, improving the accuracy, flexibility, and convenience of gas-environment control. After pressure stabilization, LabVIEW initiates RF plasma ignition and executes the RF matching algorithm aimed at minimizing reflected power and improving the stability of the plasma process. A separate software module implements the cyclic nanomaterial growth mode, in which the plasma-on time, plasma duration, and total number of synthesis cycles are predefined. This approach makes it possible to control material accumulation on the substrate and to correlate the process parameters with the morphological characteristics of the resulting nanostructures. The final module of the system is designed for optical monitoring of the nanoparticle cloud density in dusty plasma. For this purpose, the change in the intensity of laser radiation passing through the plasma region is recorded using a photodetector and a Keithley 2401 measuring unit connected to LabVIEW via RS-232 interface. The difference between the initial and modified optical signal intensity is used as a diagnostic parameter characterizing the formation and temporal evolution of nanoparticles. The developed system demonstrates that LabVIEW can be effectively applied not only for the automation of individual instruments, but also for the implementation of a complete digital control cycle for PECVD-based nanomaterial synthesis.

PECVD

Simulation Tools for Characterizing Stress Distribution in Laser Welded Dissimilar Joints

This project focuses on developing a thermo-metallurgical-mechanical modeling method to accurately predict the microstructural evolution and residual stress in laser welding between dissimilar metals, such as HSLA steel and high carbon equivalent (CE) gear steel. The method leverages a comprehensive material database to model the temperature and rate dependent phase transformations, along with their associated effects on material properties, such as thermal expansion and flow stress, throughout the welding process. A key innovation is the incorporation of phase transformation and phase-specific properties, which enhances the accuracy of residual stress predictions. The mixture material in the fusion zone due to the dissimilar metals will also be addressed in the numerical model. This is especially critical in scenarios involving phase transformations in the fusion zone and heat-affected zone (HAZ), where the phase changes can induce substantial residual stress variations. The material database has been generated using JMatPro. The modeling approach is implemented through a custom User Material (UMAT) subroutine, executed with the commercial finite element software Abaqus.

36 MATERIALS SCIENCE