Search NASASearch

SEARCH · Search NASA

Results for “Computer programming”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Oak Ridge Computing Academy: An HPC cluster deployment and management pilot

The High Performance Computing Technologies (HPCT) course is a hands-on High Performance Computing (HPC) cluster deployment and management training program offered as part of the International School for Advanced Studies (SISSA) and the International Center for Theoretical Physics (ICTP) Master in High Performance Computing (MHPC) specialization. Here, this training program introduces students to key concepts in cluster configuration. which include networking, software stack provisioning, job scheduling, and monitoring. The publicly available course materials feature several examples and underlying methods that are broadly applicable to cluster deployment and management. This paper discusses the design of a new workforce development program at the Oak Ridge National Laboratory that is based on HPCT, the Oak Ridge Computing Academy (ORCA). The ORCA pilot program was hosted by the Oak Ridge Leadership Computing Facility (OLCF) in Summer 2025. As a part of this discussion, HPCT and ORCA course contents and infrastructure are outlined, ORCA participant experiences are detailed, and potential opportunities for improvement are discussed.

Education

Metamaterials as a Platform for the Development of Novel Materials for Energy Applications

To explore the fundamental properties of metamaterials (MMs) / metasurfaces and their potential for control of energy at the sub‐wavelength scale in support of the mission of the Department of Energy and the office of Basic Energy Sciences. Electromagnetic metamaterials provide a platform for the discovery and design of new materials with novel structures, functions, and properties. The PI proposes to advance the knowledge base of these materials through fundamental investigations of the experimental and theoretical properties of metamaterials for the discovery, prediction and design of new materials with novel structures, functions, and properties. The proposed research activities emphasize a complete basic research program including the conceptual / computational design, fabrication / synthesis of the materials, and the characterization and analysis of their electromagnetic properties. The proposed project explores the fundamental properties of metamaterials / metasurfaces and their potential for energy applications. There are three main topics which will be investigated: 1) Dispersion engineering with metamaterials and metasurfaces, 2) Epsilon near zero metamaterial absorbers and emitters, and 3) All dielectric metamaterials. The program implements a complete basic research program consisting of theory / design, modeling, characterization, and analysis, in order to fully characterize metamaterials and metasurfaces, while at the same time minimizing iterations necessary to achieve the proposal goals.

36 MATERIALS SCIENCE

Ba 1−x Sr x FeO 3−δ as an improved oxygen storage material for chemical looping air separation: a computational and experimental study

Chemical looping air separation (CLAS) is a promising technology to generate oxygen-rich gas streams to enable efficient carbon dioxide capture during fossil fuel combustion or gasification. CLAS relies on the capture and release of oxygen from the atmosphere using the redox properties of an oxygen-selective solid oxide carrier. This study investigates the redox characteristics of Ba 1−x Sr x FeO 3−δ (0.0 ≤ x ≤ 0.417, 0.0 ≤ δ ≤ 0.5) using a combination of density functional theory (DFT) calculations and experimental verification using X-ray diffraction, thermogravimetric analysis, and oxygen-temperature-programmed desorption. The DFT computed energies of the Ba 1−x Sr x FeO 3−δ perovskites reveal a composition-dependent transition from hexagonal to cubic phases as the Sr-concentration or oxygen vacancy concentration increases. Oxygen vacancy formation energies of the cubic perovskites are found to be lower than those of their hexagonal counterparts. A low oxygen diffusion barrier of ∼1 eV combined with the thermodynamic preference of Ba 1−x Sr x FeO 3−δ compositions that form in a cubic phase suggests them as promising candidates for oxygen storage applications. The experimental results corroborate this finding by identifying Ba 0.75 Sr 0.25 FeO 3−δ in the cubic phase as an optimal composition offering low-temperature oxygen storage capacities comparable to that of the state-of-the-art Sr 0.75 Ca 0.25 FeO 3−δ perovskite oxygen storage material at 325 °C and 350 °C.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Machine-learned closure of URANS for stably stratified turbulence: connecting physical timescales & data hyperparameters of deep time-series models

Stably stratified turbulence (SST), a model that is representative of the turbulence found in the oceans and atmosphere, is strongly affected by fine balances between forces and becomes more anisotropic in time for decaying scenarios. Moreover, there is a limited understanding of the physical phenomena described by some of the terms in the Unsteady Reynolds-Averaged Navier–Stokes (URANS) equations—used to numerically simulate approximate solutions for such turbulent flows. Rather than attempting to model each term in URANS separately, it is attractive to explore the capability of machine learning (ML) to model groups of terms, i.e. to directly model the force balances. We develop deep time-series ML for closure modeling of the URANS equations applied to SST. We consider decaying SST which are homogeneous and stably stratified by a uniform density gradient, enabling dimensionality reduction. We consider two time-series ML models: long short-term memory and neural ordinary differential equation. Both models perform accurately and are numerically stable in a posteriori (online) tests. Furthermore, we explore the data requirements of the time-series ML models by extracting physically relevant timescales of the complex system. We find that the ratio of the timescales of the minimum information required by the ML models to accurately capture the dynamics of the SST corresponds to the Reynolds number of the flow. The current framework provides the backbone to explore the capability of such models to capture the dynamics of high-dimensional complex dynamical system like SST flows.

97 MATHEMATICS AND COMPUTING

Summary Report of the SOS26 Workshop held March 11-14, 2024

The SOS26 workshop was organized by Oak Ridge National Laboratory (ORNL) and held March 11-14, 2024, at Cocoa Beach, Florida. The SOS is a workshop organized annually, with a focus on distributed high performance computing (HPC). The technical program of the workshop was developed jointly by Sandia National Laboratories (SNL), ORNL and the Swiss National Supercomputing Center (CSCS). The 2024 SOS26 workshop theme was "Versatile HPC for the evolving and expanding needs of science" and had seven technical sessions covering HPC, data and machine learning (ML) topics. Each session consisted of four or five presentations, followed by a panel discussion. This report documents the workshop proceedings covering all the technical sessions.

97 MATHEMATICS AND COMPUTING

Dual-ion ECRAM as a stable and accurate analog synapse

Electrochemical random-access memory (ECRAM) works by tuning the bulk electronic conductance of functional materials via reversible, electrochemical insertion of ions, resulting in stable analog resistive switching, attractive for analog in-memory and neuromorphic computing. However, achieving fast programming for training and long retention for inference has been elusive. Protonic ECRAM demonstrates fast programming but insufficient retention, while oxygen-based ECRAM with excellent retention requires elevated programming temperatures. Cu-based ECRAM offers a compromise, with an activation energy (E A ) of ≈0.76 eV between protons (E A ≈ 0.4 eV) and oxygen (E A > 1 eV), enabling extensive retention and room temperature programming. Combining Cu 2+ ions with protons to form a dual-ion ECRAM, we demonstrate two distinct switching behaviors: fast switching at ≤5 V, (E A ≈ 0.45 eV) via protons, and nonvolatile, room temperature switching at ≥8 V, with E A ≈ 0.76 eV via Cu 2+ ions. In conclusion, the Cu-based state exhibits a wide conductance range, with excellent retention, low noise, and linear current-voltage behavior, achieving digital-equivalent ImageNet inference accuracy.

analog in-memory computing

Benchmarking Operators in Deep Neural Networks for Improving Performance Portability of SYCL

SYCL is a portable programming model for heterogeneous computing, so it is important to obtain reasonable performance portability of SYCL. Towards the goal of better understanding and improving performance portability of SYCL for machine learning workloads, we have been developing benchmarks for basic operators in deep neural networks (DNNs). These operators could be offloaded to heterogeneous computing devices such as graphics processing units (GPUs) to speed up computation. In this paper, we introduce the benchmarks, evaluate the performance of the operators on GPU-based systems, and describe the causes of the performance gap between the SYCL and Compute Unified Device Architecture (CUDA) kernels. We find that the causes are related to the utilization of the texture cache for read-only data, optimization of the memory accesses with strength reduction, use of local memory, and register usage per thread. We hope that the efforts of developing benchmarks for studying performance portability will stimulate discussion and interactions within the community.

Jin, Zheming [ORNL] (ORCID:000000027197780X)

pyRMG: A framework for high-throughput, large-cell DFT calculations on supercomputers

Exascale computing delivers the raw power to simulate ever larger and more chemically realistic systems, but realizing this potential requires codes that can efficiently use thousands of processors. Our real-space multigrid (RMG) density functional theory (DFT) code’s grid-decomposition approach scales nearly linearly with the number of graphics processing units (GPUs), even for simulations exceeding thousands of atoms. This scalability makes RMG a compelling tool for high-throughput DFT studies of materials that would otherwise be bottlenecked in other codes (for example, by global fast Fourier transforms in plane-wave DFT). However, the limited workflow infrastructure for RMG has thus far constrained its adoption to a small user community. In this work, we present pyRMG, a Python package designed to streamline the setup and execution of RMG DFT calculations. Built on the pymatgen and ASE (Atomic Simulation Environment) computational materials science Python packages, pyRMG automates input generation and convergence checking, and it integrates with modern job schedulers (e.g., Flux) on leadership-class platforms such as Frontier and Perlmutter. Here, we demonstrate pyRMG for a high-throughput study of strain effects in 2D 2L-Bi 2 Se 3 /2L-NbSe 2 heterostructures, which offers chemical insights into this system and shows that RMG-based workflows can converge with limited user intervention.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

On the uncertainties in helium generation predictions for fission and fusion alloys

With ongoing advances in fusion and advanced fission reactors, quantifying irradiation effects in materials is critical. Transmutation-induced helium in cladding and structural materials can drive swelling and embrittlement, thereby reducing these components’ lifespans. Yet most studies ignore the considerable uncertainties in predicting helium generation rates. In this work, we created a code wrapper, F-SCATTER, that automatically performs simulations in FISPACT-II. We used this tool to investigate potential variance in helium generation rate, or He/dpa, calculations based on deviations in alloy composition, irradiating neutron flux spectrum, computational methodology, and nuclear data sources. We used 12 wt% Cr HT9 steel as the reference case and observed a 6.5%–98.3% He/dpa spread based on compositional variation within a single chemical specification, a 1.8%–11.5% He/dpa variation upon the incorporation of a 15% artificial uncertainty in flux at each energy, and a He/dpa difference as high as 231% when using ENDF/B-VIII.0 versus TENDL-2021 data libraries. Similar results were found for other prominent iron-based alloys, including Grade 91, castable nano-structured alloy, and 316H—where additional variations exist based on reactor type (e.g. thermal, fast, or fusion) and alloying elements such as carbon, nitrogen, and nickel. Based on the simulated results, we conclude that a significant part of the heat-to-heat variability in swelling responses of Fe-based alloys can be driven by impurity content in alloy compositions, and, therefore, chemical control should be a key element in supply chain design for advanced nuclear energy systems. Furthermore, we provide critical recommendations on best practices for evaluating and reporting helium production and lattice damage rates when computing predictions with multiphysics programs such as FISPACT-II.

FISPACT-II

Quantum Stochastic Programming [SWR-26-040]

The Quantum Stochastic Programming tool contains quantum computing algorithms for two-stage stochastic optimization, with a focus on the Unit Commitment (UC) problem in power systems. The algorithms combine Discrete Quantum Annealing (DQA) with Quantum Amplitude Estimation (QAE) to compute expected-value objective functions over a probability distribution of wind-power scenarios. Based on: arXiv 2402.15029 - "Quantum algorithms for the two-stage stochastic unit commitment problem"

Maack, Jonathan [National Laboratory of the Rockie

Building research capacity in Southern Appalachian mountain wetlands

This project increased research capacity in Southern Appalachian wetlands by facilitating new collaborations between the PI (UNC Asheville) and DOE and academic research partners. Mountain wetlands are among the rarest natural communities in the Southern Appalachians; most of the sites contain rare or endangered species and have been incorporated into or are being considered for the new Mountain Bogs National Wildlife Refuge (established in 2015). PI Wilcox’s research since 2010 has focused on wetland hydrology at over a dozen priority field sites in western North Carolina, eastern Tennessee, and northern Georgia. This project provided time and funding to learn about DOE programs and resources, including meteorological data, watershed modeling programs, analytical capabilities, and computing facilities. PI Wilcox was able to meet potential research partners at professional conferences and the DOE-BER-ESS PI meeting and invite several of those potential partners to visit UNC Asheville and nearby field sites in Western North Carolina. These connections have resulted in three new collaborative research proposals, and a new cooperative agreement with the US Fish and Wildlife Service that includes two new field sites.

54 ENVIRONMENTAL SCIENCES

Radar clutter suppression

I have been working on something called clutter, which is an amount of “noise” caused by the surrounding environment in a radar system photo or image. The radar system is called “synthetic aperture radar” or SAR. Our group is trying to clarify the photo and hence data of an SAR image by removing clutter (and noise) from the photograph. This is done by utilizing advanced signal processing techniques and the MATLAB programming language on our computer.

Evans, Bradley [External Affiliation]

The DUNE Phase II Detectors

The international collaboration designing and constructing the Deep Underground Neutrino Experiment (DUNE) at the Long-Baseline Neutrino Facility (LBNF) has developed a two-phase strategy for the implementation of this leading-edge, large-scale science project. The 2023 report of the US Particle Physics Project Prioritization Panel (P5) reaffirmed this vision and strongly endorsed DUNE Phase I and Phase II, as did the previous European Strategy for Particle Physics. The construction of DUNE Phase I is well underway. DUNE Phase II consists of a third and fourth far detector module, an upgraded near detector complex, and an enhanced > 2 MW beam. The fourth FD module is conceived as a 'Module of Opportunity', aimed at supporting the core DUNE science program while also expanding the physics opportunities with more advanced technologies. The DUNE collaboration is submitting four main contributions to the 2026 Update of the European Strategy for Particle Physics process. This submission to the 'Detector instrumentation' stream focuses on technologies and R&D for the DUNE Phase II detectors. Additional inputs related to the DUNE science program, DUNE software and computing, and European contributions to Fermilab accelerator upgrades and facilities for the DUNE experiment, are also being submitted to other streams.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

ReEDS Performance Improvement

The Regional Energy Deployment System (ReEDS) is an open-source, spatially explicit, long-term capacity expansion model for the bulk electric power system of the contiguous United States, encompassing multiple scenarios with technological and political assumptions (see https://github.com/NREL/ReEDS-2.0). With the increased needs for capabilities, higher temporal and spatial resolutions to model the evolution of the power system with modern technologies and low-carbon pathways, ReEDS' model solution times have increased significantly from 4-6 hours in 2018 to 18-48+ hours in 2023 . Also, the model size for commonly-run ReEDS scenarios reached 22 and 28 million equations and variables, respectively. These runtimes can be especially challenging under certain scenario settings (e.g., very high temporal or spatial resolution) or with limited computational power. In this presentation, we will discuss several methods we used to improve model runtime, including data preparation, model modification, and solver tuning. The implementation of these methods shrank the model size to 7.2 and 7.3 million equations and variables, respectively. Furthermore, this led to a 77% reduction in the model's run time for commonly-run ReEDS scenarios. We will discuss the process of identifying areas for solve time improvements and how the specific enhancements for the ReEDS model might be applied to other similar large-scale models.

ENERGY PLANNING, POLICY, AND ECONOMY,MATHEMATICS A

Enabling Scientific Applications with Performance-Portability and High-Productivity for Multi-GPU Programming with JACC.Multi

This work bridges the gap between multi-GPU computing and high-productivity, performance-portable programming solutions. Our goal is to enhance scientific applications with a productive and portable solution—program once, deploy everywhere—for multi-GPU programming with no cost to programmability. To accomplish this, we implemented JACC.Multi, which is part of the Julia for ACCelerators (JACC) performance-portable framework. JACC. Multi is the only high-level, portable metaprogramming solution that targets multi-GPU environments and is integrated in a readily accessible programming language (e.g., Julia language). With transparent GPU-to-GPU communication, JACC. Multi is optimized for scientific application workloads and is portable for NVIDIA and AMD accelerators. For the evaluation, we use two modern multi-GPU systems: Hudson, which features two NVIDIA H100 Hopper GPUs per node, and Frontier, which features four AMD MI250X GPUs per node, each with two Graphics Compute Dies (GCDs) for a total of eight GCDs per node. Additionally, as part of the evaluation, we use JACC (one GPU), MPI+JACC, and JACC. Multi codes that implement well-known and widely used scientific algorithms/kernels such as the conjugate gradient algorithm and an explicit forward Euler solver that requires GPU-to-GPU communication. Overall, JACC. Multi codes achieve better performance than MPI+JACC codes and significant speedups over JACC (one GPU), with up to 1.9× on Hudson and 6× on Frontier.

Valero Lara, Pedro [ORNL] (ORCID:0000000214794310)

Vibrational Signatures of Electronic Properties in Renewable-Energy Catalysis

The objective of this research program is to discover and develop new approaches for ab initio computational simulations of molecular motion. Specifically, this program aims to decipher the connection between the underlying electronic structure and the accordant molecular vibrations in reactive ions and radicals for renewable-energy purposes. Recent experimental progress in ion sources and optical spectroscopies has unearthed considerable new inner-sphere detail for these complexes, but the connection between these spectral signatures and mechanistic information often remains elusive, to the continued frustration of experimentalists. For this purpose, new anharmonic vibrational frequency methods, along with a publicly deployed software package, will be developed. Working closely with committed experimental collaborators, this conceptual and computational framework will be used to explain the results of new spectroscopy experiments, focusing specifically on the inner-shell mechanisms of renewable-energy catalysis. The oxidation half of catalytic water-splitting chemistry will be a central focus, along with fundamental studies of the manner in which strong ions and radicals activate solvent as a chemical species. The resulting products of the research program will include openly available software and algorithms for the ab initio simulation of challenging vibrational spectra, as well as critical mechanistic insight into energy-focused catalytic processes that are opaque to other existing analytical techniques.

42 ENGINEERING