Search NASA⌕ Search

SEARCH · Search NASA

Results for “Computer systems performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 613 records · Page 34

Evolution at the Edge: Real-Time Evolution for Neuromorphic Engine Control

Neuromorphic computing systems are attractive for real-time control at the edge because of their low power operation, real-time processing capabilities and their potential ability to do online learning. In this work, we describe an approach for performing real-time evolution of spiking neural networks for neuromorphic systems at the edge called Neuromorphic Optimization using Dynamic Evolutionary Systems or NODES. We apply this approach to real-time combustion engine control and develop an engine-specific hardware platform for NODES called FireBox. We demonstrate how the real-time evolution approach works in simulation and the performance of networks trained in simulation on the physical engine.

Maldonado Puente, Bryan [ORNL] (ORCID:000000033880↗

Machine Learning-Based Extreme Data Reduction for Prompt Supernova Pointing at DUNE

One of the goals of the Deep Underground Neutrino Experiment (DUNE) is to use the massive underground liquid argon time projection chamber (LArTPC) detectors at its far site for multimessenger astronomy (MMA), in the detection of neutrinos from core-collapse supernovae (SNe). Its current baseline trigger strategy detects activity in the detector that is consistent with supernova (SN) neutrinos and saves the raw data for further offline analysis but provides no prompt pointing information crucial for optical follow-ups by other observatories. This approach is based on the assumption that prompt pointing determination using raw data is computationally prohibitive. In this article, we demonstrate a proof-of-concept based on applying extreme data reduction on the buffered SN data in the DUNE data acquisition (DAQ) system’s front-end computers using a machine learning (ML) workflow. This reduces the data by ~5 orders of magnitude, allowing a full track reconstruction to be carried out quickly on a single server. The total time to perform the ML-based data reduction and the full track reconstruction is less than the time to transfer the SN data back to Fermilab or a high-performance computing (HPC) center. This shows that prompt processing of raw SN data is possible and, in fact, trivial once the data have been reduced to reject radiological backgrounds, paving the way to a high-quality SN pointing trigger that is based on fully reconstructed data instead of trigger primitives (TPs).

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Development & Experimental Validation of a Generalized Resistance-Capacitance Model for Numerical Simulation of Phase-Change Material Embedded Heat Exchangers

Latent heat thermal energy storage (LHTES) using phase change material (PCM) has attracted increased attention as a viable solution for overcoming the mismatch between energy supply and demand for renewable energy-based systems. PCM-embedded heat exchangers (PCM-HX) have the potential to significantly improve thermal performance due to high storage capacity and low temperature variation during the phase change process. Most models for simulating LHTES heat transfer use Computational Fluid Dynamics (CFD) simulations, which have high computational costs resulting from considering the complex and time-dependent physics relevant to PCM-HXs. In this paper, a Generalized Resistance Capacitance-based Model (GRCM) was developed to predict the thermal performance of arbitrary PCM-HXs in a computationally efficient manner without compromising modeling accuracy. The GRCM is exercised for three case studies: (i) verification for a single-slabbed finned PCM-HX, (ii) verification and validation for a copper foam/paraffin composite PCM-HX, and (iii) validation for a straight tube annular finned PCM-HX. The copper foam PCM-HX uses an electric heater at the top of HX, while the other two configurations utilize water as heat transfer fluid. For the single-slabbed finned PCM-HX melting case, the mean deviation in average PCM temperature predicted by the GRCM compared to the CFD model was between 0.56 – 0.73 K, with maximum temperature deviation of 2.68 K. For the HTF outlet temperature, the validation results showed that GRCM prediction matches very well with experimental data, with mean temperature deviation of 0.24 K during melting case, while for solidification case was 0.34 K. These results showcase the GRCM’s capability for accurately reproducing the thermal characteristics of PCM-HXs with considerably lower computational effort.

42 ENGINEERING↗

Discovery, Design, Synthesis and Testing of High Performance Structural Alloys (Final Technical Report)

The overarching goal of this project is to understand the phase stability and mechanical behavior of non-stoichiometric multi-principal element alloy (MPEA) materials. In order to identify suitable alloys, we plan to use a combinatorial thin film screening approach, in collaboration with scientists at Lawrence Berkeley National Laboratory who are performing computational work as well as complementary experimental work. Specific tasks within the scope of this project include the fabrication, using thin film deposition from six sputtering targets, of combinatorial samples with multi-dimensional gradients in composition and microstructure. These samples are studied to screen MPEA systems for promising candidate alloys with specific composition(s), based on characterization of composition, structure and mechanical behavior across the thin film. We want to produce single-phase MPEAs with chemical homogeneity in a given thin film region, simple grain structures, and no intermetallic phases present. Gradient films facilitate first-pass screening for desirable characteristics and inform the next stage of work that involves fabrication of bulk MPEA specimens for (tensile) mechanical testing and characterization. To make the bulk alloys, metal (elemental) pieces are melted to form MPEAs, followed by heat treatment to homogenize the composition and microstructure. A subset of alloys is also cast, using vacuum arc melting, to yield larger samples (diameter ~1 cm and length ~5-10 cm) and these allow us to assess viability of scale-up for the alloys in structural applications. Further processing plans include rolling and heat treatment to recrystallize selected bulk MPEAs and grow grains to different extents, in order to investigate size effects in the mechanical behavior of MPEAs. Microspecimen testing will be performed (primarily in tension) to assess the mechanical behavior over a range of temperatures. The deformation microstructure of mechanically tested alloys will be characterized using transmission electron microscopy (TEM) to provide a scientific basis for understanding the structure-property relationships in MPEA mechanical behavior.

36 MATERIALS SCIENCE↗

Characterization and Optimization of the Fitting of Quantum Correlation Functions

This case study presents a characterization and optimization of an application code for extracting parton distribution functions from high energy electron-proton scattering data. Profiling this application code reveals that the phase-space density computation accounts for 93% of the overall execution time for a single iteration on a single core. When executing multiple iterations in parallel on a multicore system, the application spends 78% of its overall execution time idling due to load imbalance. We address these issues by first transforming the application code from Python to C++ and then tackling the application load imbalance via a hybrid scheduling strategy that combines dynamic and static scheduling. These techniques result in a 62% reduction in CPU idle time and a 2.46x speedup in overall execution time per node. In addition, the typically enabled power-management mechanisms in supercomputers (e.g., AMD Turbo Core, Intel Turbo Boost, and RAPL) can significantly impact intra-node scalability when more than 50% of the CPU cores are used. This finding underscores the importance of understanding system interactions with power management, as they can adversely impact application performance, and highlights the necessity of intra-node scaling tests to identify performance degradation that inter-node scaling tests might otherwise overlook.

Chuang, Pi-Yueh [Virginia Tech,Dept. of Computer S↗

Millimeter-Wave Superconducting Qubit

Manipulating the electromagnetic spectrum at the single-photon level is fundamental for quantum experiments. In the visible and infrared ranges, this can be accomplished with atomic quantum emitters, and with superconducting qubits such control is extended to the microwave range (below 10 GHz). Meanwhile, the region between these two energy ranges presents an unexplored opportunity for innovation. We bridge this gap by scaling up a superconducting qubit to the millimeter-wave range (near 100 GHz). Working in this energy range greatly reduces sensitivity to thermal noise compared to microwave devices, enabling operation at significantly higher temperatures, up to 1 K. This has many advantages by removing the dependence on rare 3⁢ He for refrigeration, simplifying cryogenic systems, and providing orders-of-magnitude higher cooling power, lending the flexibility needed for novel quantum sensing and hybrid experiments. Using low-loss niobium trilayer junctions, we realize a qubit at 72 GHz cooled to 0.87 K using only 4 ⁢He. We perform Rabi oscillations to establish control over the qubit state, and measure relaxation and dephasing times of 15.8 and 17.4 ns, respectively. This demonstration of a millimeter-wave quantum emitter offers exciting prospects for enhanced sensitivity thresholds in high-frequency photon detection, provides new options for quantum transduction and for scaling up and speeding up quantum computing, enables integration of quantum systems where 3 ⁢He refrigeration units are impractical, and, importantly, paves the way for quantum experiments exploring a novel energy range.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Shielding Benchmark Comparison - MCNP6.2

This report documents the calculations performed for a shielding code comparison between various Department of Energy (DOE) sites. These shielding calculations were performed using an experiment drawn from the International Handbook of Evaluated Criticality Safety Benchmark Experiments, published by the Organisation for Economic Cooperation and Development/Nuclear Energy Agency (OECD/NEA). The benchmark selected for comparison is ALARM-CF-AIR-LAB-001 (“Neutron Fields in the Three-Section Concrete Labyrinth from Cf-252 Source, Benchmark ALARM CF AIR LAB-001”). The Y-12 submission for this code comparison was performed with Monte Carlo N Particle (MCNP) Transport Code System, Version 6.2 and Automated Variance Reduction Generator (ADVANTG).

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Advanced Power Electronics and Electric Machines

NLR's advanced power electronics and electric machines (APEEM) research group develops world-class experimental and modeling capabilities to design and evaluate efficient and reliable power electronics and electric machines, as well as innovative thermal management systems to maximize their performance. They also design, fabricate, and characterize advanced power electronics packaging and are developing state-of-health monitoring techniques.

25 ENERGY STORAGE↗

Dynamical Downscaling of Earth System Model Data for Energy System Analysis

Assessing energy resources (e.g., solar, wind, and hydro) under future scenarios requires datasets with sufficient spatial and temporal detail to capture variability and extreme events. While global-scale Earth System Model (ESM) projections are widely used, their coarse resolution limits direct application to regional energy system analyses. Dynamical downscaling offers a robust approach to generate physically consistent, fine-scale datasets that better represent local atmospheric processes impacting energy resources. In this work, we present a two-stage approach for producing high-resolution historical and future projections over the contiguous United States (CONUS). First, we optimize the Weather Research and Forecasting (WRF) model configuration for energy-relevant variables - solar irradiance, wind speed, and precipitation - by conducting ERA5-driven simulations at 8-km and 28-km resolution. Multiple physics schemes and model configurations within the WRF are evaluated against observational datasets including the National Solar Radiation Database (NSRDB), the Parameter-elevation Regressions on Independent Slopes Model (PRISM), and the Stage IV multi-radar/multi-sensor precipitation product for the CONUS domain. Using the best-performing configuration, we dynamically downscale MPI-ESM1-2-HR simulations for 2000-2060 under SSP2-4.5 and SSP5-8.5 scenarios at 4-km spatial and hourly temporal resolution. This presentation will provide a comprehensive analysis of the results from multiple numerical experiments and high-resolution ESM projections. In addition, we will discuss potential applications of our high-resolution datasets within the energy sector and outline future research avenues dedicated to evaluating how extreme weather events influence system performance and resilience.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Hydrodynamic Analysis and Optimization of Aquantis Marine Turbine: Cooperative Research and Development (Final Report)

The primary aim of this proposal is to improve the accurate prediction of hydrodynamic performance and dynamic load responses of the AQ10 floating axial-flow tidal turbine with a tri-cat mooring configuration. The validation of reduced-order modeling approaches with high-fidelity model will be implemented. Additionally, the frequency response domain, Response Amplitude Floating Wind (RAFT) toolbox plus an optimizer expanded for marine hydrokinetic turbines under the Submarine Hydrokinetic And Riverine Kilo-megawatt. Systems (SHARKS) program will be used for designing and exploring different key design parameters (platform dimension, mooring layout and its parameters) of next marine hydrokinetic (MHK) turbine generation.

16 TIDAL AND WAVE POWER↗

Continuous Counter‐Current Microfluidic Liquid–Liquid Extraction Achieved Using a Pair of Wettable Screen Meshes

Continuous counter‐current microfluidic liquid–liquid extraction performs separations by flowing immiscible liquids in opposing directions within a single flow channel. In principle, this flow arrangement enables a large number of theoretical separation units in a small footprint, without using interstage valving, pumping, and phase separation. Despite its potential for excellent separation performance, this microfluidic scheme rarely appears in literature due to the requirement for capillary forces to be greater than hydrodynamic forces for stable flow. We present a novel microfluidic device and flow approaches that overcome this force‐balance challenge, enabling stable, long‐duration continuous counter‐current flow. Additionally, we cover a suite of methodologies for quantifying the performance of the microfluidic device, revealing the number of theoretical equilibrium stages achieved. The enabling technologies include a woven mesh screen‐based microfluidic device architecture that is easily fabricated outside of a clean room, surface functionalization strategies to promote conjugate (organic/aqueous) wettability, flow approaches to eliminate bubbles and carryover, and computer‐aided flow automation with optical measurement of extraction performance. The reported experiments lasted for over 36 h, terminated only at experiment conclusion, where the device still exhibited good performance. Automated Raman spectroscopy was used for solute quantitation of the ternary system tert‐butanol in a toluene/water matrix, a ternary system that was specifically chosen to analyze the device's performance with a small solute partition ratio and to enable in‐line Raman measurements of solute concentrations in both phases. The microfluidic device possessed a 55 mm contact length and a 38.5 µL internal volume. During counter‐current flow, we observed approximately 37 equilibrium stages (37 ± 13) based on a best‐fit of the solute fraction remaining in the aqueous phase using a Kremser Group Method analysis.

36 MATERIALS SCIENCE↗

CFD modeling of non-catalytic, partial-oxidation engine reformer for flare mitigation

Flaring associated natural gas is commonly employed in the oil and gas industry to reduce methane (CH 4 ) emissions but generates carbon dioxide (CO 2 ) and harmful pollutants, significantly contributing to air pollution and posing risks to public health. To mitigate this impact, M2X Energy Inc. has developed a small-scale, modular gas-to-methanol system. This system features an engine reformer that performs fuel-rich partial oxidation of wellhead gas to produce syngas—a mixture of carbon monoxide (CO) and hydrogen (H 2 )—followed by a downstream reactor for methanol synthesis. This study focused on computational fluid dynamics (CFD) modeling of the engine reformer to simulate partial oxidation chemistry, predict the rich-burn operating limit, and assess syngas quality, ultimately aiding in design and operational optimization. The CFD model, developed within a Reynolds-Averaged Navier-Stokes (RANS) turbulence framework, incorporated sub-models for turbulent combustion, a chemical mechanism with polycyclic aromatic hydrocarbon (PAH) pathways, and soot emissions to accurately capture the fuel-rich, turbulent jet ignition and combustion processes. Model validation against experimental data showed good agreement across pre- and main-chamber pressures, apparent heat release rates, and exhaust gas concentrations of key species (H 2 , CO, CO 2 , CH 4 ) for varying intake equivalence ratios. Here, the model identified a rich-burn operating limit near a fuel-air equivalence ratio of 2.35, consistent with experimental observations. Furthermore, syngas quality analysis revealed that extending the rich-burn limit through engine reformer optimization could enhance syngas production, contributing to higher methanol synthesis efficiency.

Computational Fluid Dynamics↗

Spectral gaps of two- and three-dimensional many-body quantum systems in the thermodynamic limit

We present an expression for the spectral gap, opening up new possibilities for performing and accelerating spectral calculations of quantum many-body systems. We develop and demonstrate one such possibility in the context of tensor network simulations. Our approach requires only minor modifications of the widely used simple update method and is computationally lightweight relative to other approaches. We validate it by computing spectral gaps of the 2D and 3D transverse-field Ising models and find strong agreement with previously reported perturbation theory results. Published by the American Physical Society 2024

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

CO2 Enhanced Oil Recovery Evaluation System (CO2_E_EvSystem), Version 2025

The United States Department of Energy’s (DOE’s) Office of Fossil Energy and the National Energy Technology Laboratory (NETL) developed a suite of screening-level, techno-economic models/tools, known as the CO2_E_EvSystem, to evaluate technical aspects and costs of using carbon dioxide (CO2) enhanced oil recovery (EOR) to store CO2 and produce oil. CO2_E_EvSystem has three software components and several input and output files. The software components are CO2_E_EvTool, CO2_Prophet, and CO2_E_COM. Almost all computational work is performed by CO2_Prophet and CO2_E_COM, which are both Fortran programs. The primary role of CO2_E_EvTool is to manage input and output files for the two programs, run the two programs, allow multiple oilfields to be evaluated in a single run, and generate files that summarize the results for all the oilfields run. This version of the CO2_E_EvSystem includes a residual oil zone dataset from the San Andres formation, Permian Basin, to demonstrate the system and guide users on attributes needed for dataset inputs.

AS↗

Neural operator transformers capture bifurcating drift-wave turbulence in fusion plasma simulations

Self-consistent modeling of turbulence-driven transport is critical for optimizing confinement in magnetically confined fusion plasmas, such as tokamaks and stellarators. In particular, capturing the long-term co-evolution of turbulence, flow, and background plasma profiles remains computationally challenging. Direct numerical simulation of these multiscale, highly nonlinear processes is often demanding and impractical for real-time control or design optimization. To address this bottleneck, we investigate transformer-based neural operator partial differential equation surrogates for emulating the dynamics of drift-wave turbulence bifurcation mediated by zonal flows, using the modified Hasegawa–Wakatani (MHW) model as a prototypical system. We find that the finetuned neural operator model has excellent performance in capturing the multi-spatiotemporal-scales of MHW turbulence bifurcation and is robust to testing on rare and out-of-distribution dynamics. Specifically, we demonstrate that a single unified model accurately predicts both quasi-steady-state turbulence and a wide range of dynamical transition processes, such as nonlinear saturation, spontaneous suppression of turbulence, and the emergence of macroscopic zonal flows, over time horizons vastly exceeding the local turbulence correlation time. This computationally efficient approach establishes a strong foundation for fast, AI-based modeling of complex, multiscale phenomena in magnetized fusion plasmas.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Dispatch Manager for NEML2 Constitutive Model Calculations Embedded in MOOSE

This report describes the extended capabilities of the NEML2 constitutive modeling library, including a flexible and efficient work dispatching system designed to leverage both CPU and GPU resources. This enhancement addresses one of the primary computational challenges in large-scale simulations: the ability to distribute and execute batches of material model evaluations across heterogeneous computing devices. The new dispatch system introduces a modular set of dispatcher and scheduler classes that coordinate the flow of data and execution between devices. The dispatcher is responsible for efficiently packaging work, managing device-specific memory operations, and synchronizing results. This modularity allows for extensibility, making it straightforward to integrate additional computing backends in the future. From an implementation standpoint, the dispatcher system interfaces seamlessly with NEML2's existing models. They handle device-aware tensor operations, optimize memory transfers, and support asynchronous execution when applicable. This design ensures that batches of material points can be evaluated concurrently, substantially improving throughput compared to previous single-device or serial implementations. These improvements not only enhance the raw performance of NEML2 but also improve its usability in multiscale and high-fidelity simulations, where the simultaneous evaluation of large material point batches is critical. Benchmarks included in the report demonstrate the system’s scalability, highlighting its effectiveness when leveraging modern GPU architectures.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Numerical Modeling of a Two-Stage Ocean Current Turbine

The Equinox Ocean Turbines (EQOT) current energy converter has a unique design with power generation in two small-diameter turbines attached to the tips of a large-diameter passive rotor. This configuration offers some key advantages for capturing ocean currents. With no centrally placed generator, almost no reaction torque is required at the nacelle of the main large-diameter rotor, and the small-diameter tip turbine generators operate at a higher speed and lower torque. The physics that determine the performance and loads on the turbine are also unique. The interactions of the flow field between the two stages and the general architecture of the system cannot be captured with traditional mid-fidelity modeling tools. For design iterations and large sets of load cases, it is important to have mid-fidelity models that can capture the important phenomenon with enough accuracy to identify global trends. This work uses a limited set of high-fidelity computational fluid dynamics (CFD) simulations to help inform the selection of and construction of a custom mid-fidelity model. Mid-fidelity modeling approaches were verified by comparing key turbine performance quantities to those found with the CFD model. Hydrodynamic interactions of the two-stage rotor were identified through high-fidelity CFD modeling. This highlighted the impact of the main rotor tip vortex and wake on the secondary rotor apparent inflow. This results in a relative flow rotation and sharp deficit, that change the optimal secondary rotor rotation speed and adds unsteadiness to the blade loading respectively. Multiple mid-fidelity approaches were evaluated for their ability to capture these effects. A simple approximation of the combined-stage performance based on single-stage BEM provides a reasonable rough prediction, especially near the peak TSR values, with some larger discrepancy at higher TSRs. Predicting the combined-stage performance based on single-stage CFD data improves this prediction across the TSR range. Although the combined-stage modeling in OLAF was not successful in this stage of the project, it showed promise as a mid-fidelity method, assuming the parameters can be tuned to account for the significant differences in time and length scales between the main and secondary rotors. This may be addressed through code changes in future work. A significant finding from the OLAF work was the agreement between the vortex core radius values found independently via a parameter space search and via CFD. The technique of using single-stage secondary rotor BEM, with a custom inflow taken from single-stage main rotor CFD or OLAF, provides an efficient method to capture one-way coupled flow interactions. This method provided generally good predictions of the impact of the flow rotation on the secondary rotor but struggled to accurately predict the peaks of the unsteady load progression. Future work could include some superposition of a tuned main rotor trailing edge viscous wake into the custom inflow to better predict this interaction.

16 TIDAL AND WAVE POWER↗

Intelligent Surrogate Model Development: Boosting Computational Efficiency for Autonomous Control of Advanced Reactors

Advanced reactors promise enhanced safety, greater efficiency, and waste reductions. To fully realize these benefits, it is crucial to address the need for autonomous or semi-autonomous control systems that require fewer operators. This research primarily supports the MARVEL autonomous control system, which requires real-time operation. However, the current RELAP5 reactor thermal hydraulic transient simulation is excessively time-consuming. Therefore, this study aims to leverage deep learning techniques to develop a surrogate model, providing a more efficient and accurate alternative for real-time performance. The model was trained using a combination of one-timestep prediction and scheduled sampling. It was then used for recursive prediction of the reactor state. This developed surrogate model significantly improves computational efficiency, achieving a 12 times acceleration.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN↗