Search NASASearch

SEARCH · Search NASA

Results for “Computer codes”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

MFC 5.0: An exascale many-physics flow solver

Many problems of interest in engineering, medicine, and the fundamental sciences rely on high-fidelity flow simulation, making performant computational fluid dynamics solvers a mainstay of the open-source software community. Previous work MFC 3.0 was made a published, documented, and open-source solver via Bryngelson et al. Comp. Phys. Comm. (2021) with numerous physical features, numerical methods, and scalable infrastructure. MFC 5.0 is a significant update to MFC 3.0, featuring a broad set of well-established and novel physical models and numerical methods, as well as the introduction of GPU and APU (or superchip) acceleration. Here, we exhibit state-of-the-art performance and ideal scaling on the first two exascale supercomputers, OLCF Frontier and LLNL El Capitan. Combined with MFC’s single-accelerator performance, MFC achieves exascale computation in practice, and achieved the largest-to-date public CFD simulation at 200 trillion grid points as a 2025 ACM Gordon Bell Prize finalist. New physical features include the immersed boundary method, N-fluid phase change, Euler–Euler and Euler–Lagrange sub-grid bubble models, fluid-structure interaction, hypo- and hyper-elastic materials, chemically reacting flow, two-material surface tension, magnetohydrodynamics (MHD), and more. Numerical techniques now represent the current state-of-the-art, including general relaxation characteristic boundary conditions, WENO variants, Strang splitting for stiff sub-grid flow features, and low Mach number treatments. Weak scaling to tens of thousands of GPUs on OLCF Summit and Frontier and LLNL El Capitan achieves efficiencies within 5% of ideal to over 90% of their respective system sizes. Strong scaling results for a 16-times increase in device count show parallel efficiencies over 90% on OLCF Frontier. MFC’s software stack has undergone further improvements, including continuous integration, which ensures code resilience and correctness through over 300 regression tests; metaprogramming, which reduces code length while maintaining performance portability; and code generation for computing chemical reactions

Computational fluid dynamics

3D probabilistic fracture mechanics / computational fluid dynamics simulation of a reactor pressure vessel under transient conditions

Reactor pressure vessels (RPVs) are safety-critical light-water-reactor components that, under irradiation, experience long-term material degradation in the form of embrittlement. This can increase their susceptibility to fracture under thermal-shock conditions, which could occur during off-normal transients such as loss-of-coolant accidents (LOCAs). During a LOCA, the most severe conditions for the RPV occur when emergency core cooling water is injected through the cold legs into the water-and-steam-filled RPV. The rapid cooling of the downcomer and internal RPV surface causes decreased temperature and elevated thermally driven tensile stresses in the RPV wall. This, combined with long-term material embrittlement, may cause fracture initiation at pre-existing flaws, challenging the integrity of the RPV. Assessing RPV integrity during transients with a large spatial variation in the coolant temperature requires a modeling approach that considers the effects of spatially varying coolant temperature on the fracture probability of a population of flaws distributed throughout the RPV, accounting for spatially varying embrittlement. Here, the present study addresses this need by demonstrating first-of-its-kind coupling of a high-fidelity 3D computational fluid dynamics code with 3D probabilistic fracture mechanics This was accomplished using representative models of a pressurized-water reactor subjected to small- and medium-break LOCA conditions, both of which can result in large spatial temperature variations. While the observed impact of accounting for 3D effects was minimal under the small-break LOCA this study indicates a significant increase in the probability of fracture initiation under the medium-break LOCA when 3D effects are considered, relative to a spatially uniform cooling scenario.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Gradient Coding With Iterative Block Leverage Score Sampling

Gradient coding is a method for mitigating straggling servers in a centralized computing network that uses erasure-coding techniques to distributively carry out first-order optimization methods. Randomized numerical linear algebra uses randomization to develop improved algorithms for large-scale linear algebra computations. In this study, we propose a method for distributed optimization that combines gradient coding and randomized numerical linear algebra. The proposed method uses a randomized ℓ 2 -subspace embedding and a gradient coding technique to distribute blocks of data to the computational nodes of a centralized network, and at each iteration the central server only requires a small number of computations to obtain the steepest descent update. The novelty of our approach is that the data is replicated according to importance scores, called block leverage scores, in contrast to most gradient coding approaches that uniformly replicate the data blocks. Furthermore, we do not require a decoding step at each iteration, avoiding a bottleneck in previous gradient coding schemes. We show that our approach results in a valid ℓ 2 -subspace embedding, and that our resulting approximation converges to the optimal solution.

97 MATHEMATICS AND COMPUTING

High-Fidelity CFD Simulation of Mixed Convection and Forced Convection in a Pebble Bed Test Reactor Core

The Hermes low-power [35-MW(thermal)] reactor will be built and operated by Kairos Power LLC (KP) to demonstrate its fluoride salt-cooled high-temperature reactor (FHR) technology. In the KP FHR, the reactor core is composed of randomly packed pebbles with TRISO fuel particles inside with FLiBe flow upward through the core acting as a coolant. Previous numerical and experimental studies have been limited to either a small-size bed or to a lack of detailed measurements for heat transfer. Here, to address the lack of high-fidelity heat transfer data in a real-size FHR core, in this study, we simulated a pebble bed core with 34 374 pebbles randomly packed, similar to the Hermes reactor's size. The core radius was 14 times that of the pebble diameter, while the core height was 45 times. In this work, we were particularly interested in a mixed convection regime, where buoyancy is important. Therefore, we performed several large-eddy simulations at different Reynolds numbers (160 to 1000) with gravitational force included. The spectral element computational fluid dynamics code NekRS with graphics processing unit acceleration was used for this study. The low-Mach number approximation was applied to address property changes in the FLiBe and to account for buoyancy. A pure hexahedral mesh with 60 million elements was generated by the Voronoi cell method. At the polynomial order of 5, the total degrees of freedom was 7.5 billion. The developed case in this work is the first of its kind in terms of size and complexity. The local numerical data across the domain were obtained and compared with empirical correlations. After examining the data, we found the following conclusions. For pressure drop, the Reger correlation predicted less than a 5% error. On the other hand, for heat transfer, the Wakao correlation outperformed the others. Based on our findings, we recommend the use of the Wakao correlation for the Nusselt number calculation, and for pressure drop, the KTA (Kerntechnischer Ausschuss) correclation, among the available experimental correlations. In conclusion, the Reger direct numerical simulation-driven correlation for pressure drops should also be considered, given its best agreement with our calculations.

Mixed Convection

Developing Source Term Database for Advanced Reactors

A source term database is crucial to informing nuclear emergency response measures, enabling emergency responders to assess the potential severity of nuclear and radiological consequences. In recent times, various advanced reactor designs have come into operation, are under construction, or are being designed and developed. This report documents an effort carried out to develop a source term database for advanced reactors. The report covers key design features of these reactors and discusses radioactivity buildup and source term inventories of dose-significant radionuclides in the reactor core. For neutronic and depletion analyses, we used the SCALE code system, a computational suite for reactor physics, depletion, criticality, and sensitivity/uncertainty quantification. We used SCALE/TRITON to perform depletion calculations to predict cycle length and discharge burnup and to generate the ORIGEN reactor library. Subsequently, we used SCALE/ORIGAMI to calculate radioactivity buildup and, thereby, the source term inventories at the targeted discharge burnup, using the ENDF/B-VII.1 nuclear data library. This report covers several advanced reactors, including the KLT-40S, RITM-200N, VOYGR, and eVinci. However, other reactors, such as the RITM-200S and ARC-100, have yet to be investigated and will be explored in future efforts.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Software Quality Assurance for the MOOSE-Based Open-Source Multiphysics Code Cardinal - An Expanded CI Testing Suite

Cardinal is a wrapping of the GPU-oriented spectral element Computational Fluid Dynamics (CFD) code NekRS and the Monte Carlo particle transport code OpenMC within the Multiphysics Object-Oriented Simulation Environment (MOOSE). Cardinal provides high-resolution thermal-hydraulics and/or radiation transport feedback to MOOSE multiphysics simulations. Multiphysics feedback is implemented in a geometry-agnostic manner which eliminates the need for rigid one-to-one mappings. A generic data transfer implementation also allows NekRS and OpenMC to couple to any MOOSE application, enabling a broad set of multiphysics capabilities. Cardinal simulations can also leverage combinations of MPI, OpenMP, and GPU resources. Cardinal continuous development and improvement efforts have led to the software being considered as a high-fidelity design and licensing tool for key areas of nuclear reactor relevant physics, including neutron transport, fluid flow, heat transfer, and mechanical processes. The fast development and expansion of the software from a pure R&D framework towards its application in the nuclear industry and regulation require a focus on developing, enhancing and, maintaining Cardinal’s software quality through strict adherence to a Software Quality Assurance (SQA) framework and SQA program. To facilitate compliance with SQA standards, the Cardinal SQA Program has been initiated during Fiscal Year 2023 (FY23). During the development of the Cardinal SQA Program, multiple gaps have been identified. These gaps are primarily related to model verification and code pedigree as they relate to the use of Cardinal as a safety analysis tool. These gaps have been captured in a report published in 2023. A second report highlighted the progress made during Fiscal Year 2024 (FY24) and described Argonne’s effort to document and integrate software verification within Cardinal’s software development process. This report documents a snapshot of the verification test cases currently available for Cardinal and NekRS in their assimilation into a Continuous Integration (CI) platform. Following the CI practice permits the integrating of source code changes frequently and ensuring that the integrated codebase clears the verification testing for the software. It should be noted that the SQA program itself, including the program plans, procedures, configuration management, and testing strategies, need to be developed in a future step of this task.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Practical Implementation of GPU-based Computing at the Grid Edge for Resilience Scenarios

This paper presents a practical implementation of GPU-accelerated computing at the grid edge to enhance power system resilience through next-generation smart meters. Advanced Metering Infrastructure (AMI) systems rely predominantly on centralized processing architectures, which limit real-time response capabilities during grid disturbances. This work proposes the integration of GPU-enabled computational platforms directly within smart meter to enable local execution support for power system analytics, fault detection algorithms, and optimization routines. The proposed framework uses the Julia programming language to leverage highperformance parallel computing capabilities while maintaining code portability and development efficiency. We use two experimental scenarios to benchmark the computational feasibility of this approach: sparse linear system solutions representative of power flow analyses, and multi-stage production cost simulations incorporating unit commitment and economic dispatch operations. Results demonstrate that computationally intensive power system algorithms, such as those supporting resilience scenario calculations, can be effectively executed at the distribution edge using commercially available embedded GPU hardware. Keywords—GPU acceleration, edge computing, smart meters, grid resilience, AMI, resilience.

De Souza, Reubun [School of Electrical Engineering

A Performance-Portable MultiGPU Implementation of 3D Euler Equations using ProtoX and IRIS

Computational scientists often face challenges when developing and optimizing code for high-performance computing (HPC), especially when trying to leverage GPUs. Given the heterogeneity of the nodes that comprise many modern HPC facilities, considerable demand exists for performance portable solutions for the core computational kernels used in many scientific computing libraries. In this work, we demonstrate a fourth-order finite volume method–based implementation of the Euler equations, which are an integral part of computational fluid dynamics. Our performance-portable multiGPU implementation for Euler equations uses ProtoX to generate kernels and IRIS for portability. ProtoX is a domain-specific language that uses a structured-grid partial differential equation library called Proto as its front end and the SPIRAL code generation system as its back end to generate optimized kernels for different architectures. Optimized kernels generated by ProtoX are orchestrated through the IRIS intelligent runtime system to provide portability. Two levels of optimizations within the IRIS runtime— directed acyclic graph fusion and task fusion—are explored to efficiently utilize computing resources in a multiGPU environment. Performance improvement through these optimizations is showcased by comparing the base ProtoX-IRIS implementation on AMD GPUs (Frontier node) and on NVIDIA GPUs (NVIDIA DGX-1).

Mankad, Het

SCALE Non-LWR Models for NRC Volume 3

This dataset contains input and result files of computational simulations with the SCALE code system. The simulations cover radionuclide inventory and reactivity analyses of various advanced reactors. Users wanting to reproduce results from this dataset are required to obtain a license to the SCALE code system for which details on the distribution can be found here: https://www.ornl.gov/scale/releases

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS

SCALE Non-LWR Models for NRC Volume 5

This dataset contains input and result files of computational simulations with the SCALE code system. The simulations cover radionuclide inventory generation, criticality calculations, and dose rate/shielding analyses of various advanced reactors. Users wanting to reproduce results from this dataset are required to obtain a license to the SCALE code system for which details on the distribution can be found here: https://www.ornl.gov/scale/releases

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS

High-Temperature Gas-Cooled Reactors Multiphysics Simulation Demonstration and Code Validation

This study presents a comprehensive benchmarking and verification effort of several thermal-hydraulic and multiphysics capabilities for high-temperature gas-cooled reactor applications. The first part of this effort focuses on the running-in verification of Griffin’s multiphysics capabilities, specifically for simulating the evolution of pebble-bed reactor cores from startup to equilibrium. Since Fiscal Year 2024, improvements and enhancements have been implemented in Griffin, including simplifying the process to specify streamlines and developing the online cross-section generation capability. In the absence of validation data, code-to-code comparisons are conducted with kugelpy, showing good agreement for integral quantities like k-eff predictions and predictions for maximum power density. However, accuracy issues are noted for more detailed quantities like the spatial distribution of fission rate densities which will require further work to address. The second part of this report presents an improved System Analysis Module (SAM) core channel model where the effects of cross flow are considered during the pressurized loss of forced cooling transient, resulting in an improved agreement of the predicted pebble temperature with respect to the predictions from the SAM 2D porous media model. Additionally, the wall channeling effect due to variable porosity at the near wall region of the core is also investigated. Furthermore, to demonstrate Griffin’s online cross-section generation capability, a Multiphysics simulation is performed by coupling Griffin to the SAM core channel model. In the third part of the report, as a part of the Organisation for Economic Co-operation and Development/Nuclear Energy Agency (OECD/NEA) thermal-hydraulic code validation benchmark activity for a high-temperature gas-cooled reactor, the High Temperature Test Facility (HTTF) is investigated first using the NekRS computational fluid dynamics (CFD) code to study the flow mixing phenomenon in the lower plenum of the facility. Then, code-to-code and code-to-data comparisons are performed for Test PG27, which is a pressurized conduction cooldown (PCC) test, using five different codes by six organizations from five countries. The different simulations show good agreements in terms of the general trend but there are differences in some results such as the peak temperatures of different regions and heat removal rate.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Model Validation and Uncertainty Quantification on the KRUSTY Microreactor Design Using GRIFFIN Neutron Transport Code [Poster]

Argonne National Laboratory (ANL) and INL have developed a GRIFFIN steady state neutronics model for the multiphysics simulations of the Kilopower Reactor Using Sterling TechnologY (KRUSTY) microreactor in the Multiphysics Object Oriented Simulation Environment (MOOSE). The reliability of such deterministic neutronics models can be validated by comparing with computations from Monte Carlo codes (e.g. MCNP, SERPENT, OpenMC, Shift, etc). Furthermore, potential modeling/design improvements can be identified by incorporating uncertainty quantification (UQ), which can be performed by MOOSE’s Stochastic Tools Module (STM). KRUSTY is a prototype for a 5-kW thermal nuclear-powered space reactor. Its primary components consist of nuclear fuel, heat pipes, a control rod, a reflector, and the shielding. The fuel consists of 3 stacked U-7.65Mo cylinders with a hole in the center for the control rod. 8 liquid sodium heat pipes transfer fission energy from the solid fuel block to the Sterling power conversion system where the energy is extracted, and the cooled sodium flows back to the core via capillary action . The movable Boron Carbide control rod regulates the neutron population during startup or when a reactor temperature boost is needed . The beryllium oxide reflector is in 3 places in the reactor; it surrounds the core axially, it lies beneath the core on a platen, and it is present in the shim. The axial and lower reflectors rest on an adjustable stainless-steel platen that moves upward to cover the fuel and help the reactor reach criticality. Lastly, radial stainless steel surrounds the core offering protection from radiation exposure .

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS

Plasma rotation and diamagnetic drift effects on the resistive wall modes in the negative triangularity tokamaks

Abstract It was found previously that the negative triangularity (NT) configuration is more MHD-unstable for low n modes than the positive triangularity (PT) case, although the situation is reversed for intermediate n modes and the NT configuration becomes more stable for intermediate n modes ( n = 3 − 10 ) (Zheng et al 2021 Nucl. Fusion 61 116014). Here, n is the toroidal mode number. In this work, we extend the studies to include the rotation effects, as well as the diamagnetic drift effects, to see how the resistive wall modes (RWMs) in the NT configuration are affected as compared with the PT configuration. This is particularly motivated by noting that the wall interface with the plasma is quite different between the NT and PT configurations. It affects the plasma rotation and diamagnetic drift effects on the low n RWM. We consider the DIII-D-NT-experiment equilibrium reconstructed by the EFIT code. Based on the equilibrium g-file, the extended equilibria are constructed with the VMEC code by varying the beta values while keeping the pressure and poloidal current flux profiles basically unchanged. The bootstrap current contribution to the equilibria is taken into account with the Sauter formula. The MHD stability is then computed using the AEGIS code with the rotation and diamagnetic drift effects taken into account. We found that, although the NT configuration is less stable for n = 1 MHD modes, the rotation and diamagnetic drift stabilization effects on RWMs are more effective in the NT configuration than in the PT one. Note that even in the PT case, the stabilization of RWMs by the rotation and kinetic effects is critical. Because the low-n RWMs in the regular NT case are more unstable, the rotation and diamagnetic drift stabilization effects found in this research are important for the NT tokamak concept.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Geometric Structure and Transversal Logic of Quantum Reed–Muller Codes

Designing efficient and noise-tolerant quantum computation protocols generally begins with an understanding of quantum error-correcting codes and their native logical operations. The simplest class of native operations are transversal gates, which are naturally fault-tolerant. Here, in this paper, we aim to characterize the transversal gates of quantum Reed–Muller (RM) codes by exploiting the well-studied properties of their classical counterparts. We start our work by establishing a new geometric characterization of quantum RM codes via the Boolean hypercube and its associated subcube complex. More specifically, a set of stabilizer generators for a quantum RM code can be described via transversal X and Z operators acting on subcubes of particular dimensions. This characterization leads us to define subcube operators composed of single-qubit π/2 k Z -rotations that act on subcubes of given dimensions. We first characterize the action of subcube operators on the code space: depending on the dimension of the subcube, these operators either (1) act as a logical identity on the code space, (2) implement non-trivial logic, or (3) rotate a state away from the code space. Second, and more remarkably, we uncover that the logic implemented by these operators corresponds to circuits of multi-controlled-Z gates that have an explicit and simple combinatorial description. Overall, this suite of results yields a comprehensive understanding of a class of natural transversal operators for quantum RM codes.

Reed–Muller (RM) codes

Large language model evaluation for high–performance computing software development

We apply AI-assisted large language model (LLM) capabilities of GPT-3 targeting high-performance computing (HPC) kernels for (i) code generation, and (ii) auto-parallelization of serial code in C ++, Fortran, Python and Julia. Our scope includes the following fundamental numerical kernels: AXPY, GEMV, GEMM, SpMV, Jacobi Stencil, and CG, and language/programming models: (1) C++ (e.g., OpenMP [including offload], OpenACC, Kokkos, SyCL, CUDA, and HIP), (2) Fortran (e.g., OpenMP [including offload] and OpenACC), (3) Python (e.g., numpy, Numba, cuPy, and pyCUDA), and (4) Julia (e.g., Threads, CUDA.jl, AMDGPU.jl, and KernelAbstractions.jl). Kernel implementations are generated using GitHub Copilot capabilities powered by the GPT-based OpenAI Codex available in Visual Studio Code given simple + + prompt variants. To quantify and compare the generated results, we propose a proficiency metric around the initial 10 suggestions given for each prompt. For auto-parallelization, we use ChatGPT interactively giving simple prompts as in a dialogue with another human including simple “prompt engineering” follow ups. Results suggest that correct outputs for C++ correlate with the adoption and maturity of programming models. For example, OpenMP and CUDA score really high, whereas HIP is still lacking. We found that prompts from either a targeted language such as Fortran or the more general-purpose Python can benefit from adding language keywords, while Julia prompts perform acceptably well for its Threads and CUDA.jl programming models. Finally, we expect to provide an initial quantifiable point of reference for code generation in each programming model using a state-of-the-art LLM. Overall, understanding the convergence of LLMs, AI, and HPC is crucial due to its rapidly evolving nature and how it is redefining human-computer interactions.

97 MATHEMATICS AND COMPUTING

Integration of scanning probe microscope with high-performance computing: Fixed-policy and reward-driven workflows implementation

The rapid development of computation power and machine learning algorithms has paved the way for automating scientific discovery with a scanning probe microscope (SPM). The key elements toward operationalization of the automated SPM are the interface to enable SPM control from Python codes, availability of high computing power, and development of workflows for scientific discovery. Here, we build a Python interface library that enables controlling an SPM from either a local computer or a remote high-performance computer, which satisfies the high computation power need of machine learning algorithms in autonomous workflows. We further introduce a general platform to abstract the operations of SPM in scientific discovery into fixed-policy or reward-driven workflows. Furthermore, our work provides a full infrastructure to build automated SPM workflows for both routine operations and autonomous scientific discovery with machine learning.

47 OTHER INSTRUMENTATION

Efficient Routing of Quantum LDPC Codes on Programmable 2D Toric Architectures

Quantum low-density parity-check codes are promising candidates towards scalable fault-tolerant quantum computation. Among these, bivariate bicycle (BB) codes offer superior encoding rates and large code distance compared to surface codes. However, their requirement on long-range stabilizer measurements poses significant challenges for implementation on realistic hardware with limited connectivity, such as superconducting circuit platforms. In this work, we introduce a novel hardware-software co-design that leverages a programmable communication network architecture to address these limitations. Our approach utilizes a 2D toric network of oscillators as a flexible communication fabric linking qubits at each site. Such architecture significantly reduces the number of long-range couplers required from O ( n ) to O (√ n ). Dual-rail qubits, along with native gates including Swap-Wait-Swap gates and beamsplitter SWAPs, ensure that long-range two-qubit gates can be executed with high fidelity and low latency. To further enhance performance, our qubit layout and routing algorithm utilize symmetries of the codes and enable maximum parallelism for long-range two-qubit gates, maintaining a low syndrome extraction cycle duration and scalability over the code length. We perform circuit-level simulation with realistic noise modeling based on experimental hardware parameters, observing an logical error rate per logical qubit per cycle of 3.06% for [[18,4,4]] BB code, 2.6× less than the existing experimental result. These findings provide a practical roadmap and identify key technological advancements needed to achieve low-overhead fault-tolerant quantum computing at scale.

Liu, Kun [Yale Univ., New Haven, CT (United States

On a Simplified Approach to Achieve Parallel Performance and Portability Across CPU and GPU Architectures

This paper presents software advances to easily exploit computer architectures consisting of a multi-core CPU and CPU+GPU to accelerate diverse types of high-performance computing (HPC) applications using a single code implementation. The paper describes and demonstrates the performance of the open-source C++ matrix and array (MATAR) library that uniquely offers: (1) a straightforward syntax for programming productivity, (2) usable data structures for data-oriented programming (DOP) for performance, and (3) a simple interface to the open-source C++ Kokkos library for portability and memory management across CPUs and GPUs. The portability across architectures with a single code implementation is achieved by automatically switching between diverse fine-grained parallelism backends (e.g., CUDA, HIP, OpenMP, pthreads, etc.) at compile time. The MATAR library solves many longstanding challenges associated with easily writing software that can run in parallel on any computer architecture. This work benefits projects seeking to write new C++ codes while also addressing the challenges of quickly making existing Fortran codes performant and portable over modern computer architectures with minimal syntactical changes from Fortran to C++. We demonstrate the feasibility of readily writing new C++ codes and modernizing existing codes with MATAR to be performant, parallel, and portable across diverse computer architectures.

97 MATHEMATICS AND COMPUTING