Search NASA⌕ Search

SEARCH · Search NASA

Results for “coded computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Dynamically generated decoherence-free subspaces and subsystems on superconducting qubits

Abstract Decoherence-free subspaces and subsystems (DFS) preserve quantum information by encoding it into symmetry-protected states unaffected by decoherence. An inherent DFS of a given experimental system may not exist; however, through the use of dynamical decoupling (DD), one can induce symmetries that support DFSs. Here, we provide the first experimental demonstration of DD-generated decoherence-free subsystem logical qubits. Utilizing IBM Quantum superconducting processors, we investigate two and three-qubit DFS codes comprising up to six and seven noninteracting logical qubits, respectively. Through a combination of DD and error detection, we show that DFS logical qubits can achieve up to a 23% improvement in state preservation fidelity over physical qubits subject to DD alone. This constitutes a beyond-breakeven fidelity improvement for DFS-encoded qubits. Our results showcase the potential utility of DFS codes as a pathway toward enhanced computational accuracy via logical encoding on quantum processors.

Physics↗

Invariant regimes of Spencer scaling law for magnetic compression of rotating FRC plasma

Abstract The scaling laws for the magnetic compression of a toroidally rotating field reversed configuration (FRC) have been investigated in this work. The magnetohydrodynamics (MHD) simulations of the magnetic compression on rotating FRCs employing the NIMROD code (Sovinec et al 2004 J. Comput. Phys. 195 355), are compared with the Spencer’s one-dimensional (1D) theory (Spencer et al 1983 Phys. Fluids 26 1564) for a wide range of initial flow speeds and profiles. The toroidal flow can influence the scalings directly through the alteration of the compressional work as also evidenced in the 1D adiabatic model, and indirectly by reshaping the initial equilibrium. However, in comparison to the static initial FRC equilibrium cases, the pressure and the radius scalings remain invariant for the magnetic compression ratio B w 2 / B w 1 up to 6 in presence of the initial equilibrium flow, suggesting a broader applicable regime of the Spencer scaling law for FRC magnetic compression. The invariant scaling has been proven a natural consequence of the conservation of angular momentum of both fluid and magnetic field during the dynamic compression process.

Ma, Yiming↗

Prediction of performance and turbulence in ITER burning plasmas via nonlinear gyrokinetic profile prediction

Burning plasma performance, transport, and the effect of hydrogen isotope (H, D, D-T fuel mix) on confinement has been predicted for ITER baseline scenario (IBS) conditions using nonlinear gyrokinetic profile predictions. Accelerated by surrogate modeling (Rodriguez-Fernandez et al 2022 Nucl. Fusion 62 076036), high fidelity, nonlinear gyrokinetic simulations performed with the CGYRO code (Candy et al 2016 J. Comput. Phys. 324 73), were used to predict profiles of T i , T e , and n e while including the effects of alpha heating, auxiliary power (NBI + ECH), collisional energy exchange, and radiation losses inside of $r/a$ = 0.9. Predicted profiles and resulting energy confinement are found to produce fusion power and gain that are approximately consistent with mission goals ($P_\textrm{fusion} = 500$ MW at Q = 10) for the baseline scenario and exhibit energy confinement that is within 1σ of the H-mode energy confinement scaling. The power of the surrogate modeling technique is demonstrated through the prediction of alternative ITER scenarios with reduced computational cost. These scenarios include conditions with maximized fusion gain and an investigation of potential resonant magnetic perturbation (RMP) effects on performance with a minimal number of gyrokinetic profile iterations required (3–6). These predictions highlight the stiff ITG nature of the core turbulence predicted in the ITER baseline and demonstrate that $Q \gt$ 17 conditions may be accessible by reducing auxiliary input power while operating in IBS conditions. Prediction of full kinetic profiles allowed for the projection of hydrogen isotope effects around ITER baseline conditions. The gyrokinetic fuel ion species was varied from H, D, and 50/50 D-T and kinetic profiles were predicted. Results indicate that a weak or negligible isotope effect will be observed to arise from core turbulence in IBS conditions. The resulting energy confinement, turbulence, and density peaking, and the implications for ITER operations will be discussed.

gyrokinetics↗

Gyrokinetic profile prediction and validation of a negative triangularity plasma in ASDEX Upgrade

In this work, gyrokinetic simulations are performed with the CGYRO code (Candy et al 2016 J. Comput. Phys. 324 73–93) for a negative triangularity H-mode plasma in ASDEX Upgrade, and compared with experimental measurements. The PORTALS framework (Rodriguez-Fernandez et al 2024 Nucl. Fusion 64 076034) is used to accelerate the prediction of kinetic profiles for this plasma, using surrogate modeling and Bayesian optimization. Ion heat flux, electron heat flux, and electron particle flux are simultaneously matched across the simulated radial regime of the plasma (normalized radius $r/a = 0.35-0.90$), and the resulting ion temperature, electron temperature, and electron density profiles match well with the experimental profile data within this radial range. A synthetic Correlation Electron Cyclotron Emission diagnostic is applied to find well-matched electron temperature fluctuation properties between simulation and experiment. The flux-matched profiles provide a basis for investigation of the turbulence nature across the plasma radius, revealing the dominance of Trapped Electron Mode turbulence at $r/a = 0.35$, the dominance of Ion Temperature Gradient turbulence at $r/a = 0.55$, 0.75, and 0.83, and an instability boundary at $r/a = 0.90$.

gyrokinetic simulation↗

A Simple, Scalable Large Deformation Solid Mechanics Implementation in the MOOSE Framework

This article describes a large deformation solid mechanics solver implemented as part of the freely available and open source MOOSE finite element simulation framework. The article documents the choices made in developing the solid mechanics framework and describes novel formulations for the gradient operator and constitutive modeling framework made to simplify implementations of different coordinate systems, stabilized gradient operators, and different constitutive model inputs and outputs. In the process, the article describes a new formulation that casts objective integration of the Cauchy stress as a linear transformation of the small stress rate. Finally, the article presents key implementation details and examines the parallel efficiency of the solid mechanics solver implemented in MOOSE. The implementation retains a good weak scaling efficiency beyond 1,000 parallel processes. The article includes a discussion of the factors limiting the parallel efficiency of implicit, large deformation solid mechanics codes on current high-performance computers, with the main current limitation being the scalability of the algebraic multigrid methods used to solve the linearized equilibrium equations.

Applied computing → Computer-aided design↗

Agentic AI vs ML-Based Autotuning: A Comparative Study for Loop Reordering Optimization

High Performance Computing (HPC) applications rely heavily on code optimizations to achieve good performance on modern CPU and GPU architectures. Traditional Machine Learning auto-tuning approaches have demonstrated success in exploring high-dimensional spaces, but they often require expensive compile-run evaluations and lack adaptability for large HPC applications. The recent advances in Large Language Models (LLMs) and Agentic AI systems raise intriguing questions about the potential of these approaches to address specific optimization methodologies. This work aims to answer an essential question for the HPC community: “How Agentic AI Systems Compare to Traditional ML Autotuning Techniques?” To address this question, we present a comparative analysis between a traditional ML-based optimization approach and an Agentic AI system, evaluating their respective capabilities and limitations for loop-level optimization. In addition, we introduced a new Agentic AI system named LoopGen-AI using three different Large Language Models: GPT-4.1, Claude 4.0, and Gemini 2.5. A key finding is that LoopGen-AI achieves competitive per-formance with only a few program runs, the reasoning logs from the agents revealed that their decisions rely heavily on the combination of semantic understanding of the target kernel with dynamic feedback from the environment, highlighting a promising new dimension in performance tuning. In contrast, ML-based autotuners focus on statistical exploration, and require orders of magnitude more runs to reach peak performance. Additionally, our analysis shows that prompt engineering, particularly using Persona + Context Manager patterns, significantly impacts the effectiveness of Agentic AI. Our results indicate that while Agentic AI systems are not yet a complete replacement for ML-based autotuners, it can effectively complement traditional methods.

Rosas, Miguel Romero↗

Tardigrade-examples V0.2.0

Tardigrade-examples (LANL code O4735) is a repository of computational workflows that exercise the Tardigrade software package.

Allard, Thomas [Los Alamos National Laboratory]↗

ARCS: Agentic Retrieval-Augmented Code Synthesis with Iterative Refinement

Agentic Retrieval-Augmented Code Synthesis with Iterative RefinementIn supercomputing, efficient and optimized code generation is essential to leverage high-performance systems effectively. We have developed Agentic Retrieval-Augmented Code Synthesis (ARCS), an advanced framework for accurate, robust, and efficient code generation, completion, and translation. ARCS integrates Retrieval-Augmented Generation (RAG) with Chain-of-Thought (CoT) reasoning to systematically break down and iteratively refine complex programming tasks. An agent-based RAG mechanism retrieves relevant code snippets, while real-time execution feedback drives the synthesis of candidate solutions. This process is formalized as a state-action search tree optimization, balancing code correctness with editing efficiency. Evaluations on the Geeks4Geeks and HumanEval benchmarks demonstrate that ARCS significantly outperforms traditional prompting methods in translation and generation quality. By enabling scalable and precise code synthesis, ARCS offers transformative potential for automating and optimizing code development in supercomputing applications, enhancing computational resource utilization

Bhattarai, Manish [Los Alamos National Labs]↗

FY24: Progress Report on the MELCOR Modeling of the Liquid Salt Test Loop

This report outlines the activities conducted in FY24 focused on updating the liquid salt test loop (LSTL) model through the integration of a new test section featuring 16 radial tubes coupled to a filter section. Some comparative analyses of the updated model with existing experimental data for the LSTL was made and benchmarked against other computational tools, such as the ORNL code, SAM. These actions are part of a comprehensive validation effort and promote collaboration among laboratories participating in the Molten Salt Reactor (MSR) campaign.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Investigation of Frozen Chemistry for Molten Salt Reactors (MSRs)

Understanding accident progression and the potential/conditions for fission product release from fuel is necessary to evaluate safety for any nuclear reactor system. Molten Salt Reactors (MSRs) under development need such analysis to support safety evaluations. Fission product chemistry specific to MSR concepts is a critical area that introduces distinct considerations relative to the current state-of-knowledge in reactor safety, primarily developed for water-moderated nuclear reactor systems. In Light Water Reactor (LWR) systems, it is necessary to capture the chemical interaction of fission products with the reactor evironment, containment and confinement systems. The overall effects at this point are relatively well understood for the purposes of performing safety evaluations. A key insight from LWR studies is that fission product chemical behavior can be reasonably captured by modeling approaches where the chemistry is "frozen". These modeling approaches assume that radionuclide reaction and speciation can be represented by chemical classes, each with characteristic transport behavior that is invariant under a broad range of thermochemical conditions. However, radionuclides can exhibit a range of behavior in the liquid salt-melt phase of the coolant used in MSRs. Radionuclides, salt, and the metal containment surfaces (i.e. pipes) can co-exist in dynamic equilibrium that could evolve with small system mass changes. A detailed investigation to the degree the equilibrium state can dynamically evolve with changes in the conditions of the molten salt mixture has not been previously conducted. It is currently not well understood where frozen chemistry assumptions are valid. Expanding the state-of-knowledge in this regard is relevant to better assessing the range of chemical effects that should be incorporated as part of MSR safety assessments. This investigation used the Oak Ridge Isotope GENeration (ORIGEN) module of the Standardized Computer-Analysis for Licensing Evaluation (SCALE) code to generate simulated radionuclide inventories for the MSR Experiment (MSRE) and then modeled reactor chemical speciation using the Molten Salt Thermodynamic Database – Thermochemical (MSTDB-TC) coupled with Thermochimica. The effect of composition variation during decay of fission product inventory in a molten salt over a period of 500 days prolonged post- at multiple temperatures was studied. Mass fractions for fluorine and berilium were varied in order to probe the effects of free fluorine control. Finally, speciation of fluoride reactors were showed by comparing MSRE readionuclide inventories with a FLiBe based molten salt breeder reactor (MSBR). The results showed that fission product mass change has little effect on phase mass changes and vapor pressures for fluoride species, but differ with varying carrier and fuel salt compositions. However, iodine species were found to have a vapor pressure not only dependent on temperature, but also the free fluorine potential, releasing iodine when the free fluorine potential is equal to the iodine inventory. This observation, however, arose under free fluorine potentials that are very unlikely to be realized in typical molten salt mixtures. Despite this observation, temperature was found to be the dominant parameter that drove phase change and fission product species vapor pressure. The results indicate that the current frozen chemistry approach is adequate for MSR analysis.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Highlight of IC project: w25_dreamxd

(left) The Earth’s radiation belts are donut-shaped regions containing MeV electrons (color contour for density) trapped by the magnetic field (white curves). We use DREAMxD code to model their dynamics. (right) We develop a new method to accelerate a key piece of the code – diffusion coefficient calculation. Comparing the compute time in node hours required by the standard approach to the compute time required for the fast method, it shows that the new method can be 100x faster for a large problem size.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A Picture is Worth a Thousand Data Points: Introduction to Visualization

They say a picture is worth a thousand words. My response to that? A picture is also worth a thousand data points! However, not all pictures are created equal: a good visualization tells a story and helps the viewer to understand the data. A polished visualization can help you In the first half of this workshop, I will discuss the seven sins of visualization, and how to avoid them. I will introduce guidelines on how to make excellent visualization choices. In the second half of the workshop, I will guide the participants in an interactive session on creating meaningful visualizations with just a few lines of code.

code↗

Accelerating detector simulations with Celeritas: profiling and performance optimizations

Celeritas is a GPU-optimized MC particle transport code designed to meet the growing computational demands of next-generation HEP experiments. It provides efficient simulation of EM physics processes in complex geometries with magnetic fields, detector hit scoring, and seamless integration into Geant4-driven applications to offload EM physics to GPUs. Recent efforts have focused on performance optimizations and expanding profiling capabilities. This paper presents some key advancements, including the integration of the Perfetto system profiling tool for detailed performance analysis and the development of track-sorting methods to improve computational efficiency.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Reimagining Disassembly Interfaces With Visualization: Combining Instruction Tracing and Control Flow With DisViz

In applications where efficiency is critical, developers may examine their compiled binaries, seeking to understand how the compiler transformed their source code and what performance implications that transformation may have. This analysis is challenging due to the vast number of disassembled binary instructions and the many-to-many mappings between them and the source code. These problems are exacerbated as source code size increases, giving the compiler more freedom to map and disperse binary instructions across the disassembly space. Interfaces for disassembly typically display instructions as an unstructured listing or sacrifice the order of execution. Here, we design a new visual interface for disassembly code that combines execution order with control flow structure, enabling analysts to both trace through code and identify familiar aspects of the computation. Central to our approach is a novel layout of instructions grouped into basic blocks that displays a looping structure in an intuitive way. We add to this disassembly representation a unique block-based mini-map that leverages our layout and shows context across thousands of disassembly instructions. Finally, we embed our disassembly visualization in a web-based tool, DisViz, which adds dynamic linking with source code across the entire application. DizViz was developed in collaboration with program analysis experts following design study methodology and was validated through evaluation sessions with ten participants from four institutions. Participants successfully completed the evaluation tasks, hypothesized about compiler optimizations, and noted the utility of our new disassembly view. Our evaluation suggests that our new integrated view helps application developers in understanding and navigating disassembly code.

Computer science↗

SIMD Programming for the SMASH Shock Physics Code

Many modern CPUs that are available to the NNSA as mission computing resources support vector instruction sets. Making good use of vector instructions, referred to as “vectorization”, is often critical to getting the best performance from these CPUs. While other codes choose to rely on compiler auto-vectorization, the SMASH shock physics code chooses to leverage APIs for explicit vectorization. These APIs are similar to directly calling the CPU vendor’s vector intrinsics, with the additional benefit of being vendor-agnostic. This document explains what the SIMD APIs are and how to use them in developing SMASH.

97 MATHEMATICS AND COMPUTING↗

Multi-GPU porting of a phase-change cascaded lattice Boltzmann method for three-dimensional pool boiling simulations

The Lattice Boltzmann method (LBM) has proven effective in simulating phase-change phenomena, such as melting, solidification, evaporation, and boiling. In this work, we develop a highly parallelized multi-GPU implementation of LBM for three-dimensional pool boiling simulations. The code is based on the OpenACC programming model, which enables the code to be deployed efficiently on multi-core CPUs, GPUs, and potentially other accelerators, without the need for architecture-specific rewrites. To support large-scale simulations, the domain is decomposed and distributed across multiple compute nodes using MPI. We demonstrate that the code exhibits excellent scaling properties, with ideal strong-scaling running with up to 256 GPUs on the MareNostrum5 cluster.

97 MATHEMATICS AND COMPUTING↗

Gamma Source Verification for GAMSRC and GAMSOR

Nuclear reactors that rely upon the fission reaction have two modes of thermal energy deposition in the reactor system: neutron absorption and gamma absorption. The gamma rays are typically generated by neutron capture reactions or during the fission process which means the primary driver of energy production is of course the neutron interactions. The GAMSOR program was first built in the mid 1980s to properly account for the gamma heating in an operating reactor core on core internals. The GAMSOR code is sequence of DIF3D calculations to compute the neutron and gamma flux and combine them to define both the neutron and gamma heating throughout the modeled domain. The goal of this manuscript is to present the software verification of GAMSOR. The first step of the GAMSOR sequence of calculations involves of a modified version of DIF3D (called DIF3D-GAMSOR) which generates a gamma source distribution for the follow-on DIF3D gamma transport calculation (step 2). This modified version of DIF3D increases the burden of maintenance and verification work on GAMSOR as one must reverify the DIF3D capabilities which is undesirable. Because the calculation of the gamma source is the only unique aspect of GAMSOR beyond the regular DIF3D capabilities, that part was put in a standalone code called GAMSRC such that one can use the verified DIF3D code in step 1 followed by GAMSRC to carry out the same GAMSOR calculation step. As a consequence, this manuscript is focused on verification of the gamma source files generated by GAMSRC. The verification of the modified version of DIF3D (DIF3D-GAMSOR) will be done less rigorously in that it will be verified that it produces the same output that GAMSRC does and thus GAMSRC is equivalent to GAMSOR on the problems studied here. Hand calculations and independent numerical calculations of the gamma source generation are used for the verification work. This work follows the same methodology of GAMSRC. As expected, the results agree well with those calculated by GAMSRC as will be shown.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Fault Tolerant Decoding of QLDPC-GKP Codes with Circuit Level Soft Information

Concatenated bosonic-stabilizer codes have recently gained prominence as promising candidates for achieving low-overhead fault-tolerant quantum computing in the long term. In such systems, analog information obtained from the syndrome measurements of an inner bosonic code is used to inform decoding for an outer code layer consisting of a discrete-variable stabilizer code such as a surface code. The use of Quantum Low-Density Parity Check (QLDPC) codes as an outer code is of particular interest due to the significantly higher encoding rates offered by these code families, leading to a further reduction in overhead for large-scale quantum computing. Recent works have investigated the performance of QLDPC-GKP codes in detail, and the use of analog information from the inner code significantly boosts decoder performance. However, the noise models assumed in these works are typically limited to depolarizing or phenomenological noise. In this paper, we investigate the performance of QLDPC-GKP concatenated codes under circuit-level noise, based on a model introduced by Noh et al. in the context of the surface-GKP code. To demonstrate the performance boost from analog information, we investigate three scenarios: (a) decoding without soft information, (b) decoding with precomputed error probabilities but without real-time soft information, and (c) decoding with real-time soft information obtained from round-to-round decoding of the inner GKP code. Results show minimal improvement between (a) and (b), but a significant boost in (c), indicating that real-time soft information is critical for concatenated decoding under circuit-level noise. We also study the effect of measurement schedules with varying depths and show that using a schedule with minimum depth is essential for obtaining reliable soft information from the inner code.

Borah, Shantom K. [Arizona U. (main)]↗